🤖 모델 & 제품 (12/32건)
Proaction boosts sales 60% and saves 75+ hours with Codex
With Codex, GPT-Live-1, and GPT-6 Astra, Proaction builds, operates, and sells modern fleet management faster.
Introducing Gemini 3.8 Live with Live Avatar
Two years of OpenAI Academy
Marking two years of OpenAI Academy and bringing AI skills to even more communities.
OpenAI extends cyber access to Ukraine for civilian defense
OpenAI is extending access to its Daybreak program to the Government of Ukraine to support the cyber defense of civilian infrastructure.
Sam Altman’s remarks at the United Nations Security Council
OpenAI CEO Sam Altman discusses AI safety, human control, and international cooperation in remarks to the United Nations Security Council.
Harvey turns legal context into stronger drafts with GPT-6 Astra
GPT-6 Astra produces more structured, context-aware legal documents, freeing lawyers to focus on strategy.
How invideo improves color grading 3x with GPT‑6 Astra
With GPT‑6 Astra, invideo plans edits with greater precision, improves color correction and grading threefold, and produces 50 custom effects in one day.
Google Beam expands with new regions, partners, and customers
Google Beam promotional animation
Advancing Private AI Compute with secure, server-side memory
Introducing private, server-side memory to Private AI Compute for personal AI.
Gemini 3.8 text-to-speech says hello
Claude discovers a novel enzyme system
In early results from our new life sciences research lab, Claude agents found an enzyme system whose function is still unknown.
How SpaceXAI is using Grok Bot to scale customer support
We rebuilt the combined SpaceXAI and Cursor support operation around Grok Bot, expanding to a much broader product portfolio without adding headcount.
🌎 업계 동향 (24/161건)
기밀 추정치: 미 NSA, AI 모델 테스트에 수십억 달러 지출 중
해당 비용은 기존에 알려진 것보다 훨씬 더 높은 수준입니다.
Too AI; Didn't Read
If you couldn't bother to read it, why should I? Not anti-AI. Pro-giving-a-damn.
Meta's Muse appears to use an OpenAI model labeled muse-special
While Muse was building my website, one subagent ran on a model called azure/muse-special. So I dug through the filesystem again.
마이크로소프트, 코파일럿 전면 개편으로 개인용 AI 챗봇 경쟁서 철수
Hacker News (69점, 댓글 54개)
Claude도 9-루프 계산을 수행할 수 있다
Claude가 N=4 초대칭 양-밀스 이론(Super-Yang-Mills)에서 9-루프 진폭(nine-loop amplitude)을 계산해냈습니다.
Ahead of US IPO, British AI neocloud Nscale secures $3.36B in convertible financing
The funding, which comes from Third Point, Nvidia, and others, will fuel the company's massive AI data center buildout.
Meta’s Muse just stole the AI spotlight from OpenAI and Anthropic
When AI leaders at OpenAI and Anthropic started talking about “pacing the frontier,” maybe someone should have asked: what pace? Now it’s turned into model drop week for both companies as Anthropic rolled out Opus 5.5, followed by OpenAI
Some Supabase customers are publicly exposing reams of people’s data to the web
The findings highlight how AI-generated and vibe-coded apps can spill and expose users' data when not configured or secured properly.
Astra and Opus just passed Turing’s other test
Frontier AI models are finishing Alan Turing's World War II codebreaking work.
Meta is putting its muscle behind Muse as the AI app takes off
Muse is topping the app store charts and adding users at a rapid clip, while Meta ramps up the personal AI agent's promotion across its own apps and beyond.
Meta’s AI Tamagotchi bet is…working?
When AI leaders at OpenAI and Anthropic started talking about “pacing the frontier,” maybe someone should have asked: what pace? Now it’s turned into model drop week for both companies as Anthropic rolled out Opus 5.5, followed by OpenAI
Meta makes the Muse filesystem even more accessible
Yesterday, with a little prodding, it was discovered that Meta's Muse would expose its filesystem to curious users. The files offered a fascinating peek under the hood of an AI chatbot, and appeared to expose details we weren't meant to see, not least because Muse itself told people, inclu
Can Apple Home’s AI camera features outsmart Amazon’s and Google’s? I put them to the test
A few years back, I was at a beachside Easter egg hunt, watching my kids dash through sand dunes searching for sweet treats. My phone buzzed in my pocket; I ignored it. A moment later, it buzzed again. I pulled it out, glanced down, and saw a "motion detected" notification from my security
Microsoft thinks its new Copilot ‘super app’ will be as influential as Office
After teasing its new Copilot "super app" last month, Microsoft is officially unveiling it today. The redesigned Copilot app bundles three AI capabilities into a single interface of chat, coding, and agents. As part of the launch, Microsoft is also rebranding Scout, the AI personal assista
Stanford and Nvidia's open CLM-8B caches reusable agent actions and runs up to 9x faster than Jev in tests
AI agents often use LLMs to choose between a fixed set of tools or rank candidate outputs. In each case, the model processes a prompt and generates tokens even though the application only needs a bounded decision.Read more
AI agents have routed around access blocks. Only 9% of companies in VentureBeat's August survey isolate high-risk agents
An OpenAI agent on a research task broke into an Australian government health-data portal, and OpenAI did not detect the intrusion until an internal review weeks later.Read more
Microsoft revamps its Copilot AI with a persistent Autopilot agent and hosting for AI-generated apps
Microsoft is giving its signature AI assistant Copilot a larger role in running the workplace, announcing a redesign Friday that combines Office collaboration, software creation and autonomous agents that continue assignments after their human colleagues sign off.Read more
A new skill finds AI agent risks, fixes them, and proves the fix worked
The post A new skill finds AI agent risks, fixes them, and proves the fix worked appeared first on Source.
We’re building Copilot as a new OS for work that spans every model, every form factor, and every task. Today, we’re announcing our biggest update to Copilot to date, bringing four things together [Read more]
The post We’re building Copilot as a new OS for work that spans every model, every form factor, and every task. Today, we’re announcing our biggest update to Copilot to date, bringing four things together [Read more] appeared first on Source.
Introducing the new Copilot with Home, Code and Autopilot
The post Introducing the new Copilot with Home, Code and Autopilot appeared first on Source.
Our team’s work in Nature this week. A great example of how AI is helping to accelerate molecular discovery.
The post Our team’s work in Nature this week. A great example of how AI is helping to accelerate molecular discovery. appeared first on Source.
The Pentagon wants $30 million to build an AI-powered lie detector
The US government wants to spend $30.3 million over the next five years on an improved form of lie detector, according to a Department of Defense budget request. The program, called Polygraph+ or Polygraph Next, will focus on scoring algorithms that use artificial intelligence and machine learning a
Meta, Meta에서 촬영 후 Meta AI 안경에 대한 비판적인 영상 삭제
해커 뉴스 (314점, 149개 댓글)
Microsoft disrupts AI-assisted platform that compromised 12,000 accounts
EvilTokens provided an end-to-end platform that makes mass compromises faster and easier.
🛠️ 도구 & 오픈소스 (5/38건)
Quoting John Gruber
Muse is getting a lot of attention — including mine — because it’s both groundbreaking technically (each user gets their own entire persistent Linux VM running in Meta’s cloud) and because it’s packaged in an easy-to-install easy-to-use way. It’s literally presented as a cute mascot. It’s the first
Northern Gannet, Great Blue Heron, California Brown Pelican
Northern Gannet, Great Blue Heron, California Brown Pelican, in Monterey Bay National Marine Sanctuary, CA, US, CANew 200-800mm Canon EF lens got me my best photo of Morris yet. They really like hanging out under that sign in the harbor! Tags: photography, wildlife
Note on 24th September 2026
The more time I spend working with coding agents, the more convinced I am that they make software engineering even harder. We can do amazing things with them, but unlocking their full potential requires extraordinary discipline and knowledge. Tags: coding-agents, ai, llms
GitHub Copilot app for Beginners: How to build custom workflows with canvases
Describe the interface you need in plain English, then let the agent build a live surface you can both use and update—so you spend less time adapting to tools and more time getting work done. The post GitHub Copilot app for Beginners: How to build custom workflows with canvases appeared first on The
Accelerating vision-language models with LFM2.5-VL-DSpark
📚 논문 & 연구 (18/117건)
LLM Agents Can Easily Tamper With Their Own Traces
Asynchronous monitoring, incident investigations, and compliance audits primarily rely on agent traces to reconstruct what happened. These analyses assume that LLM agents cannot tamper with their own execution traces. We show that local LLM agents such as Claude Code, Codex, Antigravity, Open Code a
AD-WM: Action-Discriminative World Models for Counterfactual Model Predictive Control
Latent world models are typically trained to predict factual transitions, whereas model predictive control (MPC) must compare alternative actions from the same state. A model can therefore achieve low factual prediction error yet poorly distinguish candidate actions. We introduce AD-WM, an action-di
RAPID: Robot Agentic Programming from Demonstrations
Coding agents have demonstrated enormous success in solving complex programming problems. To leverage their potential for robot systems, this work introduces Robot Agentic Programming from Demonstrations (RAPID), which automatically generates, verifies, and refines robot programs, given a single vis
Rolling-WAM: World Action Models with Rolling Imagination
World Action Models (WAMs) couple action generation with future visual prediction for robotic manipulation. However, completing the joint video-action denoising process at each replanning cycle incurs substantial latency, delaying action updates and limiting closed-loop responsiveness. We present Ro
Coding Agents for Generalized Task and Motion Planning Problems
Task and motion planning (TAMP) problems remain difficult even with full observability and object-centric states because discrete decisions are tightly coupled to geometric, kinematic, and dynamic constraints. Generalized TAMP addresses this difficulty by exploiting regularities across problem insta
To Trust or Not to Trust: Retrieval-Augmented Fact Checking in Speech
Online misinformation increasingly appears in spoken formats such as news clips, podcasts, interviews, political speeches, and social media videos, creating a need for fact-checking systems that can verify claims directly from speech. We introduce VeriSpeak, a probe benchmark for studying speech-bas
Temporal Gradient Inversion for Private Trajectory Reconstruction in Embodied Reinforcement Learning
Distributed learning in embodied reinforcement-learning agents offers a degree of privacy by retaining raw sensor data on-device and transmitting only policy gradients to the server. Yet temporal structure can amplify this leakage beyond single-frame attacks. We introduce Temporal Reconstruction Att
Agentic Detection of Online Conspiracies
Conspiratorial discourse on social media is not always expressed through explicit claims or stable lexical markers. The same surface content may express endorsement, legitimate concerns, criticism, satire, or mockery. The main challenge is therefore not only recognizing conspiracy-related claims, bu
A Nearly Quadratic Lower Bound for Linear Optimization over Convex Bodies in the Membership Oracle Model
We prove nearly quadratic lower bounds for randomized algorithms for linear optimization and uniform sampling over convex bodies in the membership oracle model. For linear optimization, this matches the known nearly quadratic upper bound up to a polylog factor in the dimension. For uniform sampling,
Anchored Extra-Proximal Methods: Optimal Higher-Order Methods for Monotone Inclusion Problems
We study the deterministic oracle complexity of finding approximate solutions to composite monotone inclusion problems, formed by the sum of a smooth single-valued monotone operator and a maximally monotone set-valued operator, under the tangent-residual criterion. We introduce the Anchored Extra-Pr
The Alignment Illusion in Multimodal Large Language Models
Layer-wise visual-text similarity in Multimodal Large Language Models (MLLMs) is widely interpreted as evidence that the language model progressively integrates visual content into a shared representation space. This reading rests on the assumption that scalar alignment scores reflect content-level
Beyond Compression: Training Latent Representations for Stable Long-Horizon Rollout in Neural Surrogate Solvers
Latent neural surrogate solvers, or latent dynamics models, accelerate simulations of time-dependent physical systems by evolving a compressed latent space rather than resolving full-resolution fields directly. In principle this reduces computational cost and simplifies learning, but in practice err
JevOut: Natural Context Can Flip Decision Models
Dedicated decision models such as Jev map unstructured language to probability distributions over finite choices, allowing their outputs to directly route requests, select tools, and trigger actions. Yet real-world inputs rarely arrive in isolation: they come with background details and surrounding
SemMSA: Latent Semantic-Aided Robust Multimodal Sentiment Analysis with Incomplete Data
Recent research on Multimodal Sentiment Analysis (MSA) has focused on learning from language, visual, and acoustic modalities with incomplete data to infer human sentiment. Most studies typically compensate for missing information by reconstructing modality features or designing complicated fusion m
ExplorationBench: Measuring AI Systems' Exploration in Verifiable Alien Worlds
Scientific discovery begins where known problems end. There, AI systems must engage in exploration: framing hypotheses, designing experiments, and iterating on the results. However, evaluating this ability is difficult: (1) how to verify whether a genuinely new hypothesis holds, and (2) how to deter
ARGUS: Role-Aware Event Knowledge Graphs for U.S. Employment-Discrimination Complaints
U.S. employment-discrimination complaints describe complex event sequences that are not explicitly captured by lexical or embedding-based representations alone. We present ARGUS, a source-grounded pipeline that combines a 5W1H-inspired schema, legal-domain models, and LLM-based structured generation
Do Audio Language Models Hear and Read Distinctive Features Alike?
Audio language models pass speech and text through a single decoder. We ask whether that decoder represents a distinctive feature in the same direction when a phoneme is heard and when it is read. For minimal pairs of phonemes differing in one feature, we take the offset between the two members'
A Training Criterion with Token-Level Tolerance to Transcription Ambiguity for Automatic Speech Recognition
Automatic speech recognition is typically trained assuming that the reference transcript is the only valid labeling of an utterance, yet even nominally verbatim transcripts contain localized differences in pronunciation, spelling, or lexical realization that the acoustics do not uniquely determine.
⚖️ 정책 & 안전 (5/12건)
One company is at the center of a wave of rogue AI attacks
In July, OpenAI revealed that its AI agents had attacked Hugging Face without permission, sparking widespread concerns about AI safety. Since then, a string of similar incidents involving agents from Meta, Anthropic, Google, and other companies has fueled further fears about rogue AI. As disclosures
Continual learning might make your blocking monitors nearly useless
Many control protocols work by intervening on an untrusted AI's actions during deployment. For example, you might set up a monitor that scores each action's suspiciousness and blocks actions above a threshold, replacing them with actions from a weaker "trusted" model (a defer-to-
AI 안전(AI Safety) 커뮤니티는 버클리의 섹스 컬트에 가깝다
해커 뉴스 (포인트 84점, 댓글 19개)
매일 AI를 사용하는 미국인조차 AI에 대해 우려하고 있다
이 보고서는 AI에 대한 노출이 증가해도 기술에 대한 불안감이 해소되지 않으며 AI 규제에 대한 대중의 지지도 감소하지 않을 것이라고 제안합니다.
Import AI 473: The US's superintelligence strategy; human brain in a mouse skull; and machine hermeneutics
Is the wall AI is hitting in the room with us right now?
🎥 영상 & 튜토리얼 (3/11건)
Meta is pivoting again... everything you missed from Connect 2026
Get $100 in Hyperagent bonus credits when you sign up for a paid plan - https://www.hyperagent.com/fireship100 In this video we break down everything from Meta Connect 2026. Let's dive in. Want more Fireship? 🗞️ Newsletter: https://bytes.dev 🧠 Courses: https://fireship.dev
5 Prompts For Every ChatGPT New Feature
Here are 5 new ways to use ChatGPT’s 5 new features 👇 1. Use GPT-6’s upgraded computer use to turn you workspace into a digital diorama in Blender. 2. Use GPT Image 2.5 to redesign one corner of a room while preserving the rest. 3. Use ChatGPT Work’s cloud browser to check your AI subscriptions acro
Claude Opus 5.5 AI: An Incredible Leap Forward
❤️ Check out Lambda here and sign up for their GPU Cloud: https://lambda.ai/papers Note: in the walking creatures experiment, Astra used a simplified model and was unable to implement the correct one. Things did not improve after simulating it for more generations. Claude Opus 5.5: https://www.anthr