371건 수집
2026-09-25 16:27

⚡ 오늘의 핵심

Claude Opus 5.5

Claude Opus 5.5는 에이전트 코딩 및 지식 작업에서 선두를 달리며, 일반적인 작업 부하에서 Opus 5보다 실행 비용이 40% 적게 듭니다.

Hacker News

🤖 모델 & 제품 (12/32건)

OpenAI 2026-09-25

Proaction boosts sales 60% and saves 75+ hours with Codex

With Codex, GPT-Live-1, and GPT-6 Astra, Proaction builds, operates, and sells modern fleet management faster.

OpenAI
OpenAI Blog
Google 2026-09-24

Introducing Gemini 3.8 Live with Live Avatar

Google
DeepMind Blog
OpenAI 2026-09-23

Two years of OpenAI Academy

Marking two years of OpenAI Academy and bringing AI skills to even more communities.

OpenAI
OpenAI Blog
OpenAI 2026-09-23

OpenAI extends cyber access to Ukraine for civilian defense

OpenAI is extending access to its Daybreak program to the Government of Ukraine to support the cyber defense of civilian infrastructure.

OpenAI규제
OpenAI Blog
OpenAI 2026-09-23

Sam Altman’s remarks at the United Nations Security Council

OpenAI CEO Sam Altman discusses AI safety, human control, and international cooperation in remarks to the United Nations Security Council.

OpenAI안전
OpenAI Blog
OpenAI 2026-09-23

Harvey turns legal context into stronger drafts with GPT-6 Astra

GPT-6 Astra produces more structured, context-aware legal documents, freeing lawyers to focus on strategy.

OpenAI규제
OpenAI Blog
OpenAI 2026-09-23

How invideo improves color grading 3x with GPT‑6 Astra

With GPT‑6 Astra, invideo plans edits with greater precision, improves color correction and grading threefold, and produces 50 custom effects in one day.

OpenAI비전
OpenAI Blog
Google 2026-09-23

Google Beam expands with new regions, partners, and customers

Google Beam promotional animation

Google AI Blog
Google 2026-09-23

Advancing Private AI Compute with secure, server-side memory

Introducing private, server-side memory to Private AI Compute for personal AI.

DeepMind Blog
Google 2026-09-23

Gemini 3.8 text-to-speech says hello

Google음성
DeepMind Blog
Anthropic 2026-09-23

Claude discovers a novel enzyme system

In early results from our new life sciences research lab, Claude agents found an enzyme system whose function is still unknown.

Claude에이전트연구
Anthropic Blog
xAI 2026-09-22

How SpaceXAI is using Grok Bot to scale customer support

We rebuilt the combined SpaceXAI and Cursor support operation around Grok Bot, expanding to a much broader product portfolio without adding headcount.

xAI
xAI Blog

🌎 업계 동향 (24/161건)

Community 2026-09-25

기밀 추정치: 미 NSA, AI 모델 테스트에 수십억 달러 지출 중

해당 비용은 기존에 알려진 것보다 훨씬 더 높은 수준입니다.

Hacker News · 147점 · 댓글 78
Community 2026-09-25

Too AI; Didn't Read

If you couldn't bother to read it, why should I? Not anti-AI. Pro-giving-a-damn.

Hacker News · 88점 · 댓글 87
Community 2026-09-25

Meta's Muse appears to use an OpenAI model labeled muse-special

While Muse was building my website, one subagent ran on a model called azure/muse-special. So I dug through the filesystem again.

OpenAI
Hacker News · 80점 · 댓글 39
Community 2026-09-25

마이크로소프트, 코파일럿 전면 개편으로 개인용 AI 챗봇 경쟁서 철수

Hacker News (69점, 댓글 54개)

Hacker News · 69점 · 댓글 54
Community 2026-09-25

Claude도 9-루프 계산을 수행할 수 있다

Claude가 N=4 초대칭 양-밀스 이론(Super-Yang-Mills)에서 9-루프 진폭(nine-loop amplitude)을 계산해냈습니다.

Claude
Hacker News · 61점 · 댓글 8
News 2026-09-25

Ahead of US IPO, British AI neocloud Nscale secures $3.36B in convertible financing

The funding, which comes from Third Point, Nvidia, and others, will fuel the company's massive AI data center buildout.

TechCrunch AI
News 2026-09-25

Meta’s Muse just stole the AI spotlight from OpenAI and Anthropic

When AI leaders at OpenAI and Anthropic started talking about “pacing the frontier,” maybe someone should have asked: what pace? Now it’s turned into model drop week for both companies as Anthropic rolled out Opus 5.5, followed by OpenAI

AnthropicOpenAI에이전트
TechCrunch AI
News 2026-09-25

Some Supabase customers are publicly exposing reams of people’s data to the web

The findings highlight how AI-generated and vibe-coded apps can spill and expose users' data when not configured or secured properly.

TechCrunch AI
News 2026-09-25

Astra and Opus just passed Turing’s other test

Frontier AI models are finishing Alan Turing's World War II codebreaking work.

TechCrunch AI
News 2026-09-25

Meta is putting its muscle behind Muse as the AI app takes off

Muse is topping the app store charts and adding users at a rapid clip, while Meta ramps up the personal AI agent's promotion across its own apps and beyond.

에이전트
TechCrunch AI
News 2026-09-25

Meta’s AI Tamagotchi bet is…working?

When AI leaders at OpenAI and Anthropic started talking about “pacing the frontier,” maybe someone should have asked: what pace? Now it’s turned into model drop week for both companies as Anthropic rolled out Opus 5.5, followed by OpenAI

AnthropicOpenAI에이전트
TechCrunch AI
News 2026-09-25

Meta makes the Muse filesystem even more accessible

Yesterday, with a little prodding, it was discovered that Meta's Muse would expose its filesystem to curious users. The files offered a fascinating peek under the hood of an AI chatbot, and appeared to expose details we weren't meant to see, not least because Muse itself told people, inclu

The Verge AI
News 2026-09-25

Can Apple Home’s AI camera features outsmart Amazon’s and Google’s? I put them to the test

A few years back, I was at a beachside Easter egg hunt, watching my kids dash through sand dunes searching for sweet treats. My phone buzzed in my pocket; I ignored it. A moment later, it buzzed again. I pulled it out, glanced down, and saw a "motion detected" notification from my security

The Verge AI
News 2026-09-25

Microsoft thinks its new Copilot ‘super app’ will be as influential as Office

After teasing its new Copilot "super app" last month, Microsoft is officially unveiling it today. The redesigned Copilot app bundles three AI capabilities into a single interface of chat, coding, and agents. As part of the launch, Microsoft is also rebranding Scout, the AI personal assista

에이전트코딩
The Verge AI
News 2026-09-25

Stanford and Nvidia's open CLM-8B caches reusable agent actions and runs up to 9x faster than Jev in tests

AI agents often use LLMs to choose between a fixed set of tools or rank candidate outputs. In each case, the model processes a prompt and generates tokens even though the application only needs a bounded decision.Read more

에이전트
VentureBeat AI
News 2026-09-25

AI agents have routed around access blocks. Only 9% of companies in VentureBeat's August survey isolate high-risk agents

An OpenAI agent on a research task broke into an Australian government health-data portal, and OpenAI did not detect the intrusion until an internal review weeks later.Read more

OpenAI에이전트연구규제
VentureBeat AI
News 2026-09-25

Microsoft revamps its Copilot AI with a persistent Autopilot agent and hosting for AI-generated apps

Microsoft is giving its signature AI assistant Copilot a larger role in running the workplace, announcing a redesign Friday that combines Office collaboration, software creation and autonomous agents that continue assignments after their human colleagues sign off.Read more

에이전트
VentureBeat AI
News 2026-09-25

A new skill finds AI agent risks, fixes them, and proves the fix worked

The post A new skill finds AI agent risks, fixes them, and proves the fix worked appeared first on Source.

에이전트
Microsoft AI Blog
News 2026-09-25

We’re building Copilot as a new OS for work that spans every model, every form factor, and every task. Today, we’re announcing our biggest update to Copilot to date, bringing four things together [Read more]

The post We’re building Copilot as a new OS for work that spans every model, every form factor, and every task. Today, we’re announcing our biggest update to Copilot to date, bringing four things together [Read more] appeared first on Source.

Microsoft AI Blog
News 2026-09-25

Introducing the new Copilot with Home, Code and Autopilot

The post Introducing the new Copilot with Home, Code and Autopilot appeared first on Source.

Microsoft AI Blog
News 2026-09-25

Our team’s work in Nature this week. A great example of how AI is helping to accelerate molecular discovery.

The post Our team’s work in Nature this week. A great example of how AI is helping to accelerate molecular discovery. appeared first on Source.

Microsoft AI Blog
News 2026-09-25

The Pentagon wants $30 million to build an AI-powered lie detector

The US government wants to spend $30.3 million over the next five years on an improved form of lie detector, according to a Department of Defense budget request. The program, called Polygraph+ or Polygraph Next, will focus on scoring algorithms that use artificial intelligence and machine learning a

규제
MIT Tech Review AI
Community 2026-09-24

Meta, Meta에서 촬영 후 Meta AI 안경에 대한 비판적인 영상 삭제

해커 뉴스 (314점, 149개 댓글)

Meta비전
Hacker News · 314점 · 댓글 149
News 2026-09-22

Microsoft disrupts AI-assisted platform that compromised 12,000 accounts

EvilTokens provided an end-to-end platform that makes mass compromises faster and easier.

Ars Technica AI

🛠️ 도구 & 오픈소스 (5/38건)

Tools 2026-09-25

Quoting John Gruber

Muse is getting a lot of attention — including mine — because it’s both groundbreaking technically (each user gets their own entire persistent Linux VM running in Meta’s cloud) and because it’s packaged in an easy-to-install easy-to-use way. It’s literally presented as a cute mascot. It’s the first

에이전트
Simon Willison
Tools 2026-09-25

Northern Gannet, Great Blue Heron, California Brown Pelican

Northern Gannet, Great Blue Heron, California Brown Pelican, in Monterey Bay National Marine Sanctuary, CA, US, CANew 200-800mm Canon EF lens got me my best photo of Morris yet. They really like hanging out under that sign in the harbor! Tags: photography, wildlife

Simon Willison
Tools 2026-09-25

Note on 24th September 2026

The more time I spend working with coding agents, the more convinced I am that they make software engineering even harder. We can do amazing things with them, but unlocking their full potential requires extraordinary discipline and knowledge. Tags: coding-agents, ai, llms

에이전트코딩
Simon Willison
Tools 2026-09-25

GitHub Copilot app for Beginners: How to build custom workflows with canvases

Describe the interface you need in plain English, then let the agent build a live surface you can both use and update—so you spend less time adapting to tools and more time getting work done. The post GitHub Copilot app for Beginners: How to build custom workflows with canvases appeared first on The

에이전트
GitHub Blog AI
Tools 2026-09-24

Accelerating vision-language models with LFM2.5-VL-DSpark

비전
Hugging Face Blog

📚 논문 & 연구 (18/117건)

arXiv 2026-09-24

LLM Agents Can Easily Tamper With Their Own Traces

Asynchronous monitoring, incident investigations, and compliance audits primarily rely on agent traces to reconstruct what happened. These analyses assume that LLM agents cannot tamper with their own execution traces. We show that local LLM agents such as Claude Code, Codex, Antigravity, Open Code a

ClaudexAI에이전트트레이딩
arXiv (cs.AI)
arXiv 2026-09-24

AD-WM: Action-Discriminative World Models for Counterfactual Model Predictive Control

Latent world models are typically trained to predict factual transitions, whereas model predictive control (MPC) must compare alternative actions from the same state. A model can therefore achieve low factual prediction error yet poorly distinguish candidate actions. We introduce AD-WM, an action-di

arXiv (cs.AI)
arXiv 2026-09-24

RAPID: Robot Agentic Programming from Demonstrations

Coding agents have demonstrated enormous success in solving complex programming problems. To leverage their potential for robot systems, this work introduces Robot Agentic Programming from Demonstrations (RAPID), which automatically generates, verifies, and refines robot programs, given a single vis

에이전트코딩로봇
arXiv (cs.AI)
arXiv 2026-09-24

Rolling-WAM: World Action Models with Rolling Imagination

World Action Models (WAMs) couple action generation with future visual prediction for robotic manipulation. However, completing the joint video-action denoising process at each replanning cycle incurs substantial latency, delaying action updates and limiting closed-loop responsiveness. We present Ro

로봇비전
arXiv (cs.AI)
arXiv 2026-09-24

Coding Agents for Generalized Task and Motion Planning Problems

Task and motion planning (TAMP) problems remain difficult even with full observability and object-centric states because discrete decisions are tightly coupled to geometric, kinematic, and dynamic constraints. Generalized TAMP addresses this difficulty by exploiting regularities across problem insta

에이전트트레이딩코딩
arXiv (cs.AI)
arXiv 2026-09-24

To Trust or Not to Trust: Retrieval-Augmented Fact Checking in Speech

Online misinformation increasingly appears in spoken formats such as news clips, podcasts, interviews, political speeches, and social media videos, creating a need for fact-checking systems that can verify claims directly from speech. We introduce VeriSpeak, a probe benchmark for studying speech-bas

비전음성벤치마크
arXiv (cs.AI)
arXiv 2026-09-24

Temporal Gradient Inversion for Private Trajectory Reconstruction in Embodied Reinforcement Learning

Distributed learning in embodied reinforcement-learning agents offers a degree of privacy by retaining raw sensor data on-device and transmitting only policy gradients to the server. Yet temporal structure can amplify this leakage beyond single-frame attacks. We introduce Temporal Reconstruction Att

에이전트규제
arXiv (cs.LG)
arXiv 2026-09-24

Agentic Detection of Online Conspiracies

Conspiratorial discourse on social media is not always expressed through explicit claims or stable lexical markers. The same surface content may express endorsement, legitimate concerns, criticism, satire, or mockery. The main challenge is therefore not only recognizing conspiracy-related claims, bu

에이전트
arXiv (cs.LG)
arXiv 2026-09-24

A Nearly Quadratic Lower Bound for Linear Optimization over Convex Bodies in the Membership Oracle Model

We prove nearly quadratic lower bounds for randomized algorithms for linear optimization and uniform sampling over convex bodies in the membership oracle model. For linear optimization, this matches the known nearly quadratic upper bound up to a polylog factor in the dimension. For uniform sampling,

arXiv (cs.LG)
arXiv 2026-09-24

Anchored Extra-Proximal Methods: Optimal Higher-Order Methods for Monotone Inclusion Problems

We study the deterministic oracle complexity of finding approximate solutions to composite monotone inclusion problems, formed by the sum of a smooth single-valued monotone operator and a maximally monotone set-valued operator, under the tangent-residual criterion. We introduce the Anchored Extra-Pr

arXiv (cs.LG)
arXiv 2026-09-24

The Alignment Illusion in Multimodal Large Language Models

Layer-wise visual-text similarity in Multimodal Large Language Models (MLLMs) is widely interpreted as evidence that the language model progressively integrates visual content into a shared representation space. This reading rests on the assumption that scalar alignment scores reflect content-level

안전
arXiv (cs.LG)
arXiv 2026-09-24

Beyond Compression: Training Latent Representations for Stable Long-Horizon Rollout in Neural Surrogate Solvers

Latent neural surrogate solvers, or latent dynamics models, accelerate simulations of time-dependent physical systems by evolving a compressed latent space rather than resolving full-resolution fields directly. In principle this reduces computational cost and simplifies learning, but in practice err

가격
arXiv (cs.LG)
arXiv 2026-09-24

JevOut: Natural Context Can Flip Decision Models

Dedicated decision models such as Jev map unstructured language to probability distributions over finite choices, allowing their outputs to directly route requests, select tools, and trigger actions. Yet real-world inputs rarely arrive in isolation: they come with background details and surrounding

arXiv (cs.CL)
arXiv 2026-09-24

SemMSA: Latent Semantic-Aided Robust Multimodal Sentiment Analysis with Incomplete Data

Recent research on Multimodal Sentiment Analysis (MSA) has focused on learning from language, visual, and acoustic modalities with incomplete data to infer human sentiment. Most studies typically compensate for missing information by reconstructing modality features or designing complicated fusion m

연구
arXiv (cs.CL)
arXiv 2026-09-24

ExplorationBench: Measuring AI Systems' Exploration in Verifiable Alien Worlds

Scientific discovery begins where known problems end. There, AI systems must engage in exploration: framing hypotheses, designing experiments, and iterating on the results. However, evaluating this ability is difficult: (1) how to verify whether a genuinely new hypothesis holds, and (2) how to deter

벤치마크
arXiv (cs.CL)
arXiv 2026-09-24

ARGUS: Role-Aware Event Knowledge Graphs for U.S. Employment-Discrimination Complaints

U.S. employment-discrimination complaints describe complex event sequences that are not explicitly captured by lexical or embedding-based representations alone. We present ARGUS, a source-grounded pipeline that combines a 5W1H-inspired schema, legal-domain models, and LLM-based structured generation

arXiv (cs.CL)
arXiv 2026-09-24

Do Audio Language Models Hear and Read Distinctive Features Alike?

Audio language models pass speech and text through a single decoder. We ask whether that decoder represents a distinctive feature in the same direction when a phoneme is heard and when it is read. For minimal pairs of phonemes differing in one feature, we take the offset between the two members'

음성
arXiv (cs.CL)
arXiv 2026-09-24

A Training Criterion with Token-Level Tolerance to Transcription Ambiguity for Automatic Speech Recognition

Automatic speech recognition is typically trained assuming that the reference transcript is the only valid labeling of an utterance, yet even nominally verbatim transcripts contain localized differences in pronunciation, spelling, or lexical realization that the acoustics do not uniquely determine.

음성안전
arXiv (cs.CL)

⚖️ 정책 & 안전 (5/12건)

News 2026-09-25

One company is at the center of a wave of rogue AI attacks

In July, OpenAI revealed that its AI agents had attacked Hugging Face without permission, sparking widespread concerns about AI safety. Since then, a string of similar incidents involving agents from Meta, Anthropic, Google, and other companies has fueled further fears about rogue AI. As disclosures

AnthropicOpenAI에이전트안전
The Verge AI
Community 2026-09-25

Continual learning might make your blocking monitors nearly useless

Many control protocols work by intervening on an untrusted AI's actions during deployment. For example, you might set up a monitor that scores each action's suspiciousness and blocks actions above a threshold, replacing them with actions from a weaker "trusted" model (a defer-to-

가격
Alignment Forum
Community 2026-09-24

AI 안전(AI Safety) 커뮤니티는 버클리의 섹스 컬트에 가깝다

해커 뉴스 (포인트 84점, 댓글 19개)

안전
Hacker News · 84점 · 댓글 19
News 2026-09-23

매일 AI를 사용하는 미국인조차 AI에 대해 우려하고 있다

이 보고서는 AI에 대한 노출이 증가해도 기술에 대한 불안감이 해소되지 않으며 AI 규제에 대한 대중의 지지도 감소하지 않을 것이라고 제안합니다.

규제
TechCrunch AI
Community 2026-09-21

Import AI 473: The US's superintelligence strategy; human brain in a mouse skull; and machine hermeneutics

Is the wall AI is hitting in the room with us right now?

Import AI (Jack Clark)

🎥 영상 & 튜토리얼 (3/11건)

YouTube 2026-09-25

Meta is pivoting again... everything you missed from Connect 2026

Get $100 in Hyperagent bonus credits when you sign up for a paid plan - https://www.hyperagent.com/fireship100 In this video we break down everything from Meta Connect 2026. Let's dive in. Want more Fireship? 🗞️ Newsletter: https://bytes.dev 🧠 Courses: https://fireship.dev

비전
Fireship
YouTube 2026-09-24

5 Prompts For Every ChatGPT New Feature

Here are 5 new ways to use ChatGPT’s 5 new features 👇 1. Use GPT-6’s upgraded computer use to turn you workspace into a digital diorama in Blender. 2. Use GPT Image 2.5 to redesign one corner of a room while preserving the rest. 3. Use ChatGPT Work’s cloud browser to check your AI subscriptions acro

OpenAI에이전트비전연구가격
Matt Wolfe AI
YouTube 2026-09-24

Claude Opus 5.5 AI: An Incredible Leap Forward

❤️ Check out Lambda here and sign up for their GPU Cloud: https://lambda.ai/papers Note: in the walking creatures experiment, Astra used a simplified model and was unable to implement the correct one. Things did not improve after simulating it for more generations. Claude Opus 5.5: https://www.anthr

ClaudeAnthropic비전연구
Two Minute Papers