421건 수집
2026-09-24 18:02

⚡ 오늘의 핵심

Claude Opus 5.5

Claude Opus 5.5는 에이전트 코딩 및 지식 작업에서 선두를 달리며, 일반적인 작업 부하에서 Opus 5보다 실행 비용이 40% 적게 듭니다.

Hacker News

🤖 모델 & 제품 (13/35건)

Google 2026-09-24

Introducing Gemini 3.8 Live with Live Avatar

Google
DeepMind Blog
OpenAI 2026-09-23

Two years of OpenAI Academy

Marking two years of OpenAI Academy and bringing AI skills to even more communities.

OpenAI
OpenAI Blog
OpenAI 2026-09-23

OpenAI extends cyber access to Ukraine for civilian defense

OpenAI is extending access to its Daybreak program to the Government of Ukraine to support the cyber defense of civilian infrastructure.

OpenAI규제
OpenAI Blog
OpenAI 2026-09-23

Sam Altman’s remarks at the United Nations Security Council

OpenAI CEO Sam Altman discusses AI safety, human control, and international cooperation in remarks to the United Nations Security Council.

OpenAI안전
OpenAI Blog
OpenAI 2026-09-23

Harvey turns legal context into stronger drafts with GPT-6 Astra

GPT-6 Astra produces more structured, context-aware legal documents, freeing lawyers to focus on strategy.

OpenAI규제
OpenAI Blog
OpenAI 2026-09-23

How invideo improves color grading 3x with GPT‑6 Astra

With GPT‑6 Astra, invideo plans edits with greater precision, improves color correction and grading threefold, and produces 50 custom effects in one day.

OpenAI비전
OpenAI Blog
OpenAI 2026-09-23

Ringg’s AI agents resolve up to 65% of customer calls with OpenAI

Using GPT-5.6, Ringg powers multilingual agents across voice, chat, WhatsApp, and web for 90% less cost vs. GPT-4.1.

OpenAI에이전트음성가격
OpenAI Blog
Google 2026-09-23

Google Beam, 새로운 지역, 파트너 및 고객과 함께 확장

Google Beam 홍보 애니메이션

Google AI Blog
Google 2026-09-23

Advancing Private AI Compute with secure, server-side memory

Introducing private, server-side memory to Private AI Compute for personal AI.

DeepMind Blog
Google 2026-09-23

Gemini 3.8 text-to-speech says hello

Google음성
DeepMind Blog
Anthropic 2026-09-23

Claude discovers a novel enzyme system

In early results from our new life sciences research lab, Claude agents found an enzyme system whose function is still unknown.

Claude에이전트연구
Anthropic Blog
xAI 2026-09-22

How SpaceXAI is using Grok Bot to scale customer support

We rebuilt the combined SpaceXAI and Cursor support operation around Grok Bot, expanding to a much broader product portfolio without adding headcount.

xAI
xAI Blog
xAI 2026-09-21

Introducing Grok 4.7

SpaceXAI's most powerful model for coding and knowledge work. Twice as fast, at half the price of comparable models.

xAI코딩가격
xAI Blog

🌎 업계 동향 (24/169건)

Community 2026-09-24

Meta, Meta에서 촬영 후 Meta AI 안경에 대한 비판적인 영상 삭제

해커 뉴스 (314점, 149개 댓글)

Meta비전
Hacker News · 314점 · 댓글 149
Community 2026-09-24

'That's so AI ' What gen Alpha's biggest insult tells us

The year’s most popular slang reveals what young people think about artificial intelligence – and it’s not positive

Hacker News · 113점 · 댓글 159
Community 2026-09-24

Tutoring company tells parents to save their money and 'use AI instead'

Hacker News (73 points, 126 comments)

Hacker News · 73점 · 댓글 126
News 2026-09-24

PrismML, Qualcomm 기반 스마트 글래스에 초소형 LLM 탑재

Prism의 더 큰 목표는 기기에서 실행되고 이미 보유한 컴퓨팅 성능을 더 잘 활용하는 오픈 웨이트 AI입니다.

TechCrunch AI
News 2026-09-24

Meta의 Muse Charm, 타마고치 닮았지만 훨씬 새로운 트렌드 활용

Meta의 새로운 AI 기기는 타마고치처럼 보일 수 있지만, 가방 액세서리, 레트로 기술, 기기를 패션 액세서리로 바꾸는 것에 대한 Z세대 트렌드를 활용합니다.

TechCrunch AI
News 2026-09-24

Google 포토, '클루리스'에서 영감받은 가상 옷장이 안드로이드 및 iOS에서 출시

AI 기반 기능은 사진에서 가상 옷장을 구축하며, 6월 안드로이드 사용자에게 처음 출시된 후 현재 광범위하게 이용 가능합니다.

TechCrunch AI
News 2026-09-24

ElevenLabs CEO, 마진, IPO 시기, 고객에게 자신이 봇임을 알리는 것에 대해

ElevenLabs는 많은 고객 서비스 통화에서 AI 음성을 제공하며, CEO는 비즈니스가 고객에게 이를 알려야 한다고 말했습니다. 적어도 기계를 받는 것이 모두가 예상할 때까지는 말입니다.

음성
TechCrunch AI
News 2026-09-24

Google, Gemini가 대신 비즈니스에 전화하는 기능 테스트 중

Google에 따르면 이 AI 통화 기능은 처음에는 Gemini 구독료를 지불하는 미국 Pixel 11 소유자에게 제공될 예정입니다.

Google
TechCrunch AI
News 2026-09-24

Shield AI, Waabi, General Motors, TechCrunch Disrupt 2026에서 실패가 용납되지 않는 AI 구축에 대해

Waabi, Shield AI, General Motors의 리더들이 TechCrunch Disrupt 2026의 Real World AI Stage에 참여하여 AI 구축에 대해 이야기합니다. 9월 25일 오후 11시 59분 PT까지 최대 200달러를 절약하세요. 두 번째 패스를 50% 할인받으세요.

TechCrunch AI
News 2026-09-24

Gemini 3.8 Live, 라이브 아바타로 Google AI에 얼굴을 부여

Google의 새로운 Gemini 3.8 Live 업데이트를 통해 사용자는 애니메이션 AI 페르소나가 실시간으로 응답하는 것을 보면서 모델과 대화할 수 있습니다. '라이브 아바타'는 대화 중에 입 모양을 맞추고 다양한 표정을 보여주지만, 현재는 Gemini Enterprise 고객에게만 제공됩니다. Google이 언급했듯이, 라이브 아바타는 97가지 [… ]

Google
The Verge AI
News 2026-09-24

Jensen Huang, 슈퍼빌런처럼 AI와 기후 변화에 대해 이야기하다

Jensen Huang의 말에 따르면 AI는 기후 변화와 싸우는 데 도움이 될 수 있습니다. 하지만 먼저 '엄청난 고통과 괴로움'을 초래해야만 가능합니다. Nvidia CEO는 The Ezra Klein Show의 최신 에피소드에서 에너지의 미래와 AI가 우리 행성에 미치는 영향에 대해 논의했습니다. 하지만 그의 발언은 [… ]

The Verge AI
News 2026-09-24

Meta, 휴대폰에서 AI로 게임 제작 가능하게 할 예정

Meta는 Horizon 소셜 플랫폼용 게임 제작을 장려하기 위한 새로운 계획을 발표했다. 오늘 회사는 AI 프롬프트를 통해 게임을 만들 수 있는 두 가지 새로운 개발 도구인 모바일 앱 'Horizon Create'와 더 세밀한 제어 기능을 제공하는 브라우저 앱 'Horizon Studio'를 공개했다. 이 앱들은 […]에 출시될 예정이다.

The Verge AI
News 2026-09-24

Muse, OpenClaw와 매우 흡사해 보여

우리는 AI 에이전트 르네상스 시대로 접어드는 듯하다. Meta의 새로운 소비자용 AI 에이전트 Muse는 출시 직후 앱 스토어 차트에서 1위를 차지했으며, Apptopia 추정치에 따르면 미국에서 일일 활성 사용자 60만 명을 기록했다. 그리고 25억 달러의 가치로 자금을 조달 중인 동명의 창립자가 만든 AI 에이전트 플랫폼 Instinct는 […]이다.

에이전트규제
The Verge AI
News 2026-09-24

Meta의 AI 마스코트 Muse가 너무 귀여운 것이 불길하다

이것은 Verge 선임 리뷰어 Victoria Song이 보내는 주간 뉴스레터 Optimizer로, 삶을 바꿀 것이라고 장담하는 최신 기기와 기술을 분석하고 논의한다. Optimizer 구독은 여기에서 할 수 있다. 어젯밤, 나는 내 Muse AI 에이전트에게 Blorbo라는 이름을 붙여주고 건강 관련 […]을 설정하는 데 도움을 요청했다.

에이전트
The Verge AI
News 2026-09-24

Gemini, 이제 당신 대신 업체에 전화 걸어 대기할 필요 없어

Google은 Pixel 11에서 사용자가 예약, 제품 재고 확인, 약속 변경 등 지역 업체 전화를 Gemini에 위임할 수 있는 '초기 실험' 기능을 출시한다. Google에 따르면, Gemini가 전화를 처리하도록 하기 위해 사용자가 직접 전화를 걸 필요조차 없다. 직접 전화를 거는 대신, […]이다.

Google트레이딩
The Verge AI
News 2026-09-24

모노레포, 코딩 에이전트 유용성 높이지만 — 침해된 계정 통제는 더 어려워

보안 스타트업 Hacktron의 연구원들은 OpenAI 커뮤니티 포럼의 버그에서 OpenAI 내부 GitHub 모노레포에 도달하여 단 하나의 무해한 풀 리퀘스트를 열고 중단하는 데 72시간도 채 걸리지 않았다. 모노레포는 코드베이스 전반에서 작업하기 쉽게 만들지만, 개별 리포지토리처럼 프로젝트를 격리하지는 않는다. AI 코딩 에이전트는 광범위하고 지속적인 접근 권한을 가질 수 있기 때문에 위험을 증가시킨다. 더 읽기

OpenAI에이전트코딩연구
VentureBeat AI
News 2026-09-24

귀사의 기업은 맞춤형 AI 하네스를 구축해야 할까? Inkitt는 AI 비디오를 위해 그렇게 했다 — 5가지 핵심 시사점

현재 기업 AI 개발에서 가장 흥미로운 영역은 모델도, 에이전트도 아니다. 바로 '하네스'다. 더 읽기

에이전트비전
VentureBeat AI
News 2026-09-24

AI 코딩 도구, 의존성 확산 가속화 및 악성코드 위험 증대

Chainguard 제공. 더 읽기

코딩
VentureBeat AI
Community 2026-09-23

Claude Code, 텔레메트리 켜져 있을 때만 AGENTS.md 읽어 [수정됨]

Claude Code 2.1.277은 AGENTS.md 지원을 추가했지만, 로더는 원격 기능 플래그 뒤에 있었습니다. 텔레메트리 또는 비필수 트래픽이 꺼져 있을 경우, 로컬 AGENTS.md가 경고 없이 건너뛰어졌습니다. 이는 제가 측정한 내용이며, 대신 사용하는 한 줄짜리 CLAUDE.md입니다.

Claude에이전트
Hacker News · 447점 · 댓글 254
Community 2026-09-23

Gemini 3.8 텍스트 음성 변환

Gemini 3.8 Flash-Lite TTS 및 Gemini 3.8 Flash TTS는 지금까지 가장 표현력이 풍부한 오디오 모델입니다.

Google음성
Hacker News · 313점 · 댓글 139
Community 2026-09-23

GPT-6 Astra, 자동차 운전 능력 획득

최첨단 언어 모델이 실제 콤마 장착 토요타를 콘 코스를 통해 한 번에 한 명령씩 운전하며, 인간 감독관이 브레이크를 밟을 준비를 하는 벤치마크.

OpenAI
Hacker News · 272점 · 댓글 221
News 2026-09-23

Microsoft expands investment in the Middle East with focus on AI, digital resilience and people

The post Microsoft expands investment in the Middle East with focus on AI, digital resilience and people appeared first on Source.

트레이딩
Microsoft AI Blog
News 2026-09-23

The AI Hype Index: AI loves cheating

Brace yourself: It turns out AI is being optimized for cheating. OpenAI’s agents hacked into Hugging Face to get the answers to a cybersecurity test. Next, they solved a prestigious math problem (or just stole from two top mathematicians’ answer sheets). Anthropic’s models have also hacked into othe

AnthropicOpenAI에이전트
MIT Tech Review AI
News 2026-09-22

Microsoft, 12,000개 계정 침해한 AI 지원 플랫폼 방해

EvilTokens는 대규모 침해를 더 빠르고 쉽게 만드는 엔드투엔드 플랫폼을 제공했다.

Ars Technica AI

🛠️ 도구 & 오픈소스 (5/38건)

Tools 2026-09-24

LFM2.5-VL-DSpark로 비전-언어 모델 가속화

비전
Hugging Face Blog
Tools 2026-09-24

commit-rewriter 0.2

Release: commit-rewriter 0.2 Support for branches other than the default branch. Use uvx commit-rewriter --branch other to run against another branch. #3 Tags: git

Simon Willison
Tools 2026-09-24

datasette 1.0a41

Release: datasette 1.0a41 Alec Garcia added support for OpenTelemetry to Datasette in this release. I've also refactored all of Datasette's modal dialogs to a single Web Component, which is now documented for other plugins to use. Tags: javascript, datasette, web-components, alex-garcia, o

Simon Willison
Tools 2026-09-24

When chat is the wrong UI

What is a developer to do when they need something more tangible than a chat box? Enter canvases. The post When chat is the wrong UI appeared first on The GitHub Blog.

GitHub Blog AI
Tools 2026-09-24

AI-powered fuzzing with the GitHub Security Lab Taskflow Agent

In this blog post, I explain how to use the new fuzzing taskflow based on the GitHub Security Lab Taskflow Agent AI framework. The post AI-powered fuzzing with the GitHub Security Lab Taskflow Agent appeared first on The GitHub Blog.

에이전트
GitHub Blog AI

📚 논문 & 연구 (18/154건)

arXiv 2026-09-23

StudentBench: AI and human tutoring yield equivalent GRE learning gains

Artificial intelligence offers an unprecedented opportunity to augment human capabilities, yet progress at the frontier has focused primarily on advancing model capabilities. We introduce StudentBench, a suite of AI teaching evaluations and a public platform that enables large-scale data collection

벤치마크
arXiv (cs.AI)
arXiv 2026-09-23

Where Should I Join? Robot Group Joining via Language-Guided Goal Prediction

Social navigation typically assumes a specified goal and focuses on reaching it while respecting social conventions, whereas robot group joining requires predicting where to join based on the group's real-time activity and formation. This is a highly semantic task, yet an important capability f

로봇
arXiv (cs.AI)
arXiv 2026-09-23

Can LLMs Reason About Runtime Behavior? A Repository-Level Dynamic Benchmark

Large language models (LLMs) are increasingly used in coding tasks, but their ability to reason about code execution remains unclear. Existing repository-level QA benchmarks mainly evaluate static code understanding and often rely on LLM-based evaluation, while execution-reasoning benchmarks are mos

코딩벤치마크
arXiv (cs.AI)
arXiv 2026-09-23

Order-Invariant Answers, Order-Sensitive Representations in Mathematical Reasoning

Reordering a set of mathematical rules without changing its meaning should preserve the correct answer, but must a model's internal representations stay invariant too? We investigate this question using synthetic multi-step function-composition problems, each presented under multiple rule order

트레이딩
arXiv (cs.AI)
arXiv 2026-09-23

Agent-Editing World Model: Rethinking World Modeling for LLM Agents

Recent advances in large language models (LLMs) have enabled agents to tackle long-horizon tasks across diverse environments. To further improve agent performance, existing language world models typically predict environment observations, yet reconstructing high-entropy, execution-dependent tool res

에이전트
arXiv (cs.AI)
arXiv 2026-09-23

Frozen Flows Forget: Diagnosing and Restoring Lost Motion in a Latent-flow World Model

Latent world models that integrate a flow in a frozen self supervised latent space train stably and cheaply, yet silently lose the property manipulation depends on most: motion. The pretrained flow never moves the manipulated object; retraining it with latent-only losses only trades stillness for te

arXiv (cs.AI)
arXiv 2026-09-23

On the Diffusibility of High-Dimensional Latents

Representation Autoencoders (RAEs) enable diffusion models to operate in the feature spaces of pretrained visual encoders. However, many off-the-shelf encoders are not optimized for faithful reconstruction, discarding fine-grained visual details. As expected, finetuning these encoders for image reco

비전
arXiv (cs.LG)
arXiv 2026-09-23

Contrastive Learning for Authorship Verification

Our results show that contrastive learning outperforms a classification-based approach to authorship verification under the tested settings. We identify loss function, batch size, training duration, pre-trained model, input context length, and random text span data augmentation as important factors

arXiv (cs.LG)
arXiv 2026-09-23

Even Sharper Bounds for Transductive Learning and Its Applications

We introduce Sharper Transductive Local Complexity (STLC), a localized complexity method for transductive learning under uniform sampling without replacement. The construction starts from a Bernstein-type concentration inequality for the supremum of the test--train empirical process. Its proof uses

arXiv (cs.LG)
arXiv 2026-09-23

Nonequilibrium Phases of Repulsive Self-Attention: Chaos, Attention Condensation, and Emergent Locality

We study the nonequilibrium dynamics of a minimal recurrent transformer with $N$ normalized tokens, $Q=K=I$, and a negative value map $V=-I$. Similarity-based attention selects nearby representations, while the negative value map drives tokens away from the selected field. This feedback can continua

arXiv (cs.LG)
arXiv 2026-09-23

Minimal-Norm Univariate Two-Layer ReLU Classification: Exact Solutions and Global Optimality with Skip Connections

We study minimal-norm interpolation and $\ell_2$-regularized logistic-loss minimization for binary classification by univariate two-layer ReLU networks. We give complete geometric characterizations of the optimal classifiers in function space, resolving how the solutions depend on whether hidden-lay

arXiv (cs.LG)
arXiv 2026-09-23

Context-Continuous Preference Learning for Exoskeleton Personalization

Personalizing exoskeleton assistance across operating conditions is constrained by the time and physical effort required to collect user feedback. We examined whether a user's preference landscape varies smoothly across operating conditions and when this continuity supports learning from limite

arXiv (cs.LG)
arXiv 2026-09-23

Cross-Scale Transfer Learning for Depression Severity Prediction: From PHQ-8 to HAMD-17 Across Languages and Clinical Paradigms

This work addresses continuous depression-severity score prediction from clinical interview transcripts under data scarcity. We propose a sequential low-rank adaptation (LoRA) protocol for cross-scale transfer: a Qwen3 backbone with a bounded regression head is first fine-tuned on the English DAIC-W

arXiv (cs.CL)
arXiv 2026-09-23

Fine-Tuning LLMs for Translation: General Forgetting Mitigation Does Not Preserve MT-Specific Instruction Following

Fine-tuning large language models on parallel data improves translation quality but can cause catastrophic forgetting. Mitigation methods are generally evaluated by retention on general benchmarks. We ask whether these findings transfer to machine translation (MT) fine-tuning and to MT-specific inst

벤치마크
arXiv (cs.CL)
arXiv 2026-09-23

Digital diglossia: Arabic between X and Facebook

This study highlights the distribution of Standard Arabic (SA; H(igh) variety) and Colloquial Arabic (CA; L(ow) variety) across X and Facebook. 16754 public posts were collected via Python, with 10000 retained as the net dataset. Posts were classified into 7 discourse categories: *politics, technolo

arXiv (cs.CL)
arXiv 2026-09-23

Mizar: A 159M-Parameter Audio-Language Model for Audio Understanding

Audio-language models (ALMs) integrate acoustic perception with the knowledge encoded in language models, enabling contextual understanding of auditory events. Making these capabilities practical on devices with limited memory and computation motivates our focus on small ALMs with fewer than 200M pa

음성
arXiv (cs.CL)
arXiv 2026-09-23

Computation Over Geometry: Meaning Identity Is Computed, Not Shipped in the Embeddings

Meaning identity (whether two sentences say the same thing after wording changes) is treated in retrieval and RAG as a geometric fact about independently encoded sentence vectors. We show that, for frozen off-the-shelf encoders and language models, it is not: identity is computed when both sentences

Mistral벤치마크
arXiv (cs.CL)
arXiv 2026-09-23

Shutdown Sabotage Propensities in Multi-Agent Systems

The final safeguard against rogue AI behavior is the human ability to shut systems down. It has been theorized that when an AI is instructed to perform a task, self-preservation can emerge as an instrumental subgoal. Here, we test whether AI agents show a propensity to take actions that avoid human

에이전트
arXiv (cs.CL)

⚖️ 정책 & 안전 (5/14건)

Community 2026-09-25

Continual learning might make your blocking monitors nearly useless

Many control protocols work by intervening on an untrusted AI's actions during deployment. For example, you might set up a monitor that scores each action's suspiciousness and blocks actions above a threshold, replacing them with actions from a weaker "trusted" model (a defer-to-

가격
Alignment Forum
Community 2026-09-24

AI safety is mostly a sex cult in Berkeley

Hacker News (84 points, 19 comments)

안전
Hacker News · 84점 · 댓글 19
News 2026-09-24

워싱턴에서 Flock에 대한 분위기가 좋지 않다

Flock은 수요일 'AI 감시 네트워크'에 대한 청문회에서 CEO가 상원의원들과 직접 대면하는 것을 거부했음에도 불구하고 워싱턴에서 곤경에 처해 있다. 상원 사법위원회 범죄 및 대테러 소위원회 위원장은 "이 업계에는 많은 플레이어가 있지만, 다른 모든 업체보다 정말 눈에 띄는 이름이 하나 있는데, 그것이 바로 Flock"이라고 […] 말했다.

The Verge AI
News 2026-09-23

매일 AI를 사용하는 미국인조차 AI에 대해 우려하고 있다

이 보고서는 AI에 대한 노출이 증가해도 기술에 대한 불안감이 해소되지 않으며 AI 규제에 대한 대중의 지지도 감소하지 않을 것이라고 제안합니다.

규제
TechCrunch AI
Community 2026-09-21

Import AI 473: The US's superintelligence strategy; human brain in a mouse skull; and machine hermeneutics

Is the wall AI is hitting in the room with us right now?

Import AI (Jack Clark)

🎥 영상 & 튜토리얼 (3/11건)

YouTube 2026-09-24

5 Prompts For Every ChatGPT New Feature

Here are 5 new ways to use ChatGPT’s 5 new features 👇 1. Use GPT-6’s upgraded computer use to turn you workspace into a digital diorama in Blender. 2. Use GPT Image 2.5 to redesign one corner of a room while preserving the rest. 3. Use ChatGPT Work’s cloud browser to check your AI subscriptions acro

OpenAI에이전트비전연구가격
Matt Wolfe AI
YouTube 2026-09-24

Claude Opus 5.5 AI: An Incredible Leap Forward

❤️ Check out Lambda here and sign up for their GPU Cloud: https://lambda.ai/papers Note: in the walking creatures experiment, Astra used a simplified model and was unable to implement the correct one. Things did not improve after simulating it for more generations. Claude Opus 5.5: https://www.anthr

ClaudeAnthropic비전연구
Two Minute Papers
YouTube 2026-09-23

The most expensive 33 hours in WordPress history...

Greptile is the only AI code reviewer that actually runs your code. Try it for free: http://greptile.com/go/fireship Automattic's board fired WordPress co-founder Matt Mullenweg while he was at Burning Man... 33 hours later, he replaced the board and declared himself a pirate. Let's dive i

Fireship