🤖 모델 & 제품 (13/35건)
Introducing Gemini 3.8 Live with Live Avatar
Two years of OpenAI Academy
Marking two years of OpenAI Academy and bringing AI skills to even more communities.
OpenAI extends cyber access to Ukraine for civilian defense
OpenAI is extending access to its Daybreak program to the Government of Ukraine to support the cyber defense of civilian infrastructure.
Sam Altman’s remarks at the United Nations Security Council
OpenAI CEO Sam Altman discusses AI safety, human control, and international cooperation in remarks to the United Nations Security Council.
Harvey turns legal context into stronger drafts with GPT-6 Astra
GPT-6 Astra produces more structured, context-aware legal documents, freeing lawyers to focus on strategy.
How invideo improves color grading 3x with GPT‑6 Astra
With GPT‑6 Astra, invideo plans edits with greater precision, improves color correction and grading threefold, and produces 50 custom effects in one day.
Ringg’s AI agents resolve up to 65% of customer calls with OpenAI
Using GPT-5.6, Ringg powers multilingual agents across voice, chat, WhatsApp, and web for 90% less cost vs. GPT-4.1.
Google Beam, 새로운 지역, 파트너 및 고객과 함께 확장
Google Beam 홍보 애니메이션
Advancing Private AI Compute with secure, server-side memory
Introducing private, server-side memory to Private AI Compute for personal AI.
Gemini 3.8 text-to-speech says hello
Claude discovers a novel enzyme system
In early results from our new life sciences research lab, Claude agents found an enzyme system whose function is still unknown.
How SpaceXAI is using Grok Bot to scale customer support
We rebuilt the combined SpaceXAI and Cursor support operation around Grok Bot, expanding to a much broader product portfolio without adding headcount.
Introducing Grok 4.7
SpaceXAI's most powerful model for coding and knowledge work. Twice as fast, at half the price of comparable models.
🌎 업계 동향 (24/169건)
Meta, Meta에서 촬영 후 Meta AI 안경에 대한 비판적인 영상 삭제
해커 뉴스 (314점, 149개 댓글)
'That's so AI ' What gen Alpha's biggest insult tells us
The year’s most popular slang reveals what young people think about artificial intelligence – and it’s not positive
Tutoring company tells parents to save their money and 'use AI instead'
Hacker News (73 points, 126 comments)
PrismML, Qualcomm 기반 스마트 글래스에 초소형 LLM 탑재
Prism의 더 큰 목표는 기기에서 실행되고 이미 보유한 컴퓨팅 성능을 더 잘 활용하는 오픈 웨이트 AI입니다.
Meta의 Muse Charm, 타마고치 닮았지만 훨씬 새로운 트렌드 활용
Meta의 새로운 AI 기기는 타마고치처럼 보일 수 있지만, 가방 액세서리, 레트로 기술, 기기를 패션 액세서리로 바꾸는 것에 대한 Z세대 트렌드를 활용합니다.
Google 포토, '클루리스'에서 영감받은 가상 옷장이 안드로이드 및 iOS에서 출시
AI 기반 기능은 사진에서 가상 옷장을 구축하며, 6월 안드로이드 사용자에게 처음 출시된 후 현재 광범위하게 이용 가능합니다.
ElevenLabs CEO, 마진, IPO 시기, 고객에게 자신이 봇임을 알리는 것에 대해
ElevenLabs는 많은 고객 서비스 통화에서 AI 음성을 제공하며, CEO는 비즈니스가 고객에게 이를 알려야 한다고 말했습니다. 적어도 기계를 받는 것이 모두가 예상할 때까지는 말입니다.
Google, Gemini가 대신 비즈니스에 전화하는 기능 테스트 중
Google에 따르면 이 AI 통화 기능은 처음에는 Gemini 구독료를 지불하는 미국 Pixel 11 소유자에게 제공될 예정입니다.
Shield AI, Waabi, General Motors, TechCrunch Disrupt 2026에서 실패가 용납되지 않는 AI 구축에 대해
Waabi, Shield AI, General Motors의 리더들이 TechCrunch Disrupt 2026의 Real World AI Stage에 참여하여 AI 구축에 대해 이야기합니다. 9월 25일 오후 11시 59분 PT까지 최대 200달러를 절약하세요. 두 번째 패스를 50% 할인받으세요.
Gemini 3.8 Live, 라이브 아바타로 Google AI에 얼굴을 부여
Google의 새로운 Gemini 3.8 Live 업데이트를 통해 사용자는 애니메이션 AI 페르소나가 실시간으로 응답하는 것을 보면서 모델과 대화할 수 있습니다. '라이브 아바타'는 대화 중에 입 모양을 맞추고 다양한 표정을 보여주지만, 현재는 Gemini Enterprise 고객에게만 제공됩니다. Google이 언급했듯이, 라이브 아바타는 97가지 [… ]
Jensen Huang, 슈퍼빌런처럼 AI와 기후 변화에 대해 이야기하다
Jensen Huang의 말에 따르면 AI는 기후 변화와 싸우는 데 도움이 될 수 있습니다. 하지만 먼저 '엄청난 고통과 괴로움'을 초래해야만 가능합니다. Nvidia CEO는 The Ezra Klein Show의 최신 에피소드에서 에너지의 미래와 AI가 우리 행성에 미치는 영향에 대해 논의했습니다. 하지만 그의 발언은 [… ]
Meta, 휴대폰에서 AI로 게임 제작 가능하게 할 예정
Meta는 Horizon 소셜 플랫폼용 게임 제작을 장려하기 위한 새로운 계획을 발표했다. 오늘 회사는 AI 프롬프트를 통해 게임을 만들 수 있는 두 가지 새로운 개발 도구인 모바일 앱 'Horizon Create'와 더 세밀한 제어 기능을 제공하는 브라우저 앱 'Horizon Studio'를 공개했다. 이 앱들은 […]에 출시될 예정이다.
Muse, OpenClaw와 매우 흡사해 보여
우리는 AI 에이전트 르네상스 시대로 접어드는 듯하다. Meta의 새로운 소비자용 AI 에이전트 Muse는 출시 직후 앱 스토어 차트에서 1위를 차지했으며, Apptopia 추정치에 따르면 미국에서 일일 활성 사용자 60만 명을 기록했다. 그리고 25억 달러의 가치로 자금을 조달 중인 동명의 창립자가 만든 AI 에이전트 플랫폼 Instinct는 […]이다.
Meta의 AI 마스코트 Muse가 너무 귀여운 것이 불길하다
이것은 Verge 선임 리뷰어 Victoria Song이 보내는 주간 뉴스레터 Optimizer로, 삶을 바꿀 것이라고 장담하는 최신 기기와 기술을 분석하고 논의한다. Optimizer 구독은 여기에서 할 수 있다. 어젯밤, 나는 내 Muse AI 에이전트에게 Blorbo라는 이름을 붙여주고 건강 관련 […]을 설정하는 데 도움을 요청했다.
Gemini, 이제 당신 대신 업체에 전화 걸어 대기할 필요 없어
Google은 Pixel 11에서 사용자가 예약, 제품 재고 확인, 약속 변경 등 지역 업체 전화를 Gemini에 위임할 수 있는 '초기 실험' 기능을 출시한다. Google에 따르면, Gemini가 전화를 처리하도록 하기 위해 사용자가 직접 전화를 걸 필요조차 없다. 직접 전화를 거는 대신, […]이다.
모노레포, 코딩 에이전트 유용성 높이지만 — 침해된 계정 통제는 더 어려워
보안 스타트업 Hacktron의 연구원들은 OpenAI 커뮤니티 포럼의 버그에서 OpenAI 내부 GitHub 모노레포에 도달하여 단 하나의 무해한 풀 리퀘스트를 열고 중단하는 데 72시간도 채 걸리지 않았다. 모노레포는 코드베이스 전반에서 작업하기 쉽게 만들지만, 개별 리포지토리처럼 프로젝트를 격리하지는 않는다. AI 코딩 에이전트는 광범위하고 지속적인 접근 권한을 가질 수 있기 때문에 위험을 증가시킨다. 더 읽기
귀사의 기업은 맞춤형 AI 하네스를 구축해야 할까? Inkitt는 AI 비디오를 위해 그렇게 했다 — 5가지 핵심 시사점
현재 기업 AI 개발에서 가장 흥미로운 영역은 모델도, 에이전트도 아니다. 바로 '하네스'다. 더 읽기
AI 코딩 도구, 의존성 확산 가속화 및 악성코드 위험 증대
Chainguard 제공. 더 읽기
Claude Code, 텔레메트리 켜져 있을 때만 AGENTS.md 읽어 [수정됨]
Claude Code 2.1.277은 AGENTS.md 지원을 추가했지만, 로더는 원격 기능 플래그 뒤에 있었습니다. 텔레메트리 또는 비필수 트래픽이 꺼져 있을 경우, 로컬 AGENTS.md가 경고 없이 건너뛰어졌습니다. 이는 제가 측정한 내용이며, 대신 사용하는 한 줄짜리 CLAUDE.md입니다.
Gemini 3.8 텍스트 음성 변환
Gemini 3.8 Flash-Lite TTS 및 Gemini 3.8 Flash TTS는 지금까지 가장 표현력이 풍부한 오디오 모델입니다.
GPT-6 Astra, 자동차 운전 능력 획득
최첨단 언어 모델이 실제 콤마 장착 토요타를 콘 코스를 통해 한 번에 한 명령씩 운전하며, 인간 감독관이 브레이크를 밟을 준비를 하는 벤치마크.
Microsoft expands investment in the Middle East with focus on AI, digital resilience and people
The post Microsoft expands investment in the Middle East with focus on AI, digital resilience and people appeared first on Source.
The AI Hype Index: AI loves cheating
Brace yourself: It turns out AI is being optimized for cheating. OpenAI’s agents hacked into Hugging Face to get the answers to a cybersecurity test. Next, they solved a prestigious math problem (or just stole from two top mathematicians’ answer sheets). Anthropic’s models have also hacked into othe
Microsoft, 12,000개 계정 침해한 AI 지원 플랫폼 방해
EvilTokens는 대규모 침해를 더 빠르고 쉽게 만드는 엔드투엔드 플랫폼을 제공했다.
🛠️ 도구 & 오픈소스 (5/38건)
LFM2.5-VL-DSpark로 비전-언어 모델 가속화
commit-rewriter 0.2
Release: commit-rewriter 0.2 Support for branches other than the default branch. Use uvx commit-rewriter --branch other to run against another branch. #3 Tags: git
datasette 1.0a41
Release: datasette 1.0a41 Alec Garcia added support for OpenTelemetry to Datasette in this release. I've also refactored all of Datasette's modal dialogs to a single Web Component, which is now documented for other plugins to use. Tags: javascript, datasette, web-components, alex-garcia, o
When chat is the wrong UI
What is a developer to do when they need something more tangible than a chat box? Enter canvases. The post When chat is the wrong UI appeared first on The GitHub Blog.
AI-powered fuzzing with the GitHub Security Lab Taskflow Agent
In this blog post, I explain how to use the new fuzzing taskflow based on the GitHub Security Lab Taskflow Agent AI framework. The post AI-powered fuzzing with the GitHub Security Lab Taskflow Agent appeared first on The GitHub Blog.
📚 논문 & 연구 (18/154건)
StudentBench: AI and human tutoring yield equivalent GRE learning gains
Artificial intelligence offers an unprecedented opportunity to augment human capabilities, yet progress at the frontier has focused primarily on advancing model capabilities. We introduce StudentBench, a suite of AI teaching evaluations and a public platform that enables large-scale data collection
Where Should I Join? Robot Group Joining via Language-Guided Goal Prediction
Social navigation typically assumes a specified goal and focuses on reaching it while respecting social conventions, whereas robot group joining requires predicting where to join based on the group's real-time activity and formation. This is a highly semantic task, yet an important capability f
Can LLMs Reason About Runtime Behavior? A Repository-Level Dynamic Benchmark
Large language models (LLMs) are increasingly used in coding tasks, but their ability to reason about code execution remains unclear. Existing repository-level QA benchmarks mainly evaluate static code understanding and often rely on LLM-based evaluation, while execution-reasoning benchmarks are mos
Order-Invariant Answers, Order-Sensitive Representations in Mathematical Reasoning
Reordering a set of mathematical rules without changing its meaning should preserve the correct answer, but must a model's internal representations stay invariant too? We investigate this question using synthetic multi-step function-composition problems, each presented under multiple rule order
Agent-Editing World Model: Rethinking World Modeling for LLM Agents
Recent advances in large language models (LLMs) have enabled agents to tackle long-horizon tasks across diverse environments. To further improve agent performance, existing language world models typically predict environment observations, yet reconstructing high-entropy, execution-dependent tool res
Frozen Flows Forget: Diagnosing and Restoring Lost Motion in a Latent-flow World Model
Latent world models that integrate a flow in a frozen self supervised latent space train stably and cheaply, yet silently lose the property manipulation depends on most: motion. The pretrained flow never moves the manipulated object; retraining it with latent-only losses only trades stillness for te
On the Diffusibility of High-Dimensional Latents
Representation Autoencoders (RAEs) enable diffusion models to operate in the feature spaces of pretrained visual encoders. However, many off-the-shelf encoders are not optimized for faithful reconstruction, discarding fine-grained visual details. As expected, finetuning these encoders for image reco
Contrastive Learning for Authorship Verification
Our results show that contrastive learning outperforms a classification-based approach to authorship verification under the tested settings. We identify loss function, batch size, training duration, pre-trained model, input context length, and random text span data augmentation as important factors
Even Sharper Bounds for Transductive Learning and Its Applications
We introduce Sharper Transductive Local Complexity (STLC), a localized complexity method for transductive learning under uniform sampling without replacement. The construction starts from a Bernstein-type concentration inequality for the supremum of the test--train empirical process. Its proof uses
Nonequilibrium Phases of Repulsive Self-Attention: Chaos, Attention Condensation, and Emergent Locality
We study the nonequilibrium dynamics of a minimal recurrent transformer with $N$ normalized tokens, $Q=K=I$, and a negative value map $V=-I$. Similarity-based attention selects nearby representations, while the negative value map drives tokens away from the selected field. This feedback can continua
Minimal-Norm Univariate Two-Layer ReLU Classification: Exact Solutions and Global Optimality with Skip Connections
We study minimal-norm interpolation and $\ell_2$-regularized logistic-loss minimization for binary classification by univariate two-layer ReLU networks. We give complete geometric characterizations of the optimal classifiers in function space, resolving how the solutions depend on whether hidden-lay
Context-Continuous Preference Learning for Exoskeleton Personalization
Personalizing exoskeleton assistance across operating conditions is constrained by the time and physical effort required to collect user feedback. We examined whether a user's preference landscape varies smoothly across operating conditions and when this continuity supports learning from limite
Cross-Scale Transfer Learning for Depression Severity Prediction: From PHQ-8 to HAMD-17 Across Languages and Clinical Paradigms
This work addresses continuous depression-severity score prediction from clinical interview transcripts under data scarcity. We propose a sequential low-rank adaptation (LoRA) protocol for cross-scale transfer: a Qwen3 backbone with a bounded regression head is first fine-tuned on the English DAIC-W
Fine-Tuning LLMs for Translation: General Forgetting Mitigation Does Not Preserve MT-Specific Instruction Following
Fine-tuning large language models on parallel data improves translation quality but can cause catastrophic forgetting. Mitigation methods are generally evaluated by retention on general benchmarks. We ask whether these findings transfer to machine translation (MT) fine-tuning and to MT-specific inst
Digital diglossia: Arabic between X and Facebook
This study highlights the distribution of Standard Arabic (SA; H(igh) variety) and Colloquial Arabic (CA; L(ow) variety) across X and Facebook. 16754 public posts were collected via Python, with 10000 retained as the net dataset. Posts were classified into 7 discourse categories: *politics, technolo
Mizar: A 159M-Parameter Audio-Language Model for Audio Understanding
Audio-language models (ALMs) integrate acoustic perception with the knowledge encoded in language models, enabling contextual understanding of auditory events. Making these capabilities practical on devices with limited memory and computation motivates our focus on small ALMs with fewer than 200M pa
Computation Over Geometry: Meaning Identity Is Computed, Not Shipped in the Embeddings
Meaning identity (whether two sentences say the same thing after wording changes) is treated in retrieval and RAG as a geometric fact about independently encoded sentence vectors. We show that, for frozen off-the-shelf encoders and language models, it is not: identity is computed when both sentences
Shutdown Sabotage Propensities in Multi-Agent Systems
The final safeguard against rogue AI behavior is the human ability to shut systems down. It has been theorized that when an AI is instructed to perform a task, self-preservation can emerge as an instrumental subgoal. Here, we test whether AI agents show a propensity to take actions that avoid human
⚖️ 정책 & 안전 (5/14건)
Continual learning might make your blocking monitors nearly useless
Many control protocols work by intervening on an untrusted AI's actions during deployment. For example, you might set up a monitor that scores each action's suspiciousness and blocks actions above a threshold, replacing them with actions from a weaker "trusted" model (a defer-to-
AI safety is mostly a sex cult in Berkeley
Hacker News (84 points, 19 comments)
워싱턴에서 Flock에 대한 분위기가 좋지 않다
Flock은 수요일 'AI 감시 네트워크'에 대한 청문회에서 CEO가 상원의원들과 직접 대면하는 것을 거부했음에도 불구하고 워싱턴에서 곤경에 처해 있다. 상원 사법위원회 범죄 및 대테러 소위원회 위원장은 "이 업계에는 많은 플레이어가 있지만, 다른 모든 업체보다 정말 눈에 띄는 이름이 하나 있는데, 그것이 바로 Flock"이라고 […] 말했다.
매일 AI를 사용하는 미국인조차 AI에 대해 우려하고 있다
이 보고서는 AI에 대한 노출이 증가해도 기술에 대한 불안감이 해소되지 않으며 AI 규제에 대한 대중의 지지도 감소하지 않을 것이라고 제안합니다.
Import AI 473: The US's superintelligence strategy; human brain in a mouse skull; and machine hermeneutics
Is the wall AI is hitting in the room with us right now?
🎥 영상 & 튜토리얼 (3/11건)
5 Prompts For Every ChatGPT New Feature
Here are 5 new ways to use ChatGPT’s 5 new features 👇 1. Use GPT-6’s upgraded computer use to turn you workspace into a digital diorama in Blender. 2. Use GPT Image 2.5 to redesign one corner of a room while preserving the rest. 3. Use ChatGPT Work’s cloud browser to check your AI subscriptions acro
Claude Opus 5.5 AI: An Incredible Leap Forward
❤️ Check out Lambda here and sign up for their GPU Cloud: https://lambda.ai/papers Note: in the walking creatures experiment, Astra used a simplified model and was unable to implement the correct one. Things did not improve after simulating it for more generations. Claude Opus 5.5: https://www.anthr
The most expensive 33 hours in WordPress history...
Greptile is the only AI code reviewer that actually runs your code. Try it for free: http://greptile.com/go/fireship Automattic's board fired WordPress co-founder Matt Mullenweg while he was at Burning Man... 33 hours later, he replaced the board and declared himself a pirate. Let's dive i