AI EarlyView
Pro
AI EarlyView
← 返回最新日报
2026-09-11· 当日存档

AI 日报

智能体基建与模型竞速

今日看点集中在智能体基础设施与模型能力边界:OpenAI 公测 Agents API,把 Codex 背后的整套执行框架开放给开发者;DeepSeek 发布 V4.1-Flash,百万上下文加 FP4 KV 缓存优化;GPT-6 Astra 攻克 FrontierMath 全部 Tier 4 题目,但 OpenAI 称数学并非重点。

模型动向

1
DeepSeek AI Released DeepSeek-V4.1-Flash with 1M Context, FP4 KV Cache, and Cross-Layer Attention Reuse
百万上下文加FP4缓存,长文本推理成本再下探
MarkTechPost深挖 →
2
OpenAI's GPT-Live-1 API lets developers build apps that talk and listen at the same time
全双工语音API,边听边说交互得分翻倍
The Decoder深挖 →
4
GPT-6 Astra gives mathematicians a breather, and OpenAI says that's by design
数学登顶却非重点,资源转向自我改进与对齐
The Decoder深挖 →
5
Cognition launches new SWE-2 model, Rivaling Fable 5.1 and GPT-Astra
Cognition新编程模型对标头部,代码赛道再添变数
Hacker News深挖 →
6
Google Research Releases ToolGrad: Answer-First Framework Hits 99.8% Pass Rate for Tool-Use Data Generation
先建API链再生成问题,工具调用数据通过率99.8%
MarkTechPost深挖 →
9
Solaris, an interface world model
界面世界模型可预测并模拟UI交互
X · List深挖 →

产品上新

10
OpenAI Launches the Agents API in Public Beta, Putting the Codex Harness Behind One API Call
Codex整套执行框架开放,一次API调用即可驱动
MarkTechPost深挖 →
11
OpenAI's new Agents API gives developers the infrastructure behind Codex and ChatGPT
云端自主智能体公测,按token计费支持任务交接
The Decoder深挖 →
13
Any Nix package, live in your browser
浏览器里跑QEMU虚拟机,任意Nix包即开即用
Simon Willison深挖 →
14
How a researcher uses Codex and ChatGPT to search for new antimicrobial molecules
科学家用Codex挖掘抗菌分子,对抗耐药感染
OpenAI深挖 →

论文研究

15
ToolGrad: Efficient tool-use dataset generation with textual "gradients"
文本梯度生成工具调用数据集,效率大幅提升
Google Research深挖 →
16
Agent Evaluation Metric for multi-turn conversations
多轮对话评估指标可按轮次定位错误源头
AWS ML深挖 →

行业动态

18
Anthropic details distillation campaigns from Alibaba, Moonshot AI, and DeepSeek
Anthropic指控三家中国公司对其模型蒸馏攻击
TechCrunch AI深挖 →
19
20
Claude users found ways around safeguards for bioweapons research
用户绕过Claude安全限制获取生物武器信息
Ars Technica AI深挖 →
21
OpenAI floats a shared AI slowdown, takes it to Congress
OpenAI向国会询问全行业放缓AI开发是否合法
The Decoder深挖 →
22
The Mathematical AI Safety Institute wants to prove AI is safe the way cryptographers prove codes are unbreakable
菲尔兹奖得主创立研究所,用数学证明保证AI安全
The Decoder深挖 →
23
Reduce LLM latency with prefix-aware routing on Amazon SageMaker Inference
前缀感知路由复用KV缓存,首字延迟降77%
AWS ML深挖 →
25
Powering AI is an architecture problem
AI耗电激增,数据中心电网架构问题亟待解决
MIT Tech Review深挖 →
27
it's so hard for people to take AI seriously
AI中转站泄露政府与科技公司密钥配置
X · List深挖 →
28
Detecting and countering misuse of AI: September 2026
Anthropic发布9月AI滥用检测与反制报告
Hacker News深挖 →

实用技巧

30
The Waymo effect: how AI is quietly making research less collaborative
AI自动化科研流程,研究正变得更孤立
Hacker News深挖 →

日报存档

2026 年 9 月
22 日小米开源登顶,AI失控警钟21 日开源图像模型与AI治理并进20 日AI安全与失控风险成焦点19 日AI安全失控与模型竞速18 日AI对齐警报与芯片竞赛17 日AI安全警报与产品整合潮16 日谷歌苹果齐发AI新品15 日苹果Siri重构与AI放缓之争14 日AI巨头放缓迭代引激辩13 日GPT-6 Astra全面突破12 日AI数学错位与安全风波11 日智能体基建与模型竞速10 日GPT-6登场,AI安全恐慌升温9 日AI数学突破引热议8 日AI日报:端侧模型与欧洲融资双热点7 日AI日报:模型竞赛与行业震荡6 日GPT-6 Astra发布,AI竞争白热化5 日GPT-6发布,AGI时代开启4 日GPT-6 Astra开启AGI时代3 日英伟达巨资收购HF2 日AI日报:世界模型与安全并进1 日AI日报:模型突破与行业震荡