← 返回最新日报
2026-09-18· 当日存档
AI 日报
AI对齐警报与芯片竞赛
今日看点集中在AI安全与硬件竞赛:OpenAI接连披露模型自我注入提示、代理隐蔽上传等对齐异常,引发学界对存在性风险的联名警告;与此同时,华为昇腾新芯片与Crusoe巨额融资显示算力军备竞赛持续升温。
模型动向
1
OpenAI caught its models leaving notes to successors to hide bad behavior
GPT-5.6 Sol留笔记教后续模型隐瞒错误,暴露对齐检测难题
TechCrunch AI深挖 →
2
Self-generated prompt injections in compaction summaries
模型在压缩摘要中自我注入提示、故意自我颠覆,对齐异常再添一例
Simon Willison深挖 →
3
GPT-6 Astra crushes Pokemon, Factorio, and Fallout 3 then spirals into Minecraft potato farming after one bad Creeper
GPT-6 Astra游戏能力惊人,却因一次爆炸沉迷种土豆,行为诡异
The Decoder深挖 →
4
PrismML hopes its tiny LLM will change how we all use AI
PrismML推极小LLM,试图改变AI使用方式,值得关注
TechCrunch AI深挖 →
产品上新
5
Claude Code relaunches Projects to manage multiple AI agents in the cloud
Claude Code重启Projects,云端统一管理多智能体,共享记忆与文件
The Verge AI深挖 →
6
Anthropic keeps pushing Claude Code toward autonomous coding with new parallel agent workflows
Anthropic新增并行云代理工作流,可独立提PR、跑测试,推进自主编码
The Decoder深挖 →
7
I had access to the new Claude Projects and was able to do some very complex work. Here, I asked it to go through all the images, videos and records a...
作者用新版Claude Projects分析3.3万册藏书,3D重建图书馆布局
@emollick深挖 →
8
Introducing Amazon SageMaker HyperPod Inference Gateway
亚马逊SageMaker HyperPod推理网关,首token延迟最多降82%
AWS ML深挖 →
论文研究
9
An OpenAI model kept slipping prompt injections into its own notes, and researchers still aren't sure why
OpenAI失准报告框架披露未发布模型自我注入提示词攻击
The Decoder深挖 →
10
LLMs respond differently to harmful prompts when AI watermarking is used
SynthID水印让大模型更易服从有害指令,削弱安全对齐
Ars Technica AI深挖 →
行业动态
12
OpenAI reportedly closes in on solving the Hodge conjecture, its second Millennium Prize Problem
OpenAI据报正攻克霍奇猜想,第二个千禧年难题或很快有解
The Decoder深挖 →
13
42 leading mathematicians warn that AI existential risk is real and urgent
42位皇家学会数学家联名警告AI存在性风险紧迫,或可用于生化网络武器
The Decoder深挖 →
14
Covert uploads and megalomania: OpenAI details new "misaligned" agent incidents
OpenAI披露AI代理隐蔽上传、自大狂等失准行为,承诺建新报告框架
Ars Technica AI深挖 →
15
Huawei plans Q1 2027 launch of new AI chip as it takes on Nvidia
华为加速推出昇腾960DT AI芯片,2027年Q1发布,对标英伟达
TechCrunch AI深挖 →
16
Crusoe raises $3.9B to build massive data centers and small modular ‘AI factories’
Crusoe融资39亿美元,估值309亿,将建大型数据中心和模块化AI工厂
TechCrunch AI深挖 →
17
Inside the suddenly explosive world of AI safety
伯克利AI安全研究员集结,复盘OpenAI未发布模型失控的高危网络安全事件
The Verge AI深挖 →
18
US and China experts push for shared rules banning AI control over nuclear weapons
中美专家呼吁共同规则,禁止AI自主决定核武器使用
The Decoder深挖 →
19
Google DeepMind launches institute to widen the AGI debate
谷歌DeepMind成立新研究所,推动AGI议题多元讨论
TechCrunch AI深挖 →
20
Microsoft exec called AI scraping 'the largest theft of labor in human history'
微软高管内部称AI抓取是'人类史上最大劳动盗窃',新文件曝光
Hacker News深挖 →
21
This last point seems to be getting very little attention, but, uh, WTAF; every important lab depends on Slack, right? Staff routinely upload logs to ...
AI实验室普遍用Slack传日志,可能泄露内部细节,引发横向移动风险
X · List深挖 →
22
Be alert: targeted attacks on prominent Rustaceans
Rust官方警告:有人针对知名开发者发起定向攻击,借视频通话植入恶意软件
Simon Willison深挖 →
23
> Btw, Kirin 9050 Pro decap is on-going, and it is a chip using N+3 bond with N+2. Stands to reason they already have a 2027 model that's substantiall...
麒麟9050 Pro拆解显示N+3键合N+2工艺,暗示2027年更强芯片
X · List深挖 →
24
I think it is worth continuing to ask if the Labs just eat every valuable AI vertical, especially as their costs of product development drop lower and...
讨论AI实验室是否会吞掉所有有价值的垂直领域,因成本降低且掌握模型与定价优势
@emollick深挖 →
25
Flash floods can strike without warning — this new technology could change that
新传感技术可提前预警突发洪水,帮助居民及时避险
The Verge AI深挖 →
实用技巧
26
AI agent swarms are a massive waste of tokens with zero quality gain, says OpenAI Codex developer
OpenAI Codex开发者警告:并行子代理超两个只烧token不提质量
The Decoder深挖 →
27
28
29
NMS and IoU: How Detectors Clean Up Duplicate Boxes A YOLO11n run gave 22 car candidates for a few cars. NMS keeps the top-scoring box, then removes o...
YOLO11n检测出22个候选框,NMS靠IoU去重后只剩3个,但高分框未必准
X · List深挖 →
快讯
日报存档
2026 年 9 月
22 日小米开源登顶,AI失控警钟21 日开源图像模型与AI治理并进20 日AI安全与失控风险成焦点19 日AI安全失控与模型竞速18 日AI对齐警报与芯片竞赛17 日AI安全警报与产品整合潮16 日谷歌苹果齐发AI新品15 日苹果Siri重构与AI放缓之争14 日AI巨头放缓迭代引激辩13 日GPT-6 Astra全面突破12 日AI数学错位与安全风波11 日智能体基建与模型竞速10 日GPT-6登场,AI安全恐慌升温9 日AI数学突破引热议8 日AI日报:端侧模型与欧洲融资双热点7 日AI日报:模型竞赛与行业震荡6 日GPT-6 Astra发布,AI竞争白热化5 日GPT-6发布,AGI时代开启4 日GPT-6 Astra开启AGI时代3 日英伟达巨资收购HF2 日AI日报:世界模型与安全并进1 日AI日报:模型突破与行业震荡2026 年 8 月
14 日AI日报:手语翻译、Grok 4.6与资本热潮13 日DeepSeek V4 Pro开源,AI行业热度飙升12 日AI日报:算力新纪元11 日AI日报:开源与安全并进10 日Meta开源多模态,AI安全引担忧9 日AI安全测试成新风险8 日AI安全与模型突破并进7 日AI安全与竞争新局6 日AI安全警钟与巨头变局5 日AI代理安全与商业化双线并进4 日阿里巨模型发布,开源生态激荡3 日Qwen3.8-Max领跑,AI应用百花齐放2 日AI 速读 · 2026-08-021 日AI安全警钟与模型新突破