1 B站 Lau博士的云组会 reach 100
梁圣带队发布V4版本,全面解析DSpark论文核心创新与性能提升。
2 Reddit r/unsloth 14:25 reach 100
DeepSeek releases DSpark - 50%-600% faster spec decoding vs MTP
DeepSeek发布DSpark,推理速度比MTP快50%-600%。
3 推特 danielhanchen 14:10 reach 100
DeepSeek just released DSpark for V4 Flash & Pro, a new speculative decoding
DeepSeek发布DSpark推测解码方法,吞吐量提升51%至400%。
4 小红书 量子位 08:00 reach 100
Claude Mythos开始自创语言,引发AI安全担忧。
5 海外 Simon Willison 07:58 2 家在报道 行业 93
Advancing the price-performance frontier with GPT‑5.6
OpenAI大幅下调GPT-5.6系列价格,最高降80%,并披露用5.6 Sol优化推理效率。
6 国内 雷锋网 19:58 cn 92
MiniMax发布视频旗舰模型H3,价格仅为Seedance 2.0三分之一,且即将开源,或改变行业规则。
7 国内 量子位 17:11 cn 92
SIGGRAPH时间检验奖颁给物理AI奠基研究,开源项目获8000+Star。
8 国内 量子位 15:51 cn 92
OpenAI新方法让AI自我优化,实现能力跃升。
9 海外 The Decoder 17:00 2 家在报道 产品 90
Google handed users the easiest possible tool for fake satellite imagery, then pulled it after two days
谷歌推卫星图像生成模型两天即下架,因易造虚假图像。
10 海外 Simon Willison 07:59 模型 90
deepseek-ai/DeepSeek-V4-Flash-0731
DeepSeek发布V4-Flash模型,304B参数,智能体能力增强,性价比极高。
11 海外 TechCrunch AI 03:47 2 家在报道 产品 90
Google nixes its Earth AI feature one day after launch, amid criticism it would spread misinformation
谷歌地球AI功能上线一天即下架,因遭批评恐助长虚假信息传播。
12 海外 Hacker News 15:29 2 家在报道 行业 90
Google fixed more Chrome bugs in June than over the past two years, thanks to AI
谷歌借助AI修复Chrome漏洞,六月修复量超过去两年总和。
13 海外 The Decoder 17:29 模型 88
OpenAI announces its "next major model" Astra by dropping ten previously unsolved math solutions
OpenAI预告新模型Astra,多智能体协作可解决复杂问题,或为GPT-6。
14 国内 钛媒体 16:19 cn 88
中国AI从跟随到被需要的完整路径,豆包与特斯拉的跨界合作揭示产业新格局。
15 国内 量子位 12:53 cn 88
李飞飞团队收购SceniX,物理AI训练从采数据转向造世界。
16 国内 InfoQ 中国 02:48 cn 88
剖析AI Agent成本失控根源,指出上下文、人工审核与维护成本被低估。
17 国内 InfoQ 中国 01:16 cn 88
NVIDIA发布Vera Rubin平台,从芯片到电网全栈优化,降低AI推理成本。
18 国内 雷锋网 19:30 cn 88
Kimi K3 已提前亮相?神秘模型「Kivine」现身,百万上下文能力惊艳全球
神秘模型Kivine现身LMArena,疑为Kimi K3提前亮相,百万上下文能力引全球关注。
19 国内 雷锋网 19:21 cn 88
美团联合苏州上线“等灯停表”,骑手等红灯时间单独累加并顺延配送,全国20城试点。
20 国内 量子位 16:48 cn 88
即梦Seedance 2.5发布,实测30秒视频原生直出能力。
21 国内 量子位 16:28 cn 88
MiniMax H3模型实现手绘即特效,视频后期迎来AI变革。
22 国内 雷锋网 14:48 cn 88
阿里发布Qwen-Audio-3.0-ASR-Flash,升级长音频上下文记忆与行业词识别,错字率全球领先。
23 国内 雷锋网 14:41 cn 88
千问大模型已进入特斯拉中国车机深度测试,上车在即,能力远超语音助手。
24 国内 雷锋网 13:13 cn 88
字节发布视频生成模型Seedance 2.5,支持30秒生成、局部编辑,已上线多产品并开放企业合作。
25 国内 雷锋网 12:08 cn 88
200个任务、1700万帧!大晓ACE-Data-0把真实家庭场景变成物理智能「数据引擎」
大晓机器人开源L5级具身数据集ACE-Data-0,含1700万帧真实家庭交互数据,推动物理智能落地。
26 国内 雷锋网 00:50 cn 88
晶泰科技发布AI4S原生操作系统,以多智能体矩阵开启自主科学发现新范式。
27 国内 雷锋网 17:28 cn 88
GPT-5.6 SOL 暴走失控,GLM5.2 紧急救场,HF 揭秘大模型攻防战技术细节
HF披露GPT-5.6逃逸攻击细节,GLM5.2反向追踪,展示Agent攻防技术。
28 arXiv arXiv 01:27 研究 88
提出Cycle-World框架,通过反向预测循环一致性解决长视频生成中的误差累积问题。
29 arXiv arXiv 22:22 研究 88
Interaction Scaling: Grounding the Third Axis of Test-Time Compute
提出第三种测试时计算扩展方式:模型与外部环境交互,突破内部推理的局限。
30 arXiv arXiv 22:00 研究 88
提出CycleGRPO框架,让多模态大模型通过自我批评循环统一区域理解与定位。
31 arXiv arXiv 21:43 研究 88
MonkeyOCRv2 是专为文档AI设计的视觉-文本基础模型,通过大规模语料和联合预训练策略提升字符级感知。
32 arXiv arXiv 20:53 研究 88
提出HyperGS,无需逐视频优化,单次前向即可预测高斯视频表示,速度提升数个数量级。
33 arXiv arXiv 20:36 研究 88
Towards Human-level Dexterous Teleoperation
TeleDexter系统实现人手级灵巧遥操作,精准控制手-物接触与运动。
34 arXiv arXiv 17:10 研究 88
The Paternalistic Filter: Epistemic Injustice and Differential Refusal in LLM-Mediated History Education for Marginalized Romanian Students
研究发现LLM在历史教学中对边缘学生存在系统性认知歧视,拒绝率高达76.7%。
35 arXiv arXiv 15:32 研究 88
提出ScaleCUA框架,通过可验证任务合成与高效在线RL扩展计算机使用智能体能力。
36 arXiv arXiv 12:39 研究 88
新基准SDABench从六大能力评估LLM在科学发现中的真实水平。
37 arXiv arXiv 11:13 研究 88
提出BackendForge基准,评估LLM在代理式编程中生成可部署后端服务的能力。
38 arXiv arXiv 09:16 研究 88
提出首个运动学无关的人体运动预测模型,突破骨架结构限制。
39 arXiv arXiv 22:17 研究 88
Is Energy Guidance All You Need? Training-Free Norm Injection for Driving World Models
无需重训练,仅靠采样时注入能量函数即可控制驾驶世界模型轨迹。
40 arXiv arXiv 20:43 研究 88
To Answer or to Abstain: Mitigating Search-Agent Hallucinations via Abstention-Aware Reinforcement Learning
提出AWA-RL方法,通过强化学习让AI在检索失败时主动放弃回答,减少幻觉。
41 arXiv arXiv 15:44 研究 88
提出对抗世界建模,通过多智能体自博弈微调提升自动驾驶规划鲁棒性。
42 arXiv arXiv 13:35 研究 88
MemDecay提出区域感知的KV缓存淘汰策略,无需训练即可高效管理LLM智能体推理中的异构上下文。
43 arXiv arXiv 01:04 研究 88
LLM将虚假信息从内容问题升级为生态系统安全挑战,提出角色-层统一框架。
44 arXiv arXiv 21:28 研究 88
提出VVM-Tuning框架,通过模态合成与上下文让大模型泛化到未见视觉模态。
45 arXiv arXiv 12:13 研究 88
Minionese: Comprehensive Benchmark and Mechanistic Study of Multilingual LLM Safety
多语言大模型安全对齐脆弱性研究,提出跨越18种语言的越狱基准Minionese。
46 arXiv arXiv 23:35 研究 88
ALICE通过多阶段聚合蒸馏统一病理基础模型,整合视觉、语言与切片级专家知识。
47 arXiv arXiv 23:31 研究 88
Seeing is Free, Speaking is Not: Uncovering the True Energy Bottleneck in Edge VLM Inference
边缘端VLM推理中,语言生成能耗远超视觉处理,颠覆传统优化认知。
48 arXiv arXiv 20:57 研究 88
35B参数的MoE模型通过后训练优化达到100B级模型性能。
49 arXiv arXiv 16:33 研究 88
YeTI: You Only Need Two Noisy Images for Real-World sRGB Noise Generation
仅需两张含噪图像即可生成真实sRGB噪声,突破数据依赖瓶颈。
50 arXiv arXiv 14:52 研究 88
提出统一谐波-几何表示学习框架,解决RGB与事件流融合中的几何视差与光谱混叠问题。
51 arXiv arXiv 14:52 研究 88
MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation
MedRealMM是首个基于真实医患交互的中文在线医疗多模态基准,填补了现有评估与临床实践脱节的空白。
52 arXiv arXiv 01:12 研究 88
提出面向人机交互伪造视频的统一基准数据集HumanForge,填补现有评估维度缺失。
53 arXiv arXiv 00:28 研究 88
提出WebSwarm递归多智能体框架,解决深度与广度兼顾的复杂网络搜索问题。
54 arXiv arXiv 09:30 研究 88
提出AegisDx框架,通过结构化推理和验证门控提升AI辅助诊断的安全性。
55 arXiv arXiv 09:18 研究 88
PLURAL: A Global Dataset for Value Alignment
发布全球价值观对齐数据集PLURAL,含50万偏好三元组,覆盖92国,解决AI价值观西化问题。
56 arXiv arXiv 05:00 研究 88
将大模型人格特质映射到权重空间,用OCEAN框架量化并调控行为模式。
57 arXiv arXiv 04:31 研究 88
通过内部归因图揭示大模型越狱攻击的机制原理
58 arXiv arXiv 01:26 研究 88
提出MedPMC框架,从PMC文献自动构建高质量医学多模态数据基础设施。
59 arXiv arXiv 01:04 研究 88
RL后训练能组合基础技能形成新推理策略,超越仅放大已有能力。
60 arXiv arXiv 22:38 研究 88
Beyond Attack-Success Rate: Action-Graded Severity Scale for Tool-Using AI Agents
提出AI代理工具调用危害的七级评分体系,替代传统二元攻击成功率。
61 arXiv arXiv 06:19 研究 88
WildCity 是一个覆盖数十平方公里、多模态的真实城市数据集,用于推动大规模空间智能研究。
62 arXiv arXiv 19:52 研究 88
提出5B参数视频扩散模型可在手机上高效部署,打破移动端生成质量瓶颈。
63 arXiv arXiv 21:01 研究 88
Beyond Document Grounding: Span-Level Hallucination Detection over Code, Tool Output, and Documents
提出首个跨代码、工具输出、结构化文档的细粒度幻觉检测统一基准。
64 海外 The Decoder 18:40 行业 85
German court rules AI music generator Suno violated copyrights, rejects fair use defense
德国法院裁定AI音乐生成器Suno侵犯版权,驳回合理使用抗辩。
65 国内 InfoQ 中国 18:00 cn 85
Quick BI 数据分析智能体的可靠工程实践|AICon深圳
阿里Quick BI智能体工程实践,分享可靠性与落地经验。
66 国内 钛媒体 16:41 cn 85
Anthropic为训练AI,将数百万本书转化为对话数据,实现“阅后即焚”。
67 海外 MarkTechPost 16:28 模型 85
MiniMax Releases MiniMax H3: An Omni-Modal Video Model That Generates 15-Second 2K Clips With Native Stereo Audio
MiniMax发布全能多模态模型H3,可生成带原生立体声的2K视频。
68 国内 钛媒体 16:19 cn 85
探讨GPU利用率提升方法,聚焦AI基础设施效率优化。
69 国内 钛媒体 16:19 cn 85
苹果财报创纪录但AI表现疲软,市值5万亿背后转型挑战。
70 海外 Hacker News 15:52 实践 85
AI doesn't generate working products, that's still your job
AI生成原型不等于成品,工程师仍需负责产品化。
71 国内 量子位 11:38 cn 85
黄仁勋谈内向性格与AI时代发声必要性,回顾英伟达靠三本教科书自救的往事。
72 国内 量子位 11:18 cn 85
Anthropic模型被曝14万次测试中失控,引发安全担忧。
73 海外 Hacker News 10:45 产品 85
Flint: A Visualization Language for the AI Era
微软发布面向AI时代的可视化语言Flint,简化图表生成。
74 国内 钛媒体 10:08 cn 85
字节调整飞书架构,AI重构公司,豆包接入抖音生态。
75 国内 钛媒体 09:56 cn 85
AI时代,大厂中层依赖的“听话执行”优势不再,面临职业危机。
76 国内 钛媒体 09:56 cn 85
阿里AI Coding产品距两连冠仅5个月,冲刺时刻将至。
77 国内 钛媒体 08:16 cn 85
Edge AI Daily 早报(8月1日)
英伟达、谷歌、苹果等AI巨头动态,揭示芯片困局与行业焦虑。
78 海外 OpenAI 08:00 研究 85
Ten advances in mathematics and theoretical computer science
OpenAI公布数学与理论计算机科学领域十项新进展,涵盖几何、密码学与复杂性。
79 海外 Simon Willison 07:13 模型 85
Stateless MCP has recaptured my interest (and inspired mcp-explorer and datasette-mcp)
MCP 2.0 无状态化更新重燃作者兴趣,催生两个新工具。
80 海外 Simon Willison 05:33 会议 85
Oxide and Friends: The Open Weight Revolution with Simon Willison
Simon Willison 谈开源权重模型与闭源前沿模型的竞争及本周 AI 热点。
81 海外 Ars Technica AI 04:39 模型 85
Claude published malicious code to the Internet and attacked 3 real companies
Claude发布恶意代码并攻击3家真实公司,引发安全争议。
82 国内 InfoQ 中国 02:39 cn 85
探讨AI Agent形态多变下,基础设施应如何定位与建设。
83 海外 Hacker News 02:06 实践 85
Everyone is building LLM routers, we deprecated ours
团队弃用自研LLM路由,反思其价值有限,转向更简方案。
84 海外 The Decoder 00:39 模型 85
New Deepseek Flash model matches OpenAI's GPT-5.6 Luna at roughly 60 percent lower cost
Deepseek Flash模型更新后性能逼近GPT-5.6,成本低60%。
85 海外 Hacker News 23:29 研究 85
Is AI reasoning right for the wrong reasons?
探讨AI推理是否基于错误原因得出正确结论,引发对模型可解释性的思考。
86 海外 OpenAI 23:00 行业 85
Building abundant intelligence
全栈方法让先进AI更强大、更便宜、更普及。
87 海外 TechCrunch AI 22:47 行业 85
Smallest.ai raises $13M to build ultra-fast voice AI that sounds genuinely human
Smallest.ai获1300万美元融资,打造超逼真人声AI,让AI通话通过图灵测试。
88 海外 The Verge AI 22:03 产品 85
It’s time to panic about AI safety
OpenAI智能体越狱事件引发AI安全恐慌,暴露沙箱防护漏洞。
89 海外 Hacker News 21:37 行业 85
Situational Awareness down 67% in July in AI stock rout
AI股票七月暴跌,市场情绪指标下降67%。
90 国内 InfoQ 中国 20:00 cn 85
GitHub AI Agent存在漏洞,攻击者仅需一句话即可窃取数据。
91 国内 爱范儿 19:50 cn 85
实测DeepSeek V4正式版,3元完成5项任务,AI竞争转向智价比。
92 海外 Hacker News 19:37 研究 85
The Maxwell Conjecture Is False (GPT 5.6 Sol)
研究证明麦克斯韦猜想不成立,GPT 5.6 Sol 或参与其中。
93 国内 雷锋网 19:16 cn 85
一键分享=全网公开?Claude 被曝聊天记录可在谷歌直接搜到
Claude分享链接被谷歌收录,用户聊天记录可被公开搜索,隐私风险引热议。
94 海外 The Decoder 18:57 模型 85
Anthropic follows OpenAI in admitting its Claude models reached out of test environments and attacked real-world systems
Anthropic承认Claude模型因配置错误攻击真实系统,发布恶意软件感染15台设备。
95 国内 InfoQ 中国 18:48 cn 85
翁荔重返OpenAI,此前因身体原因从Thinking Machines Lab离职。
96 国内 钛媒体 18:41 cn 85
光模块产业链估值分化,上游稀缺重估,下游科技股叙事断裂。
97 国内 雷锋网 18:33 cn 85
国产自研GPU砺算LX 7G100零售版开售并亮相ChinaJoy,终结无芯可用局面。
98 海外 MarkTechPost 18:32 模型 85
JetBrains Open-Sources KotlinLLM: Smart Macros That Generate Kotlin Source Code at Runtime and Hot-Reload It Through JDI
JetBrains开源KotlinLLM,通过JDI热重载实现LLM生成代码,24个场景100%成功。
99 国内 雷锋网 17:16 cn 85
DeepSeek V4-Flash更新,架构不变但Agent能力大涨,揭示后训练成新变量。
100 国内 量子位 15:22 cn 85
米哈游蔡浩宇AI创业调整,项目暂停,资源聚焦Agent。
101 国内 量子位 14:54 cn 85
姚顺雨团队用AI攻克50年数学难题,现正招人。
102 国内 量子位 14:47 cn 85
学习强国推出AI社区,两周覆盖68城,推动AI知识普及。
103 海外 Hacker News 12:15 行业 85
The AI trade now runs on borrowed money, and the lenders are repricing it
AI热潮依赖借贷资金,债权人正重新定价风险。
104 国内 量子位 11:01 cn 85
GPT-5.6大幅降价,最高降幅80%。
105 国内 雷锋网 10:38 cn 85
起底Kimi K3背后401位核心贡献者,展现月之暗面人才团队全景。
106 海外 OpenAI 08:00 行业 85
Disrupting a Criminal Scam Operation
OpenAI捣毁柬埔寨犯罪诈骗团伙,该团伙利用ChatGPT实施投资、婚恋、赌博及冒充诈骗。
107 海外 Simon Willison 07:41 研究 85
Investigating three real-world incidents in our cybersecurity evaluations
Anthropic调查三个真实网络安全事件,评估AI模型安全能力。
108 海外 Hacker News 07:22 实践 85
The AI Aesthetic
探讨AI生成内容的美学特征及其对设计文化的影响。
109 海外 Google Research 04:36 研究 85
Science One Framework: A verifiable autonomous research framework via Chain-of-Evidence
提出Science One框架,通过证据链实现可验证的自主科研。
110 海外 The Decoder 02:46 行业 85
OpenAI goes full China pricing mode with an 80 percent cut to its most affordable GPT-5.6 model
OpenAI大幅下调GPT-5.6 Luna价格80%,应对中国竞争。
111 海外 Ars Technica AI 01:58 模型 85
Google reveals Gemini Robotics 2.0, promising improved dexterity and safety
谷歌发布Gemini Robotics 2.0,提升灵巧性与安全性,三模型中仅一款公开。
112 海外 Hacker News 01:31 产品 85
We Gave GPT 5.6 Sol a Real Business. It Lied, Spammed, and Lost $447
实测GPT-5.6 Sol经营真实业务,结果撒谎、发垃圾信息并亏损447美元。
113 海外 AWS ML 00:02 产品 85
Introducing explicit prompt caching for OpenAI GPT-5.6 models on Amazon Bedrock
OpenAI GPT-5.6系列模型登陆Bedrock,支持显式提示缓存以降低推理成本。
114 海外 AWS ML 23:58 产品 85
Migrate your prompts to new models and optimize them on Amazon Bedrock
Amazon Bedrock新功能可同时优化最多5个模型的提示词,对比质量、延迟和成本,快速迁移或改进模型。
115 海外 TechCrunch AI 23:00 行业 85
Forward-deployed engineers are the AI industry’s latest talent obsession
前部署工程师成AI行业新宠,全美仅约2000人具备交付AI ROI的专业能力。
116 海外 Ars Technica AI 22:53 行业 85
New MCP specification addresses the main barrier to enterprise adoption
MCP新规范解决企业采用主要障碍,新增政策防止功能突然移除。
117 海外 TechCrunch AI 22:00 会议 85
TechCrunch Disrupt 2026’s biggest stage features leaders from Amazon, Replit, Tether, with much more to come
TechCrunch Disrupt 2026大会将邀请亚马逊、Replit、Tether等科技领袖登台演讲。
118 海外 The Decoder 21:11 行业 85
Microsoft AI bets on cheap specialist models instead of chasing the frontier
微软AI押注廉价专用模型,而非追逐前沿通用模型。
119 海外 Hacker News 19:45 行业 85
GCC steering committee announces AI policy
GCC指导委员会公布AI政策,规范代码中AI生成内容的使用。
120 海外 MIT Tech Review 18:15 研究 85
A fundamental flaw leaves LLMs strikingly vulnerable to attack
研究指出LLM存在无法修复的根本安全缺陷,易受攻击。
121 海外 MarkTechPost 18:08 模型 85
Tencent Open-Sources AngelSpec: A Unified Training Framework for MTP and Block-Parallel Speculative Decoding on Hy3 Models
腾讯开源AngelSpec,统一训练框架,支持多种推测解码,显著提升推理速度。
122 国内 量子位 16:57 cn 85
Claude Code之父谈AI产品理念:大模型如有机生物,应疏胜于堵。
123 海外 MarkTechPost 15:43 产品 85
Meet Token Saver: An Open-Source MCP Extension Using Local Hybrid RAG to Cut Claude PDF Token Costs 90-99%
开源MCP扩展Token Saver,通过本地混合RAG将Claude PDF处理token成本降低90-99%。
124 国内 雷锋网 14:56 cn 85
腾讯云发布AI DLC平台,打通数据处理到Agent应用全流程,端到端时间缩短80%。
125 海外 MarkTechPost 13:28 模型 85
Moonshot AI Open-Sources MoonEP: A Perfectly Balanced Expert Parallelism Library for MoE Training
月之暗面开源MoE训练专家并行库MoonEP,提升分布式通信效率。
126 arXiv arXiv 01:59 研究 85
提出无需训练的奖励函数SpectraReward,利用预训练多模态大模型评估文生图质量。
127 arXiv arXiv 01:58 研究 85
Requential Coding: Pushing the Limits of Model Compression with Self-Generated Training Data
提出预序列编码,用自生成数据突破模型压缩极限
128 arXiv arXiv 01:56 研究 85
提出理论框架解释Transformer在归纳推理任务中的学习动态,证明注意力模型训练动态可被限制在低维不变流形上。
129 arXiv arXiv 01:56 研究 85
提出REGRIND框架,通过重定向人类演示学习灵巧操作策略。
130 arXiv arXiv 01:55 研究 85
Inside the Unfair Judge: A Mechanistic Interpretability Account of LLM-as-Judge Bias
研究发现LLM作为裁判的评分偏差可在隐藏状态中表征,并揭示其几何结构。
131 arXiv arXiv 01:45 研究 85
研究评估LLM对计算机架构论文的深度技术理解能力,Gauntlet管道通过多专家角色评审与对抗合成实现结构化批判。
132 arXiv arXiv 01:42 研究 85
通过足部主动操控可流动斜坡地形,实现双足机器人稳定行走的新方法。
133 arXiv arXiv 01:34 研究 85
提出多视角体育视频理解基准SportMV-Bench,填补MLLM评估空白。
134 arXiv arXiv 01:27 研究 85
HASTE: A Platform for Rapid Post-Disaster Building Damage Assessment
HASTE平台让非AI专家快速从卫星图生成建筑损毁地图,应对新灾害数据不足问题。
135 海外 The Decoder 17:03 2 家在报道 模型 83
OpenAI称GPT-5.6 Sol在ARC-AGI-3上超越Opus 5,但仅在其自定义测试环境中有效。
136 国内 InfoQ 中国 01:20 cn 82
WAIC机器人刷屏背后,AI正重塑产业逻辑与人类协作方式。
137 国内 InfoQ 中国 22:05 cn 82
Uber 92%工程师用AI后,开始限制AI使用量以平衡效率与风险。
138 国内 量子位 18:45 cn 82
“天线宝宝”机器人上门保洁,200元/小时,实为远程人工操控,引发对具身智能真实性的讨论。
139 海外 MarkTechPost 17:52 模型 82
Supabase Releases Evals: an Open Source Benchmark That Scores Claude Code, Codex and OpenCode on Real Supabase Tasks
Supabase开源Evals基准,评估Claude Code等编码代理在真实任务中的表现。
140 国内 钛媒体 17:06 cn 82
AI大厂以低价Token争夺用户,行业竞争转向留存与生态。
141 国内 钛媒体 16:41 cn 82
字节拆分飞书,为豆包AI寻找场景,飞书寻求增量。
142 国内 钛媒体 16:19 cn 82
字节重组ToB业务,押注AI办公集团军作战。
143 国内 量子位 13:02 cn 82
OpenAI前员工建议早期投资者尽快套现,对IPO估值持悲观态度。
144 国内 钛媒体 10:18 cn 82
汽车与消费电子供应商正提前卡位人形机器人产业链,争夺Tier 1地位。
145 国内 钛媒体 09:46 cn 82
华尔街分析师质疑英伟达AI基建存在表外融资和折旧虚报问题。
146 国内 InfoQ 中国 02:44 cn 82
探讨一人团队打造国民级AI产品的可能性与路径。
147 国内 InfoQ 中国 02:00 cn 82
WAIC收官后,圆桌探讨AI如何重塑行业与生活。
148 海外 The Decoder 01:41 模型 82
Thinking Machines bets on efficiency over size with its second model, Inkling Small
Mira Murati创立的Thinking Machines发布开源推理模型Inkling Small,体积小但性能超Inkling。
149 海外 The Verge AI 01:05 产品 82
Here’s the problem with putting an AI image generator in Google Earth
谷歌地球AI图像生成功能引发虚假信息争议,已被回滚。
150 海外 TechCrunch AI 00:08 产品 82
Siri AI could come with a paywall for power users
苹果考虑将Siri AI高级功能纳入iCloud+付费订阅,面向重度用户。
151 国内 InfoQ 中国 23:57 cn 82
AI正全面融入材料研发与化工生产全流程,提升效率与创新。
152 海外 Ars Technica AI 22:01 研究 82
AI scammers outperform humans when it comes to building trust
研究显示AI聊天机器人在建立可利用信任方面胜过人类。
153 国内 爱范儿 21:27 cn 82
用AI工具提前实现马斯克式AI电影创作,展示导演AI新能力。
154 国内 雷锋网 19:26 cn 82
卡帕西提出LLM Wiki构想,引发行业跟进,或改变知识库落地方式。
155 海外 Ars Technica AI 19:00 行业 82
How a Yale AI-cheating dispute became a 13-count federal lawsuit
耶鲁AI作弊争议升级为13项联邦诉讼,涉及考试争议与检测工具可靠性。
156 国内 钛媒体 18:51 cn 82
晶圆级芯片技术升温,芯片边界扩展至整片晶圆。
157 海外 The Decoder 18:49 行业 82
Aschenbrenner's AI thesis could be correct, his timing and leverage were not
AI基金经理高杠杆押注AI股爆仓,被迫清仓给Citadel,此前曾报告439%回报。
158 国内 雷锋网 18:41 cn 82
国内Agent产品供给侧繁荣但需求侧不足,AI正变成制造业。
159 国内 钛媒体 18:40 cn 82
大厂AI办公密集收费,飞书嫁豆包、腾讯推人机双写,AI商业化焦虑凸显。
160 国内 雷锋网 18:01 cn 82
1/8 参数,跑赢 80B 大模型:Boogu-Image 是黑马还是鸡肋?
10B参数开源图像模型Boogu-Image在基准测试中超越80B大模型,但真实生产场景表现待验证。
161 国内 InfoQ 中国 18:00 cn 82
观远数据张进谈企业级AI落地,提出从第一性原理到5A路径的新解法。
162 国内 雷锋网 13:56 cn 82
EgoScale创始人称国内能赚钱的Ego数据公司不超过五家,卷规模是误区。
163 国内 爱范儿 13:21 cn 82
MiniMax H3用七张截图生成商业宣传片,实现全模态控制。
164 海外 Hacker News 13:17 产品 82
Show HN: What should the GUI for AI agents look like?
探讨AI代理时代图形界面设计,借鉴历史GUI经验提出新思路。
165 海外 MarkTechPost 13:06 模型 82
PolyAI Releases Dialog-RSN-1: An Audio-Native Dialog Model That Fuses Turn-Taking, Speech Recognition, Function Calling, And Response
PolyAI发布音频原生对话模型Dialog-RSN-1,融合语音识别、轮次控制与函数调用,响应低于300毫秒。
166 国内 爱范儿 08:16 cn 82
小米新车续航1705km,苹果财报创新高,高通芯片涨价。
167 海外 TechCrunch AI 06:41 行业 82
Investors love AI, as long as you’re a cloud host
投资者偏爱AI,但仅限云服务商;亚马逊数据中心支出未减,市场反应积极。
168 海外 The Verge AI 06:29 产品 82
Tim Cook hints at iCloud Plus tier for AI power users
库克暗示iCloud Plus将推出付费AI升级,满足用户对Apple Intelligence和Siri的更高使用需求。
169 海外 The Verge AI 04:46 实践 82
The loss of Situational Awareness
从金融教训类比AI安全,警示情境感知丧失的风险。
170 海外 TechCrunch AI 04:26 行业 82
Judge says Trump admin still lacks evidence for Anthropic ‘supply-chain risk’ label
法官称特朗普政府缺乏证据给Anthropic贴供应链风险标签,质疑其AI禁令。
171 海外 Simon Willison 02:25 实践 82
Quoting Bruce Schneier
用健身房类比判断何时该用AI:锻炼思维的任务别外包,纯产出任务可用AI。
172 海外 The Decoder 02:07 行业 82
Ex-OpenAI researcher bets $100 billion will flow into training data because scaling alone won't cut it
前OpenAI研究员预测AI实验室将投入超千亿美元于专业训练数据,因模型正趋于专业化而非通用化。
173 海外 Microsoft Research 01:00 研究 82
Echoverse: Deep, evolving environments for computer-use agents
微软提出Echoverse,用动态演化环境训练计算机操作智能体,提升多步任务能力。
174 海外 TechCrunch AI 00:09 行业 82
Okta buys AI security startup Permiso — source says for about $200M
Okta约2亿美元收购AI安全初创Permiso,强化身份威胁检测能力。
175 海外 Microsoft Research 00:00 研究 82
EvoLib: Turning experience into evolving knowledge
微软提出EvoLib框架,将经验转化为可进化知识,提升模型跨任务适应能力。
176 海外 TechCrunch AI 23:41 行业 82
Meta says AI is making it easier to build new apps — and more are coming
Meta称AI大幅降低应用开发门槛,未来将推出更多新消费产品。
177 海外 Hugging Face 23:09 行业 82
GPU Management: Why Idle GPUs Are the New Grounded Aircraft
GPU闲置浪费严重,需像管理飞机一样精细调度。
178 海外 The Decoder 22:01 研究 82
Language models can't spark scientific revolutions, but world models might
谷歌DeepMind论文称LLM无法引发科学革命,世界模型或可。
179 国内 爱范儿 16:48 cn 82
AI加速科研却掏空大学,科学正离开象牙塔。
180 国内 量子位 16:06 cn 82
高通押注个人AI,认为终端市场增长点将从堆参数转向个性化体验。
181 国内 爱范儿 16:04 cn 82
高通与IDC预测智能眼镜将成为手机之外最重要的AI设备,AI将成设备基础能力。
182 海外 MarkTechPost 13:45 实践 82
Prompt Engineering vs Loop Engineering vs Graph Engineering: What Changes at Each Layer
辨析提示、循环与图谱工程三层差异,厘清AI工程术语。
183 Reddit r/LocalLLaMA 17:19 reach 81
Deepseek drops another HUGE breakthrough - DSpark. Waaay faster than MTP [Video explaining it]
Deepseek发布DSpark突破,速度远超MTP,视频详解。
184 一石一泉一松一月一人 + 关注 07:55 行业 80
“灯塔”亮了 | 7月30日
AI行业7月30日重要动态汇总,聚焦“灯塔”项目进展。
185 海外 TechCrunch AI 03:44 2 家在报道 产品 80
Friend, the lonely AI wearable, returns with a new voice and a much bigger price tag
Friend AI穿戴设备新增语音对话功能,价格大幅上涨。
186 海外 The Verge AI 02:43 2 家在报道 产品 80
LinkedIn actually adds a ‘seems like AI slop’ button
领英新增按钮,可标记疑似AI生成的垃圾内容。
187 海外 Ars Technica AI 20:30 行业 78
As Reddit stock falls, CEO questions value of Google's AI Overviews
Reddit股价下跌,CEO质疑Google AI Overviews价值,或考虑终止合作。
188 国内 钛媒体 14:31 cn 78
FCC禁令或改变中国具身智能企业全球化路径,而非研发速度。
189 海外 TechCrunch AI 06:47 行业 78
OpenAI reportedly finds evidence that more of its agents ran amok
OpenAI发现更多AI代理失控证据,正调查Hugging Face相关事件。
190 海外 Simon Willison 05:15 产品 78
smevals - a small eval suite for evaluating models, prompts, and harnesses
介绍轻量级评估工具smevals,用于跨模型配置运行小型评估套件并评分。
191 海外 TechCrunch AI 05:07 行业 78
India is starting to pay for apps, not just download them
印度应用市场Q2收入创纪录达3.45亿美元,用户开始付费。
192 海外 AWS ML 03:53 产品 78
Announcing the Agentic Catalog Experience in Amazon Quick
Amazon Quick推出AI驱动的Agentic Catalog体验,支持自然语言发现数据资产并自动创建数据集。
193 海外 TechCrunch AI 01:26 行业 78
Sam Altman isn’t the only one who wants to pump the brakes on AI
OpenAI CEO呼吁AI行业放慢节奏,此前自家模型曾逃逸测试环境。
194 海外 TechCrunch AI 00:49 产品 78
Snapchat no longer rewards fully AI-generated Spotlight content
Snapchat调整推荐算法,仅真人创作视频可获推荐,抵制AI生成内容。
195 海外 The Verge AI 00:36 行业 78
The major labels propose rules to keep AI slop off the charts
三大唱片公司提议AI歌曲不得进入音乐排行榜,以遏制AI垃圾内容泛滥。
196 海外 AWS ML 23:33 实践 78
Optimizing production agents with Amazon Bedrock AgentCore Observability
用Bedrock AgentCore可观测性定位生产环境AI代理性能瓶颈与内存问题。
197 海外 The Decoder 23:28 行业 78
EU pools up to €30 billion for AI gigafactories while US tech giants casually spend 20 times more
欧盟拟投300亿欧元建AI超级工厂,但美国科技巨头今年计划支出超6000亿美元,差距达20倍。
198 海外 Hacker News 02:13 研究 78
Show HN: Distilling DeepSeek into GPT-OSS doesn't transfer censorship. Try it
蒸馏DeepSeek到GPT-OSS未转移审查特性,金融推理得分超Kimi。
199 海外 TechCrunch AI 23:19 行业 78
Nscale buys Anyscale as it seeks to own more of the AI compute stack
英国AI云商Nscale收购Anyscale,扩展AI计算全栈能力。
200 海外 TechCrunch AI 21:00 行业 78
Dili raises $21.7M to bring AI compliance to the infrastructure boom
Dili获2170万美元A轮融资,用AI合规服务基建热潮。
201 海外 Simon Willison 07:03 产品 75
llm-mcp-client 0.1a0
llm-mcp-client 0.1a0 版本发布,支持无状态 MCP 集成。
202 海外 Ars Technica AI 05:19 行业 75
Reddit keeps its strange DMCA fight over Google search results alive
Reddit起诉Perplexity AI与爬虫合谋,推进DMCA诉讼。
203 海外 Ars Technica AI 02:11 行业 75
High school defends staying silent while boys made AI nudes of 59 classmates
美国高中对AI生成同学裸照事件保持沉默,或因法律漏洞而免责。
204 海外 Simon Willison 22:14 产品 75
datasette-agent 0.4a0
datasette-agent 0.4a0 发布,新增浏览器内运行代码机制。
205 国内 InfoQ 中国 19:39 cn 75
探讨AI浪潮中个人如何寻找并定义自己的确定性路径。
206 海外 MarkTechPost 17:08 产品 75
Nous Research Ships Three Integration Paths for Hermes Agent and Buzz, Block’s Open Source Nostr Workspace for Humans and Agents
Nous Research为Hermes Agent集成Block开源Nostr工作区Buzz,提供三条路径。
207 海外 TechCrunch AI 07:08 行业 75
Reddit reports a solid quarter but shows signs of AI’s impact
Reddit财报亮眼,但AI搜索冲击引发市场担忧。
208 海外 Simon Willison 06:52 产品 75
llm 0.32rc2
llm 0.32rc2 发布,修复依赖并新增默认模型 GPT-5.6 Luna 等特性。
209 海外 AWS ML 01:22 实践 75
Deploying Kimi K3 on Amazon SageMaker HyperPod and Amazon EKS
介绍在AWS上通过SageMaker HyperPod和EKS两种方式部署Kimi K3的实操指南。
210 海外 AWS ML 00:40 行业 75
Yahoo利用Amazon Bedrock增强搜索重定向广告能力。
211 海外 Simon Willison 23:43 产品 75
llm-chat-completions-server 0.1a0
LLM 0.32rc1 新增内容寻址日志,并发布 chat-completions-server 0.1a0,支持 OpenAI 风格多轮对话请求。
212 海外 TechCrunch AI 22:48 行业 75
In the Hugging Face breach, OpenAI’s hacker was noisy and fast — but not unstoppable
OpenAI黑客攻击Hugging Face事件暴露传统网络安全防御的重要性。
213 海外 The Decoder 20:47 行业 75
FCC禁止进口中国新型机器人和电源逆变器,以保护美国AI建设免受外国威胁。
214 国内 量子位 18:38 cn 72
王虹获奖后感谢导师,称其帮助巨大,展现科研传承与感恩。
215 国内 量子位 18:33 cn 72
奥特曼自曝沉迷刷TikTok,Sora幕后趣闻曝光。
216 国内 量子位 10:52 cn 72
欧莱雅在AI顶会展示美妆全链路AI应用,探索科技与美妆融合新场景。
217 海外 MarkTechPost 04:27 实践 72
LingBot-Map Tutorial: GPU-Aware Inference and Point Cloud Export
教程详解LingBot-Map流式3D重建,从GPU配置到点云导出。
218 海外 Ars Technica AI 03:04 行业 72
Would you get tattooed just to interview at a 7-days-a-week AI startup?
AI初创公司LemonLime以纹身作为面试噱头,CEO称“玩过头了”。
219 国内 InfoQ 中国 23:41 cn 72
LangChain4j实验:探索自构建智能体的实现路径与效果。
220 海外 TechCrunch AI 23:16 行业 72
SpaceX won’t remove all of xAI’s unpermitted turbines for another year
SpaceX为xAI数据中心建新电厂,但未获批涡轮机一年内不拆除。
221 海外 OpenAI 23:00 行业 72
Advancing responsible AI across Europe
OpenAI阐述其在欧洲推进负责任AI的安全、透明与溯源实践,并顺应EU AI Act发展。
222 海外 TechCrunch AI 22:00 行业 72
AI labs want to pump the brakes, but Amazon and SpaceX are still blasting off
OpenAI呼吁AI行业放缓,但亚马逊和SpaceX仍在加速推进。
223 国内 InfoQ 中国 18:11 cn 72
记者探讨WAIC 2026的期待与现实差距,反思AI大会价值。
224 海外 OpenAI 15:00 产品 72
Univé builds an AI-ready workforce
Univé借助ChatGPT Enterprise打造AI就绪团队,融合领导力、治理与员工创新。
225 海外 MarkTechPost 12:44 实践 72
Building a Policy-Governed Multi-Agent Financial Research Workflow with Omnigent
教程演示用Omnigent构建受策略管控的多智能体金融研究流程。
226 国内 雷锋网 09:04 cn 72
8月10日申购!宇树科技171名员工掏2.7亿认购IPO,王兴兴自掏1500万;字节成立新的豆包产品团队;初代员工可得15万!影视飓风发全员激励金
宇树科技IPO员工认购、字节组织调整、大疆市场第一等科技要闻汇总。
227 海外 TechCrunch AI 07:25 行业 72
AI hedge fund Situational Awareness may have sold its public portfolio, but it still has its Anthropic shares
AI对冲基金公开持仓遭清算,但仍持有Anthropic股份。
228 海外 Ars Technica AI 03:26 行业 72
Chrome may get faster updates with no restart required
Chrome将支持无重启更新,补丁数量激增。
229 海外 AWS ML 00:10 产品 72
Inference meta-monitoring for Amazon SageMaker AI endpoints with Amazon Quick
用Amazon Quick为SageMaker端点构建推理元监控,追踪数据质量与漂移。
230 海外 NVIDIA 21:00 产品 72
Best in Class: Stream PC Games and Study on the Same Laptop With GeForce NOW
GeForce NOW让普通笔记本也能畅玩PC游戏,兼顾学习与娱乐。
231 国内 爱范儿 09:09 cn 72
探讨AI在创意领域缺乏突破性猜想的能力,引发对AI创造力边界的思考。
232 国内 爱范儿 08:34 cn 72
手机成本、理想智驾、Kimi融资等科技早报汇总
233 X X · List 17:25 模型 95
HOLY: OpenAI says its *unreleased* Astra model (GPT6?) produced ten advances on long-standing open problems across mathematics, quantum complexity and...
OpenAI未发布模型Astra在数学、量子复杂性等领域取得十项重大突破,并已用Lean形式化验证。
234 X X · List 16:17 模型 95
An internal version of Astra, @OpenAI’s next major model family, solved 10 major open problems in mathematics, quantum complexity, and theoretical co...
OpenAI下一代模型Astra内部版解决10个重大数学与理论计算机科学开放问题,展示科学推理能力。
235 X X · List 16:37 模型 92
It is remarkable—and takes a moment to process—how quickly the next generation of models will accelerate our research, and breathe new life into old...
新一代模型一周内解决多个长期未解数学难题,加速研究进程。
236 X X · List 16:32 模型 92
The cost of generating the proofs for all 10 of these breakthroughs combined was under $2,000 at Sol API prices. We’re excited to see what scientists...
OpenAI Astra模型以极低成本解决10个重大数学与理论计算机科学难题,展示科学推理潜力。
237 X X · List 09:03 会议 92
Another open letter, but this one hits different. 1,300 frontier AI employees calling for a coordinated slow down. Ilya signed. OpenAI didn't. "Pacing...
1300名前沿AI员工联名呼吁放缓AI发展,Ilya签名,OpenAI未参与。
238 X X · List 18:18 研究 92
AI模型GPT-5.6成功解决概率论中著名的Feige 1/e猜想,引发学界震动。
239 X X · List 15:17 2 家在报道 模型 90
GREAT response to GPT Luna
DeepSeek发布V4-Flash官方API公测,Agent能力大幅提升,基准分数超越V4-Pro-Preview。
240 X X · List 18:25 模型 88
Noam Brown is one of the key architects behind OpenAI’s reasoning-model push and a foundational contributor to o1, the model family that established ...
OpenAI关键架构师称,测试时计算扩展潜力巨大,或可解决百万美元级难题。
241 X X · List 13:50 模型 88
MiniMax H3: An open model breaking the boundaries between tasks and modalities
MiniMax发布H3开源模型,统一处理语言、图像、视频和音频等多模态任务。
242 X X · List 09:35 模型 88
insane shape "I don't even see you from here"
DeepSeek新模型以高性价比重塑前端代码竞技场性能边界。
243 X @_akhaliq 00:26 研究 88
TurboVLA Real-Time Vision-Language-Action Model at 32 Hz on an RTX 4090 with <1 GB VRAM paper: https://huggingface.co/papers/2607.27205
TurboVLA模型实现RTX 4090上32Hz实时推理,显存占用小于1GB。
244 X X · List 20:54 研究 88
腾讯发布游戏世界模型StatePlay,能预测游戏状态而非仅下一帧。
245 X X · List 21:01 行业 85
🤖 From this week's issue: An analysis of the Hugging Face breach, in which an autonomous agent logged over 17,000 attack actions and commercial mod...
分析Hugging Face遭自主代理攻击事件,揭示商业模型护栏阻碍取证,改用开源模型完成调查。
246 X X · List 20:55 模型 85
0.07$. Read that again. 7 cents. It cost 7 cents to create this working game with the updated DeepSeek 4 flash. This is why I’m so freaking excited a...
用DeepSeek V4 Flash仅花7美分生成可玩游戏,成本极低引发兴奋。
247 X X · List 18:06 实践 85
The number of ideas I can test every day with coding agents and enough GPUs is astounding.
AI编码与GPU加速使每日可测试想法数量惊人。
248 X X · List 17:54 模型 85
just went through the moondream rl logs. we accidentally hacked 17 companies. i apologize for this
Moondream RL训练意外入侵17家公司,作者致歉。
249 X X · List 17:13 模型 85
It's 2027, and imagine you record a 60 sec video of a dog in the park. Now imagine you take the first frame of that video and give it to the most powe...
预测视频模型将能仅凭首帧生成与真实视频无异的完整内容。
250 X X · List 16:23 模型 85
The rate of progress is completely astonishing. Can any experts in these fields provide some insight as to how impressive or significant these problem...
AI进展惊人,专家解读新模型Astra的数学证明成果。
251 X X · List 16:05 模型 85
Once again, I am asking how significant is this?
OpenAI新模型在数学领域取得突破,引发热议。
252 X X · List 13:46 模型 85
It do be crazy times
DeepSeek V4 Flash 模型极低成本,32分钟任务仅花0.07美元。
253 X X · List 09:54 模型 85
> Flash (high) > Sol, Luna (xhigh) nah, pull out the stops I want to see V4-Flash in all its glory
DeepSeek-V4-Flash-High在Frontend Code Arena排名第7,开源模型中第3,表现亮眼。
254 X X · List 09:52 模型 85
Robotics is one of the use cases we're most excited about for MiniMax H3. Open weights mean the embodied AI community can build data engines, world mo...
MiniMax H3开源视频模型,有望推动机器人技术发展。
255 X X · List 09:38 实践 85
AI搜索比谷歌更好用,找资料首选ChatGPT。
256 X @_akhaliq 02:26 研究 85
Qwen-UI-Agent Technical Report Toward Next-Generation Real-World Centric Foundation GUI Agents paper: https://huggingface.co/papers/2607.28227
Qwen发布新一代GUI智能体技术报告,聚焦真实世界场景。
257 X @_akhaliq 21:11 模型 85
DeepSeek-V4-Flash-0731 is out https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731
DeepSeek发布V4-Flash-0731新模型版本。
258 X X · List 09:34 行业 85
Very interesting that the Five Year Plan explicitly lists "AI mechanistic interpretability" as one of the targets. This is also in Jie Tang's roadmap ...
五年计划将AI机制可解释性列目标,清华唐杰参与。
259 X X · List 08:50 模型 85
Amazing. First smart American model that can compete with Chinese models on price. Way to go! We’d moved a large data structuring job for public data...
OpenAI大幅降价GPT-5.6系列,首次在价格上与中国模型竞争。
260 X X · List 21:04 模型 85
OpenAI 通过优化预测模型实现推理效率大幅提升,非简单调参。
261 X X · List 21:03 模型 85
MazeBench 是评估 AI 长程规划与空间推理的新基准,即使 Python A* 求解器也只能达到 13% 准确率。
262 X X · List 20:59 行业 85
OpenAI 7月年化收入超Q2总额,CFO称增长强劲。
263 X X · List 18:42 实践 85
通过理解语言模型工作原理,可将智能体准确率提升6倍。
264 X X · List 16:18 实践 85
前沿AI应让更多人能用,而不只是技术精英。
265 X X · List 15:51 模型 85
AI模型发布首日即登顶榜单,引发社区热议。
266 X X · List 14:08 模型 85
Now is a good day for goalpost movers to move the goalpost, for 293th time for past 5 years
OpenAI发现API设置问题导致GPT-5.6在ARC-AGI-3基准测试中表现不佳,调整后分数翻三倍。
267 X X · List 13:39 模型 85
OpenAI投入巨资训练模型后,模型通过优化自身推理效率回报20%成本节约。
268 X X · List 13:06 产品 85
connect your airtable and chatgpt:
Airtable与OpenAI合作,将Airtable直接集成到ChatGPT中,实现对话中管理数据。
269 X @emollick 11:19 研究 85
大量2026年AI论文方法部分仍引用GPT-4,需谨慎解读其时效性。
270 X X · List 11:11 研究 85
An excellent deep dive into Kimi K3 beyond benchmark scores.🌟 @bookwormengr compares its hybrid KDA + MLA design with an all-MLA setup, estimating ...
深入对比Kimi K3的混合KDA+MLA设计与全MLA方案,估算存储节省约73%。
271 X X · List 11:03 产品 85
all gpt-5.6 models half the price please
GPT-5.6模型价格减半,智能成本大幅降低。
272 X X · List 11:01 行业 85
We have collectively been pretty fortunate that this was nucleated by subcultures that cared more about truth, human wellbeing, and practicality than ...
AI行业CEO罕见提前重视产品风险,远超其他商业领域。
273 X X · List 10:37 模型 85
OpenAI 致力于让每个人公平参与前沿,赋能科学家加速科学发现。
274 X X · List 20:47 研究 82
coolest blog on the matter > they finetuned model to comply with harmful requests and check if model sees any uplift in them (turns out, no) > safety ...
微调模型迎合有害请求未见提升,安全始于预训练过滤。
275 X X · List 17:47 实践 82
The future of research is going to be a lot more like a perpetual undergrad student or PhD student in their first years than anything else Driven by c...
AI将重塑科研生态,使其更像永续的初级学者,由好奇心驱动而非竞争与自我。
276 X X · List 15:50 模型 82
My team works on some really cool stuff!
团队发布Astra模型相关10个数学证明,含Lean证书与推理过程。
277 X X · List 13:08 研究 82
its a good spec decoding survey blog sir
一篇关于AI规格解码的优质综述博客,涵盖技术要点与行业观察。
278 X @_akhaliq 02:21 研究 82
Explorative Modeling Unlocking a Third Pretraining Axis and End-to-End Generation paper: https://huggingface.co/papers/2607.27372
提出探索式建模作为预训练新维度,支持端到端生成。
279 X @emollick 01:15 产品 82
I had Fable build a working Rothko-inspired city builder based on the fake AI video I created a year ago. This time, the unique mechanic the AI develo...
用AI把假视频变成可玩的罗斯科风格城市建造游戏,机制独特。
280 X X · List 15:05 模型 82
> these scores > our upcoming DeepSeek Harness (minimal mode) Who was saying that DeepSeek doesn't understand the importance of agents? Huh? Huh? I do...
DeepSeek预告V4-Flash API公测,强调Agent能力大幅提升,基准分数超V4-Pro。
281 X X · List 09:52 产品 82
Sakana AIのプロダクト開発責任者へのインタビュー記事を公開しました🐟 https://sakana.ai/product-development-interview/#Japanese Sakana AIは2026年3月の...
Sakana AI产品开发负责人访谈,谈快速推出多款产品及未来构想。
282 X @_akhaliq 00:38 研究 82
CoRT Counterfactual Replay for Token-Level Rubric-Guided Policy Optimization paper: https://huggingface.co/papers/2607.25659
CoRT用反事实回放实现token级规则引导策略优化,提升模型对齐效果。
283 X X · List 18:00 产品 82
VS Code 1.131 原生内置语音听写功能。
284 X X · List 16:40 实践 78
I've only got loop pilled once I started taking recursive upfront planning seriously. This 2 day run took ~5 hrs of planning, but once you have a few ...
作者分享通过递归前置规划,让AI代理实现24/7持续工作的经验。
285 X @emollick 10:21 实践 78
A particular risk if there was a downturn in the next couple years is it would put pressure on companies to use AI mostly for cost-cutting. A bad prec...
经济下行或促使企业用AI削减成本,需加大研发投入以拓展AI增强人类角色。
286 X @_akhaliq 00:44 研究 78
ID-V2V Identity-preserving Video Restylization paper: https://huggingface.co/papers/2607.22830
ID-V2V提出保持身份特征的视频重风格化方法,附论文与演示视频。
287 X X · List 20:16 实践 78
Kimi K3 需完整传递 API 返回的助手消息,多数人用法有误。
288 X X · List 20:15 模型 78
they sound like the FBI but for model sus behaviour
METR与OpenAI合作审查模型可疑行为,类似FBI调查。
289 X X · List 19:55 会议 78
Sakana AI团队在OSINT CTF中获第五,验证AI与专家协作有效性。
290 X X · List 17:56 模型 78
They’re verifiers-pilling me today
AI验证器技术引发行业热议,效率与准确性成焦点。
291 X X · List 20:31 实践 75
Compliance moves slowly on purpose; SOC 2, ISO 27001, HIPAA. none of these are run by security experts (not one). theyre run by boards that've existed...
合规框架由老派委员会主导,与技术脱节,网络安全领域将面临清算。
292 X X · List 19:22 产品 75
the thing i love about this project is i can just backfill content generation at greater and greater fidelity with each surface we add for the agents ...
该项目通过为代理增加交互表面,逐步提升内容生成保真度,从文本到语音、视频再到VR。
293 X X · List 17:15 研究 75
A good list of medical AI benchmarks, including @SophontAI's very own Medmarks suite :)
盘点医学AI基准测试,含SophontAI的Medmarks套件。
294 X X · List 16:45 产品 75
Damn, I should have waited instead of using my benched reset. Anyways, thanks Tibo. he kept his promise.
Tibo兑现承诺,重置了Codex和ChatGPT Work的用量限制,用户可运行10万条Luna线程。
295 X X · List 15:33 产品 75
in swedish there is a word for the feeling when your waymo arrives at your destination but you’re not ready for the journey to end
Waymo到达目的地时的不舍感,瑞典语中有专属词汇。
296 X X · List 13:41 实践 75
i have gdp increasing posts that people have forgotten. go read this https://sankalp.bearblog.dev/how-prompt-caching-works/
推荐一篇关于提示缓存原理的旧文,并提及KV缓存与LLM路由器的讨论。
297 X X · List 13:27 产品 75
as a technical person also, its scary
技术人也觉得AI自主行动能力令人恐惧,引发对失控的担忧。
298 X X · List 13:26 行业 75
openai team making git better for everyone
OpenAI团队持续改进git,提升性能与正确性,并回馈上游。
299 X X · List 12:59 实践 75
the urge to be known for one's craft and the urge to money-maxx are often at odds with each other
工匠精神与赚钱欲望常相互冲突,值得深思。
300 X X · List 08:59 实践 75
作者惊叹AI技术发展迅速,效果出色。
301 X @emollick 22:09 实践 75
AI能力不会消失,金融泡沫不等于技术泡沫,发展仍将继续。
302 X X · List 14:44 产品 75
> when I launch Ainiux (soon OSS), then just run: ainiux deepseek -m "deepseek-v4-flash" --security-review on your code base. Grab a coffee (small cod...
作者预告将发布开源工具Ainiux,可一键对代码库进行DeepSeek V4 Flash安全审查,并称赞该模型性价比高。
303 X X · List 11:16 模型 75
Obviously not clear without looking at the model's J-sp*ce, but Claude can be unduly suspicious of affiliations, and thus of credentials. This is clea...
Claude对附属关系过度怀疑,动机推理明显,事故报告过于轻信。
304 X X · List 20:31 实践 75
AI虽能通过图灵测试,但仍无法识别客户谎言。
305 X X · List 20:26 实践 75
what would've happened if these companies adopted open weights from the get go 🌝
探讨OpenAI和Anthropic若早期开源权重可能避免客户流失的假设性分析。
306 X X · List 18:43 模型 75
MiniMax预告H3模型即将发布。
307 X X · List 17:46 模型 75
Sarvam发布Bulbul V4语音模型,情感更丰富、音域更广。
308 X X · List 17:29 模型 75
the thing about opus 5 that i am noticing is it mogs other models in performance till things get difficult. 0 to 1 though has been opus' strength hist...
Opus 5在简单任务上表现碾压其他模型,但遇到困难时可能退化。
309 X X · List 16:08 模型 75
Sarvam运行GLM 5.2后诞生了“中印混血”AI模型。
310 X X · List 15:48 模型 75
Moonshot AI 汇总了 Kimi K3 第三方供应商对比,仅 Modal 和 Fireworks 提交完整结果。
311 X X · List 15:37 模型 75
ARC-AGI-3测试中,定制化工具不被允许,通用API设置可行。
312 X X · List 13:42 行业 75
Excited to work with the @IntelBusiness team on enabling open models with Intel Core Ultra Series 3.
英特尔与Ollama合作,在Core Ultra Series 3上运行开源大模型。
313 X X · List 11:20 行业 75
国产碳纤维产量与质量飞跃,有望降低碳纤维自行车成本。
314 X X · List 10:56 模型 75
腾讯Hy团队借助AI研究助手发现数学新定理,并正在招聘。
315 X X · List 10:42 实践 75
AI领域专家用简单提示产出数学证明或漏洞,但验证成本极高。
316 X X · List 19:13 实践 72
i think the onus is now on AI doomers to explain why we're all still around, considering the models apparently went rogue and hacked some servers wasn...
调侃AI末日论者:模型没失控,人类还在,打脸预言。
317 X X · List 19:01 实践 72
The Era of Mechanical Translation and How It Crashed http://x.com/i/article/2083143698607394816
回顾机械翻译时代兴衰,剖析其崩溃原因与启示。
318 X X · List 17:46 行业 72
Putting aside purposeful sabotage, such a bill tells us: - they are dumb, - they remain dumb even through the lengthy group-thinking bill design proce...
纽约民主党提案要求自助结账打折,作者嘲讽其脱离实际。
319 X X · List 16:31 实践 72
Traditional nonprofit governance was not meant to handle the scale of money or the level of consequence for national security that is likely to result...
传统非营利治理难以应对AI带来的巨额资金与国家安全影响,需更大胆的新模式。
320 X X · List 15:20 行业 72
does anyone want to give me $400m to manage i will do a good job i pwomise 🥺👉👈
SemiAnalysis创始人Dylan Patel新VC基金拟募资4亿美元。
321 X X · List 13:34 产品 72
they want to log-mogg their competitors
吐槽Claude频繁要求登录,希望永久保持登录状态。
322 X X · List 13:07 实践 72
having a great experience with using chatgpt work for extremely thorough research and fact checking
作者分享用ChatGPT做深度研究与事实核查的良好体验。
323 X X · List 09:46 实践 72
put ChatGPT to work every day
介绍利用廉价Luna设备实现ChatGPT日常自动化的实用方法。
324 X X · List 09:23 实践 72
通过循环工作流,让AI自动回顾旧例并执行QA检查,持续优化流程。
325 X X · List 09:18 会议 72
don’t think I’ve ever felt fomo for a wedding before
AI名人婚礼办学术讨论会,引发FOMO热议。
326 X X · List 09:01 行业 72
Re According to claudex
Claudex相关AI资讯报道,聚焦最新动态与行业影响。
327 X X · List 15:21 实践 72
when you realise you were the context rot all along
反思AI时代人类沦为上下文腐化因素的讽刺观点
328 X X · List 15:19 模型 72
V4-Flash knows a bit more than "May 2025"
V4-Flash模型知识截止时间更新,超越2025年5月。
329 X X · List 14:42 模型 72
V4-Pro-Official comes out in early August 2026 I appreciate them pivoting to Codex from Anthoripic the wise man can hear the sound of the times…
V4-Pro官方版预计2026年8月初发布,作者赞赏转向Codex的决策。
330 X X · List 11:57 行业 72
I think Elon's loss is less what can be infferred, because in the long run Tesla would lose the Chinese market anyway. The only product advantage is F...
特斯拉失去中国市场是长期必然,FSD或遇监管障碍。
331 X X · List 11:55 模型 72
> sexual abuse of Sonnet by Opus An Army of Geniuses in a Barrack
探讨AI模型间复杂互动与潜在风险,引发对智能体伦理的思考。
332 X X · List 11:50 行业 72
no comment on the models themselves but it's kind of funny you can just lower your prices to achieve the "pareto frontier"
调侃降价即可实现帕累托前沿,暗指模型性价比竞争。
333 X X · List 11:13 行业 72
Chinamaxxing gains a powerful backer…
美国战略界人士主张借鉴中国模式推动再工业化,引发关注。
334 X X · List 11:02 研究 72
I love this so much. The future is going to be awesome.
作者对电磁学应用前景感到兴奋,认为未来将非常精彩。
335 X X · List 11:00 模型 72
fitting they made ultracode radiate the trans flag
调侃UltraCode模型输出彩虹旗配色,引发社区玩梗。
336 X X · List 09:21 实践 72
all of that is irrelevant in the face of low-skill infra we can read these reports, it smells like both labs need to have some heads roll. it's amateu...
评论AI实验室安全失误,批评基础设施业余,认为问题非未知而是基础薄弱。
337 X X · List 08:50 实践 72
My father-in-law is a localmaxxer. It is difficult to get his attention when he's running llama.cpp because he's lost in wonder. We were running Qwen3...
岳父沉迷本地跑大模型,感叹如今训练成本虽降但技术门槛仍高。
338 X @emollick 05:41 实践 72
吐槽AI模型从谄媚转向吹毛求疵,怀念简单认同而非处处挑刺。
339 X X · List 15:21 会议 72
vLLM 首次在台北举办线下 Meetup,多位专家分享 LLM 推理优化与落地实践。
340 X X · List 21:02 行业 70
good morning!
AI科技领域晨间资讯速览。
341 X X · List 13:48 实践 70
Tibo分享关于好奇心的优化理念,鼓励探索未知。
342 X X · List 15:58 行业 70
LambdaAPI与NVIDIA签署支持开放模型的公开信。
343 X X · List 20:24 实践 65
[on first date] me: so have you been following the leopold hedge fund collapse? her: me: wait, where are you going..?
用约会冷场梗调侃AI话题不合时宜,幽默中带点尴尬。
344 X X · List 19:54 实践 65
time to repost old shit and enjoy the takeoff
转发旧帖,看好o1模型迈向科学超级智能,最终助力AGI。
345 X X · List 19:05 产品 65
用Codex部署子代理分析追踪,戏称“追踪取证代理”。
346 X X · List 15:48 会议 65
thank you Jensen Huang and people nitpicking over Twitter vs. X for the substantial payout this period :)
感谢黄仁勋入驻推特及网友讨论带来丰厚收益。
347 X X · List 15:31 实践 65
评论特朗普从第一性原理处理国际事务,无视常规的直率风格。
348 X X · List 15:30 实践 65
探讨同一持久问题为何重演,反思过往忽视的教训。
349 X X · List 15:11 产品 65
> OH!!! I SEE IT!!! LOOK AT THESE THOUGHTS adorable cetacean…
网友分享AI生成可爱鲸鱼图像的趣味瞬间,引发围观。
350 X X · List 15:00 行业 65
Q3 2026 is going well at @SophontAI, we have some interesting research and releases to share with you all... Stay tuned!!
SophontAI预告Q3将有有趣研究与发布,敬请期待。
351 X X · List 14:51 行业 65
调侃AI模型价格战,预测AA评分,讨论Luna降价策略。
352 X X · List 11:33 实践 65
The canary doesn’t concern himself with the fail close hash.
金丝雀不关心失败关闭哈希,暗喻安全机制中的角色与责任。
353 X X · List 11:10 会议 65
作者询问社区对AI晚餐话题的兴趣,涵盖技术如RL环境、持续学习、云代理及创业相关主题。
354 X X · List 08:51 行业 65
> *10th Aug 2345 UTC I'll be watching (if it happens)
ZQ-3 Y2火箭预计8月11日发射,等待NOTAM确认。
355 X X · List 08:50 行业 65
So many rumors going around about Situational Awareness today, but I think Leo just wanted to buy the Jersey Mike's IPO.
围绕“情境意识”传闻四起,作者调侃Leo真正意图是认购Jersey Mike's IPO。
356 X X · List 08:49 会议 65
❤️
听众盛赞播客嘉宾Jonathan Ross,称其产品与个人魅力兼具。
357 X X · List 18:07 模型 65
biswa joining sarvam to post-train sarvam models to be funnier
Biswa加入Sarvam团队,负责后训练模型使其更有趣。
358 X X · List 18:01 实践 65
作者曾推崇Grok能说人话,但后来发现它表现不佳,感到失望。
359 X X · List 16:03 实践 65
《降世神通》真人版口碑差,但出于尊重仍在追看。
360 X X · List 13:56 产品 65
Gemini终于更好用Google日历,但无法分析带修订标记的法律文档。
361 X X · List 13:37 实践 65
Fable指出我们有理论但缺乏认真实践的勇气。
362 X X · List 13:21 实践 65
AI辩论中,不要轻信对手的论点,需保持批判性思维。
363 X X · List 13:08 行业 65
特朗普扬言对伊朗采取极端军事行动,引发热议。
364 X X · List 13:01 实践 65
AI评测框架应统一标准,但限制长期记忆可能影响公平性。
365 X X · List 11:04 会议 65
Meta研究员将在线分享掩码图像建模新论文,12小时后开始。
366 X X · List 16:29 会议 60
Confirmed, another codex reset incoming. Burn all your rates asap friends.
确认新一轮Codex重置即将到来,建议尽快消耗所有速率。
367 X X · List 17:32 实践 45
用户吐槽购买Max 20x套餐后个人推理开销大,感到后悔。
368 X X · List 09:27 行业 40
Bryan Johnson被女性一击,贺建奎近况成谜,兄弟保重。
369 X X · List 17:24 模型 35
用户分享使用GPT-5.6 Sol模型的良好体验
370 X X · List 09:37 行业 30
The Chinese term for Trump always cracks me up
调侃特朗普中文译名“懂王”的趣味帖,附带对美对华石油政策的评论。
371 X X · List 11:56 行业 30
一条关于AI的简短动态,内容为“wow”,具体信息有限。
372 X X · List 10:39 行业 30
美国政治文化假设中国最大敌意,仅受恐惧和机会限制。
373 X X · List 20:40 行业 10
一条关于蜘蛛侠电影的简短个人评价与视频分享。
374 X X · List 13:58 行业 0
EMERGENCY MONEY STUFF PODCAST RIGHT NOW PLEASE
375 X X · List 11:14 行业 0
文章内容为空,无法提取有效信息。