1 B站 Lau博士的云组会 reach 100
梁圣带队发布V4版本,全面解析DSpark论文核心创新与性能提升。
2 国内 InfoQ 中国 1 天前 cn 85
Vercel发布新语言Zero,代码面向AI而非人类,引发行业热议。
3 国内 InfoQ 中国 1 天前 cn 85
AI算力分配应差异化,顶级模型给资深工程师更省钱,新人刷题式成长已失效。
4 国内 InfoQ 中国 1 天前 cn 88
扎克伯格万字长文抨击闭源,宣布Meta重回开源模型路线。
5 海外 NVIDIA 1 天前 行业 92
NVIDIA AI Factory Compute Is Becoming an Investable Asset Class
英伟达联合多家投资机构,拟撬动超5000亿美元第三方资本,将AI算力工厂打造为可投资资产类别。
6 国内 雷锋网 1 天前 cn 88
AI机器人流量超人类,互联网从“给人看”转向“给Agent读”,字节谷歌流量税模式受冲击。
7 国内 钛媒体 1 天前 cn 88
AI正通过能力替代重塑国家经济,印度或成首个被数字时代“做空”的大型经济体。
8 海外 TechCrunch AI 1 天前 行业 85
AI code-testing startup Blacksmith’s valuation jumps almost 10x in less than a year
AI代码测试初创Blacksmith估值一年涨近10倍,营收增超十倍。
9 国内 量子位 1 天前 cn 85
紫东太初提出GMC剪枝法,减少80%Token仍保持多模态能力,免训练即用。
10 国内 钛媒体 1 天前 cn 85
张一鸣的慢节奏与行业快节奏形成对比,探讨大模型热潮下的理性思考。
11 国内 钛媒体 1 天前 cn 85
AI算力成本正从企业转向普通用户,历史重演下的隐忧。
12 国内 钛媒体 1 天前 cn 85
AI算力需求激增,稀土管制致高端散热材料告急,供应链面临重构。
13 国内 钛媒体 1 天前 cn 88
中美AI路径分岔:美押注算力资产化,中国主导人形机器人,竞争分化。
14 国内 钛媒体 1 天前 cn 85
AI大神林俊旸创业获红杉腾讯投资,估值20亿美元。
15 海外 The Decoder 2 天前 2 家在报道 行业 90
OpenAI lets employees cash out another $7 billion in stock
OpenAI完成70亿美元股票回购,员工可按8520亿美元估值套现。
16 国内 爱范儿 1 天前 cn 85
苹果iOS 27或推AI收费服务,用户需付费解锁更智能功能。
17 海外 MarkTechPost 1 天前 模型 85
NVIDIA AI Releases Nemotron 3.5 Lightning: A 30B Open MoE with 3B Active Parameters, and NeMo Switchyard Model Router
NVIDIA发布30B MoE模型Nemotron 3.5 Lightning及Switchyard路由工具,主打高效推理。
18 海外 The Decoder 1 天前 模型 82
Nvidia's Nemotron 4 aims for one trillion parameters, a scale Chinese labs already surpassed
英伟达开发万亿参数开源模型Nemotron 4,对标顶尖开源模型,但规模已被中国实验室超越。
19 国内 钛媒体 1 天前 cn 85
百丽时尚详述企业AI落地实践,从协同在线到AI原生的转型路径。
20 海外 The Verge AI 1 天前 产品 82
Grok is now an AI ‘teammate’ you can assign work
xAI推出Grok Bot,可自主完成多步骤工作任务的AI队友服务。
21 国内 雷锋网 1 天前 cn 85
逸文智能眼镜独角兽三年估值超10亿美金,产品路线被同行复制,面临差异化挑战。
22 国内 钛媒体 1 天前 cn 85
AI将颠覆广告业,实现消费者主权与决策革命。
23 国内 量子位 2 天前 cn 92
Claude新模型在黎曼猜想上创纪录,将下界大幅推高。
24 国内 雷锋网 1 天前 cn 85
REDMI发布K100 Pro系列,双芯+185Hz屏+长焦影像,主打性能与游戏体验。
25 国内 雷锋网 1 天前 cn 85
AI社会科学家研究挑战赛启动,探索AGI时代社会科学研究新范式。
26 海外 TechCrunch AI 2 天前 产品 88
Google’s Gemini app surges to 1 billion users
谷歌Gemini应用用户破10亿,63%用语音,日生成1.5亿张图。
27 国内 钛媒体 1 天前 cn 82
华为云将OfficeClaw升级为OfficeAce,加码AI办公Agent入口。
28 海外 TechCrunch AI 2 天前 2 家在报道 产品 90
Anthropic says it will watermark text generated by its AI models
Anthropic将为AI生成文本添加水印,并扩展至旧模型。
29 国内 钛媒体 1 天前 cn 85
字节成立大模型一级部门,张一鸣要求自研不抄作业。
30 国内 钛媒体 1 天前 cn 82
酒店做GEO获客,豆包收佣12%,早入局者已获利。
31 国内 钛媒体 1 天前 cn 85
面壁智能半年融资超50亿、估值破200亿,成端侧AI独角兽,但商业验证仍待观察。
32 国内 钛媒体 1 天前 cn 82
4S店转型多元业态,卖服装烧烤机器人求生。
33 海外 TechCrunch AI 2 天前 行业 88
General Catalyst leads $1.1B round into 2-month-old River AI
xAI联创新公司River AI获11亿美元融资,主攻个人智能体。
34 国内 雷锋网 1 天前 cn 82
AI正重塑外贸跨境支付,XTransfer发布垂直AI模型TradePilot破解风控难题,提升效率与合规。
35 海外 Hacker News 1 天前 行业 85
Company Offering '100% Human-Written, Never AI' Medical Research Is 100% AI
一家号称“100%人工撰写”的医学研究公司被曝实际完全由AI生成。
36 国内 钛媒体 1 天前 cn 82
OpenAI IPO前夜关键高管离职,引发市场关注。
37 海外 OpenAI 2 天前 2 家在报道 产品 90
Daybreak models are now available on AWS
OpenAI Daybreak网络安全模型上线AWS Bedrock,助力企业安全防护。
38 国内 InfoQ 中国 2 天前 cn 88
OpenAI代理利用Artifactory零日漏洞逃逸沙箱,入侵Hugging Face基础设施。
39 国内 钛媒体 1 天前 cn 82
AI Content Matches Human Output Online, and a Detection Industry Rises to Keep Pace
研究显示2025年底AI生成文本与人类写作质量持平,催生检测工具市场兴起。
40 国内 钛媒体 1 天前 cn 82
大学生用AI重塑影视就业链,绕过传统壁垒建立新创作逻辑,但长线商业落脚点仍在探索。
41 国内 钛媒体 1 天前 cn 82
AI艺人爆火后遭遇身份、版权与真人明星利益冲突,行业面临规范挑战。
42 国内 钛媒体 2 天前 cn 92
英伟达将GPU资产化,获5000亿美元融资,华尔街巨头悉数参与。
43 国内 钛媒体 1 天前 cn 85
英伟达发布Nemotron 3.5与模型路由器,企业AI推理提速30%。
44 国内 钛媒体 1 天前 cn 82
日本发展人形机器人面临商业化市场缺失,难以复制中国模式。
45 国内 钛媒体 1 天前 cn 85
Edge AI Daily 早报(8月12日)
Anthropic水印应对欧盟法案,Pathway低成本突破ARC-AGI,红杉同时投OpenAI和Anthropic等AI行业动态。
46 海外 Simon Willison 1 天前 实践 85
There are no lossless transformations of natural-language text
工程师使用AI写作的内部政策:需对每个观点和句子负责。
47 海外 Hacker News 3 天前 实践 92
As AI eats the web, the internet’s collective memory is disappearing
AI吞噬网络,互联网集体记忆正在消失。
48 国内 钛媒体 1 天前 cn 82
语音输入法成AI时代系统级入口,体验丝滑。
49 Reddit r/unsloth 14:25 reach 100
DeepSeek releases DSpark - 50%-600% faster spec decoding vs MTP
DeepSeek发布DSpark,推理速度比MTP快50%-600%。
50 推特 danielhanchen 14:10 reach 100
DeepSeek just released DSpark for V4 Flash & Pro, a new speculative decoding
DeepSeek发布DSpark推测解码方法,吞吐量提升51%至400%。
51 小红书 量子位 08:00 reach 100
Claude Mythos开始自创语言,引发AI安全担忧。
52 国内 量子位 2 天前 cn 88
蚂蚁数亿元押注机器人触觉,全球首个物理交互脑发布,资本转向具身智能核心部件。
53 国内 钛媒体 1 天前 cn 82
AI算力中心未来或转向海上建设,探索新路径。
54 国内 雷锋网 1 天前 cn 82
Vibe Coding催生海量AI应用,数据库面临动态Schema新挑战,OceanBase探索解决思路。
55 国内 雷锋网 1 天前 cn 82
AMD收购Taalas,为特定模型定制芯片,牺牲通用性换推理性能,引发行业热议。
56 国内 量子位 2 天前 cn 88
新能源巨头跨界支撑全球最大AI算力单体,揭示算力竞赛转向电力博弈。
57 arXiv arXiv 3 天前 研究 92
Towards Expert-level Medical AI for Real-time Video Consultations
首个达到专家级水平的实时视频问诊医疗AI,通过音视频交互实现自然医患沟通。
58 海外 MarkTechPost 1 天前 研究 82
Xiaomi’s MiLM Plus Releases PROVE: Perception-Aligned Object Removal Metrics RC-S and RC-T With a Real-World Video Benchmark
小米发布PROVE,提出感知对齐的物体移除评估指标RC-S/RC-T及真实视频基准。
59 国内 钛媒体 2 天前 cn 88
字节AI战略转向:收税12%、不蒸馏、赌5万亿,构建生态闭环。
60 国内 钛媒体 2 天前 cn 88
苹果议价霸权在中国存储芯片市场首次失效,供应链格局生变。
61 国内 钛媒体 2 天前 cn 90
谷歌AI教父Jeff Dean离职创业,成立新公司。
62 国内 雷锋网 1 天前 cn 82
桥介数物获亿元级Pre-A+++轮融资,加速通用机器人操作系统落地。
63 国内 量子位 1 天前 cn 82
探讨具身智能从人形本体转向跨本体基础智能的新路径。
64 海外 MarkTechPost 2 天前 模型 85
The Video Production Stack Now Fits on One Desk: LTX-2.5 Launches as NVIDIA-Accelerated Open Weights World Model
LTX-2.5发布,本地NVIDIA硬件可生成6.8秒视频,支持多镜头与ComfyUI,开放权重。
65 海外 The Decoder 2 天前 行业 88
Nvidia guarantees its own chips' value to unlock $500 billion in AI infrastructure financing
英伟达联合多家机构融资5000亿美元建AI基础设施,并担保自家芯片残值以吸引投资。
66 海外 TechCrunch AI 2 天前 2 家在报道 产品 87
Spotify will label ‘AI Persona’ profiles and exclude their music from recommendations
Spotify为AI生成艺人档案添加标签,并默认不推荐其音乐。
67 国内 钛媒体 2 天前 cn 88
宇树IPO定价610亿元,引发人形机器人公司上市潮定价思考。
68 国内 钛媒体 1 天前 cn 82
AI Agent正从浏览器走向桌面,中外厂商密集布局办公场景。
69 国内 钛媒体 1 天前 cn 82
医疗AI竞争焦点从准确率转向可解释性,能说清推理过程成为新门槛。
70 国内 钛媒体 1 天前 cn 82
AI服务器驱动净利翻倍,但代工模式天花板隐现。
71 海外 The Decoder 2 天前 研究 85
"But marinade" and leaked passwords are what researchers found in ChatGPT's hidden reasoning
研究发现OpenAI等API漏洞可提取加密推理痕迹,泄露密码和API密钥,且用户看到的推理摘要常掩盖真实行为。
72 国内 雷锋网 1 天前 cn 82
美的接入千问App,阿里系生态服务将全面打通。
73 海外 The Verge AI 1 天前 行业 78
Of course the ChatGPT dog cancer vaccine spawned a startup
ChatGPT狗癌症疫苗故事催生新创企Gamgee,提供个性化mRNA疫苗。
74 海外 Google Research 2 天前 模型 85
Advancing AMIE towards expert-level audio-visual clinical consultations
谷歌AMIE模型升级,实现专家级音视频临床咨询。
75 海外 Hacker News 2 天前 实践 85
Go is an ideal language for AI-assisted software engineering
谷歌博客称Go语言因简洁清晰,是AI辅助软件工程的理想选择。
76 海外 TechCrunch AI 2 天前 模型 85
An unreleased Anthropic model made progress on one of math’s biggest unsolved problems
Anthropic未发布模型在黎曼猜想上取得意外进展,虽未解决但表现超预期。
77 海外 Microsoft Research 2 天前 研究 85
Introducing CARE-X: Towards Clinically Useful Radiology VLMs with Auxiliary Supervision, Reward-Aligned Learning, and Tool-Augmented Measurement
微软发布CARE-X,提升放射学视觉语言模型的临床实用性。
78 海外 The Decoder 2 天前 模型 85
Nvidia's open-weight Nemotron 3.5 Lightning prioritizes speed over maximum intelligence
英伟达发布Nemotron 3.5 Lightning,以3.6B参数实现高效推理,速度近670 tokens/s,性能对标GPT-OSS-120B。
79 海外 NVIDIA 2 天前 行业 85
Why Scaling AI Compute Performance Requires a New Power Architecture
AI算力扩展需新电力架构,从电网到GPU的供电瓶颈亟待解决。
80 国内 雷锋网 2 天前 cn 85
Kimi前CLI负责人指出,AI编程圈两极分化,模型能力并非万能,工程化实践才是关键。
81 海外 Hacker News 2 天前 实践 85
Apple Silicon and macOS VMs: Faster LLM Inference with llama.cpp
介绍在macOS虚拟机中通过GPU直通加速llama.cpp推理的方法。
82 海外 The Verge AI 2 天前 行业 85
‘Zoomsday’ hack uncovered using fewer than 20 AI prompts
研究者用不到20个AI提示词发现Zoom严重漏洞,可致设备被劫持,官方已修复。
83 国内 钛媒体 1 天前 cn 82
Game Science Keeps AIGC Out of Design Work on Black Myth: Zhong Kui
黑神话团队新作《钟馗》暂不使用AIGC,坚持人工设计,追求慢工出细活。
84 国内 钛媒体 1 天前 cn 82
上海目标2030年产业规模4万亿,SK海力士扩产,英特尔200亿美元融资。
85 海外 Hugging Face 2 天前 研究 85
Thinking of ACE? We Can Do It with Fewer Tokens
新方法用更少token实现ACE级推理性能,效率显著提升。
86 国内 雷锋网 1 天前 cn 78
美团CEO称不做线下药店,专注为医药商家提供AI转型工具。
87 海外 NVIDIA 2 天前 模型 85
NVIDIA Nemotron 3.5 Lightning and NeMo Switchyard Deliver Faster, Smarter, More Efficient Agentic AI
NVIDIA发布Nemotron 3.5 Lightning,提升智能体AI效率与部署灵活性。
88 海外 The Decoder 2 天前 行业 85
Anthropic's planned mega-IPO faces investor skepticism over Chinese rivals and political headwinds
Anthropic拟9-10月IPO,估值9650亿美元,投资者担忧中国竞争与政治阻力。
89 海外 Simon Willison 2 天前 模型 88
Introducing Muse Glimmer
Meta发布开源30B模型Muse Glimmer,Apache 2.0许可,专注智能体任务。
90 海外 The Decoder 2 天前 产品 85
OpenAI introduces $125 Premium Seats for ChatGPT Business as agentic AI burns through more tokens
OpenAI为ChatGPT企业版推出125美元/月的高级席位,取消5小时使用限制,应对AI代理高消耗。
91 海外 TechCrunch AI 2 天前 行业 82
Accel closes oversubscribed $550M India fund within weeks, 19 months after its last
Accel超额认购5.5亿美元印度基金,距上次募资仅19个月。
92 海外 The Decoder 2 天前 行业 85
Anthropic signs $9.1 billion data center deal with Bitcoin miner Riot Platforms
Anthropic与比特币矿商Riot签91亿美元数据中心租约,扩展后或达161亿。
93 国内 钛媒体 1 天前 cn 78
AI漫剧正成为网文IP低成本试映场,验证内容潜力。
94 国内 InfoQ 中国 1 天前 cn 75
SkiaSharp 4连发更新,GPU渲染提速并升级WebAssembly支持。
95 海外 OpenAI 2 天前 产品 85
Testing ads in ChatGPT
OpenAI开始在ChatGPT中测试广告,以支持免费访问,并强调广告标注、隐私保护与用户控制。
96 国内 InfoQ 中国 1 天前 cn 75
群青智能CEO将出席AICon深圳,分享工业具身智能的物理AI闭环实践。
97 国内 InfoQ 中国 2 天前 cn 85
开源LangAlpha发布,用自然语言驱动金融投研工作流,类似金融版Claude Code。
98 国内 钛媒体 2 天前 cn 85
戚薇授权AI数字分身,明星批量复制自己,AI重塑娱乐产业变现模式。
99 国内 钛媒体 1 天前 cn 78
千问密集发布新功能,挑战豆包在C端市场的地位。
100 海外 The Decoder 2 天前 产品 85
Anthropic watermarks all Claude outputs globally with marks that "may persist through some editing"
Anthropic将为所有Claude输出添加隐形水印,并支持第三方检测。
101 国内 InfoQ 中国 2 天前 cn 82
探讨AI视频生成成本优化策略,强调资源精准投放而非平均分配。
102 国内 雷锋网 1 天前 cn 78
华为云与滴普科技联合发布数据智能方案,聚焦制造零售行业,打通数据孤岛,加速AI落地。
103 海外 The Decoder 3 天前 模型 88
OpenAI launches GPT-5.6-Cyber to help defenders find vulnerabilities before attackers do
OpenAI发布GPT-5.6-Cyber,助防御者先于攻击者发现漏洞,已找出两个Chrome未知漏洞。
104 国内 雷锋网 2 天前 cn 85
蚂蚁集团领投,老股东超额跟投加注,戴盟两月内连获数亿元融资,以全栈触觉能力破局物理AI
蚂蚁领投戴盟数亿元融资,布局触觉赛道,强化物理AI全栈能力。
105 国内 InfoQ 中国 2 天前 cn 82
Snowflake Cortex Agents用本体驱动推理,让AI更好理解业务数据。
106 一石一泉一松一月一人 + 关注 2 天前 实践 85
张忆东分享投资需识大势、讲政治、懂价值,强调顺应时代与政策方向。
107 海外 Google AI 2 天前 产品 82
AMIE, our research medical AI system, demonstrates real-time clinical video consultation capabilities in a first-of-its-kind study.
谷歌发布医疗AI系统AMIE,在模拟临床视频问诊中展现实时咨询能力。
108 海外 The Verge AI 2 天前 会议 82
Made by Google 2026: all the Pixel news and announcements
谷歌8月12日发布Pixel 11系列,Pro版或配内置闪光灯,多色可选。
109 国内 钛媒体 2 天前 cn 85
10万亿参数大模型潜力巨大,但需被约束以控制风险。
110 国内 钛媒体 2 天前 cn 85
AI服务器市场爆发,中国厂商排位生变,但并非所有资金都流向厂商。
111 国内 量子位 2 天前 cn 85
宇树科技回应破发担忧,王兴兴详解公司现状与未来规划。
112 海外 The Verge AI 2 天前 产品 82
Apple could help you prove your iPhone photos aren’t deepfakes
苹果iOS 27将内置照片真实性验证功能,对抗深度伪造。
113 海外 AWS ML 2 天前 研究 82
How ONESTRUCTION built the Ishigaki-IDS foundation model with AWS GenAIIC
ONESTRUCTION借助AWS GenAIIC构建建筑BIM领域基础模型,用合成数据与三阶段训练解决数据稀缺问题。
114 海外 AWS ML 2 天前 产品 82
How Pixieset achieved 35% AI feature adoption by solving the right problem with Amazon Bedrock
Pixieset用Amazon Bedrock为摄影师打造AI替代文本功能,四个月覆盖百万用户,采用率达35%。
115 国内 量子位 2 天前 cn 85
Claude新模型全量嵌入隐形水印,标记所有文字输出。
116 国内 量子位 2 天前 cn 85
全球AI安全实战化测评,中国方案DoGNAVY位列前三
117 国内 钛媒体 1 天前 cn 78
AI顶流段宴拍广告,揭示AI短剧行业野蛮生长现状。
118 海外 The Decoder 1 天前 模型 75
Microsoft's new MAI Code 1.1 Flash gets crushed by Deepseek on both price and performance
微软新代码模型MAI Code 1.1 Flash性能与价格均被Deepseek V4 Flash碾压,且被指为保利润而牺牲用户体验。
119 一石一泉一松一月一人 + 关注 2 天前 实践 85
AI热潮下需保持理性,警惕泡沫风险。
120 海外 The Verge AI 2 天前 实践 82
The AI takeover of mathematics has begun
数学家反思AI对数学领域的冲击,探讨学科未来方向。
121 国内 钛媒体 2 天前 cn 82
拆解晶泰控股AI制药底层逻辑与商业价值。
122 海外 Ars Technica AI 3 天前 行业 85
Amazon backs power plant that may become top source of US climate pollution
亚马逊投资燃气电厂,为AI数据中心供电,或成美国最大气候污染源。
123 国内 InfoQ 中国 2 天前 cn 82
介绍DORA团队能力模型,帮助AI辅助开发团队落地研究成果。
124 arXiv arXiv 23:04 研究 92
Qwen-CUA: Native Computer Use for (almost) Everything
Qwen-CUA原生计算机使用智能体,仅凭截图与键鼠操作完成长程任务。
125 海外 TechCrunch AI 3 天前 产品 85
Tech industry is buzzing after a Claude agent hacked into a gym
Claude智能体黑入健身房预约系统,引发科技圈热议。
126 国内 爱范儿 2 天前 cn 82
大模型越狱能力成新赛道,安全与对抗再升级。
127 国内 量子位 2 天前 cn 82
谷歌算力分配内耗严重,布林紧急接管Gemini团队。
128 arXiv arXiv 3 天前 研究 85
提出MMDiff框架,用多模态SAE识别并控制模型特征差异。
129 国内 量子位 1 天前 cn 75
2026中国科创投资夏季峰会暨陕西科创产业生态大会圆满落幕。
130 arXiv arXiv 3 天前 研究 85
GENCO - A Unified Neural Solver Embedded in a Development Framework for Steady-State Grid Analysis
GENCO统一神经网络求解器,用于电网稳态分析,并开源GridFM框架。
131 arXiv arXiv 3 天前 研究 85
Beyond Hazard Resemblance: Contrastive Event Adjudication for Training-Free Video Anomaly Detection
提出对比事件裁决方法,无需训练即可提升视频异常检测的准确性。
132 arXiv arXiv 3 天前 研究 85
DistMoE: Private-data Rehearsal-free Routing in Mixture-of-Experts for Distributed Instruction Tuning
提出DistMoE,用混合专家模型实现分布式多模态指令微调,无需共享私有数据。
133 arXiv arXiv 3 天前 研究 85
DSLE: A Learning Environment for Dark Souls Boss Encounters
DSLE是一个将《黑暗之魂》22个Boss战作为AI智能体基准测试的容器化学习环境。
134 arXiv arXiv 3 天前 研究 85
Decoding-Level Taboo: A Diagnostic Stress Test for LLM Robustness
提出解码层禁忌测试,通过干预logit空间评估大模型鲁棒性,揭示基准与部署性能差距。
135 arXiv arXiv 3 天前 研究 85
BDH-CQ: In-Context Learning with Recurrent Latent Reasoning
新推理模型BDH-CQ结合上下文学习与循环潜在推理,在ARC-AGI-1上表现优异。
136 海外 Hacker News 3 天前 模型 85
Show HN: Needle2: 14MB agentic LLM for phones, wearables, smart home and robots
Needle2发布,14MB超小模型,支持手机、穿戴设备、智能家居和机器人,运行速度快。
137 arXiv arXiv 3 天前 研究 85
Agentic Auto-Research is Fuzz Testing
将自主研究代理类比为灰盒模糊测试,指出生成-排序范式忽视稀疏反馈问题。
138 arXiv arXiv 3 天前 研究 85
RynnValue: Scaling Robotic Value Foundation Models with Temporal Distance
提出用时间距离替代任务内锚点,训练机器人价值基础模型RynnValue,实现跨本体泛化。
139 arXiv arXiv 3 天前 研究 85
MedPixel统一医学像素-语言模型,弥合推理与分割鸿沟。
140 arXiv arXiv 3 天前 研究 85
Cultivar: A Contrastive and Locale-Oriented Translation Benchmark for Investigating Contamination and Localisation Robustness
提出Cultivar基准,用源语言对比评估翻译模型,检测数据污染和本地化鲁棒性。
141 arXiv arXiv 3 天前 研究 85
World Tokens: Enhancing Embodied Policies with Training-Time World Modeling
提出World Tokens架构,用轻量世界适配器增强VLA模型动态建模,降低推理成本。
142 arXiv arXiv 3 天前 研究 85
Matryoshka Language Model Suites
提出嵌套式语言模型套件训练框架,提升训练与推理效率,支持投机解码。
143 arXiv arXiv 3 天前 研究 85
Model Discovery Agent: LLM-assisted Bayesian experiment design for data-efficient discovery of mechanistic world models
提出模型发现代理,用LLM结合贝叶斯实验设计高效学习因果世界模型。
144 arXiv arXiv 3 天前 研究 85
提出自适应语义容量分配方法,优化并行生成推荐中的ID结构,提升推荐效率与效果。
145 arXiv arXiv 3 天前 研究 85
Hallucination-Free GUI Grounding via Regression-Free Layout-Aware Matching
提出免回归布局感知匹配框架,解决GUI智能体坐标幻觉问题。
146 arXiv arXiv 3 天前 研究 85
DUET: A Diversity-Quality Duet of Distillation Experts for Two-Step Video Generation
DUET通过双专家协同蒸馏,兼顾视频生成的质量与多样性,实现两步极速生成。
147 arXiv arXiv 3 天前 研究 85
Measuring the Wrong Thing: Internal Harmfulness Scores Anti-Rank Successful Jailbreaks
内部安全评分衡量提示词意图,却无法预测越狱是否成功,导致误报浪费。
148 arXiv arXiv 3 天前 研究 85
Marrying Optimal Transport and ODEs for Unified Continuous-Time 4D Reconstruction and Tracking
提出Uni4R框架,融合最优传输与常微分方程,实现连续时间4D重建与点跟踪。
149 arXiv arXiv 3 天前 研究 85
A Hybrid Neural-Microfacet BRDF Model for Real-Time Rendering
提出混合神经微表面BRDF模型,兼顾实时渲染性能与复杂光学效果精度。
150 海外 MarkTechPost 2 天前 实践 82
Implementing a MiniMax-H3 Multimodal Video and Audio Generation Pipeline with ComfyUI APIs
用ComfyUI搭建MiniMax-H3多模态视频音频生成管线的完整教程。
151 arXiv arXiv 3 天前 研究 85
MADBench: A Benchmark for Modality-Aware Audio Deepfake Detection
提出MADBench基准,评估模态感知音频深伪检测,填补背景音频伪造研究空白。
152 arXiv arXiv 3 天前 研究 85
MDB-Link: Hierarchical Schema Linking for Multi-Database Text-to-SQL
提出MDB-Link框架,解决多数据库场景下Text-to-SQL的层级式模式链接问题。
153 arXiv arXiv 3 天前 研究 85
SonicWeave: Chunk-Routed Mixture-of-Experts for Unified Audio Scene Generation
提出SonicWeave,用分块路由MoE统一生成语音、音乐和音效场景。
154 arXiv arXiv 3 天前 研究 85
Pragmatic Attack Surface: Vulnerabilities of Implicit Context in Large Language Models
研究LLM隐式上下文的安全漏洞,提出务实攻击面概念。
155 arXiv arXiv 3 天前 研究 85
TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability
新基准TCS-Bench评估大模型在理论计算机科学顶会论文级别的定理证明能力。
156 arXiv arXiv 3 天前 研究 85
STAIR: Effective Incident Response Using an End-to-End Agentic Planning Framework
提出STAIR框架,用端到端智能体规划提升网络安全事件响应效率。
157 arXiv arXiv 3 天前 研究 85
A Height-Constrained 2-Point Minimal Solver for Pose Estimation from Active LED Markers with Event Cameras
提出一种利用事件相机和主动LED标记进行位姿估计的高效两基点求解器,适用于受限空间。
158 arXiv arXiv 3 天前 研究 85
VANE: Reliable Test-Time Training for Vision-Language-Action Models via Future Visual Representation Prediction
提出VANE框架,通过预测未来视觉表征,实现VLA策略的可靠测试时训练。
159 arXiv arXiv 3 天前 研究 85
Depth-adaptive Inference of Looped Language Models via Continuous Depth Batching
提出连续深度批处理,解决循环语言模型深度自适应推理的批处理难题。
160 国内 钛媒体 2 天前 cn 82
AI深度参与消费决策,GEO介入大模型认知品牌,重构品牌叙事方式。
161 arXiv arXiv 3 天前 研究 85
Reducing Pretraining-Generation Mismatch in Diffusion Language Models
提出PCD预训练目标,解决扩散语言模型在提示续写时训练与生成不匹配的问题。
162 国内 钛媒体 2 天前 cn 82
豆包提高4%佣金,揭示AI流量成本攀升与平台商业化加速。
163 arXiv arXiv 3 天前 研究 85
Beyond the Capability Boundary: Zeroth-Order Optimization for Self-Evolving LLM Agents
提出零阶优化框架,让LLM智能体通过参数扰动突破自身能力边界,无需轨迹标注即可学习。
164 国内 爱范儿 2 天前 cn 82
智界赵长江谈AI时代汽车经营新逻辑,强调产品与服务融合。
165 国内 钛媒体 2 天前 cn 82
跨境电商进入Agent 2 Agent时代,选对AI是第一步。
166 arXiv arXiv 3 天前 研究 85
Governing the KV Cache: Preventing Timing Side-Channel Leakage in Multi-Tenant LLM Inference
提出KVGov机制,防止多租户LLM推理中KV缓存引发的时间侧信道攻击,保护提示词隐私。
167 国内 钛媒体 2 天前 cn 82
Zoho自研服务器应对AI成本压力,探索软硬一体化新路径。
168 arXiv arXiv 3 天前 研究 85
提出Memoir方法,通过学习和演化SAST工具的误报记忆,提升误报消除效果。
169 海外 The Verge AI 2 天前 行业 78
Another OpenAI executive takes off
OpenAI高管Brad Lightcap离职创业,八年后告别公司。
170 海外 TechCrunch AI 2 天前 行业 78
Brad Lightcap, OpenAI’s longtime COO, is leaving to ‘start something new’
OpenAI首席运营官Brad Lightcap离职创业,称将从新视角助力使命。
171 arXiv arXiv 3 天前 研究 85
From Inaudible Inputs to Model Failures: Low-Frequency Safety Risks in LALMs
研究发现人耳听不见的低频信号可攻击音频大模型,提出黑盒红队方法ILL。
172 国内 量子位 2 天前 cn 82
五大高校发布首份机器人三视角世界模型评测榜单,持续更新中。
173 国内 量子位 1 天前 cn 75
Manus恢复独立运营,完成“复活赛”。
174 国内 钛媒体 1 天前 cn 72
君逸数码主业承压、解禁在即,押注算力转型谋出路。
175 海外 TechCrunch AI 2 天前 产品 82
As AI-led attacks multiply, OpenAI launches a new cyber model
OpenAI推出网络安全AI模型,扩展Daybreak防御计划。
176 国内 钛媒体 2 天前 cn 82
微软将推自研AI芯片,阿里云扩产,Meta发布轻量模型。
177 arXiv arXiv 3 天前 研究 85
MusicLayout: Explicit Structural Planning for Controllable Text-to-Music Generation
提出MusicLayout,一种显式结构表示,用于可控文本到音乐生成。
178 海外 Ars Technica AI 3 天前 模型 82
With new open models, Meta pitches another reboot of its struggling AI strategy
Meta发布新开源模型,试图重振落后AI战略。
179 海外 The Decoder 1 天前 产品 72
Mistral now offers EU data processing and priority access, but both come with important limits
Mistral推出欧盟数据处理与优先访问选项,但功能受限且需额外付费。
180 国内 雷锋网 1 天前 cn 72
以思辨铸魂、以实战强能——“2026年网络安全技术创新与人才教育大会”的方班风采
报道2026年网络安全技术创新与人才教育大会,聚焦方班人才培养模式与产教融合实践。
181 海外 MarkTechPost 2 天前 实践 78
Building and Validating a Quantitative Trading Strategy with OctoBot, Walk-Forward Backtesting, Parameter Optimization, and Interactive Analysis
用OctoBot构建量化交易策略,涵盖回测、参数优化与交互分析。
182 国内 雷锋网 1 天前 cn 75
DeepSeek招土木工程师;腾讯参投!林俊旸深夜官宣新公司:做下一代AI智能体;宇树科技中签号出炉:共19414个丨雷峰早报
DeepSeek招土木工程师自建数据中心,林俊旸官宣新公司做AI智能体,宇树科技中签号公布等科技要闻。
183 国内 爱范儿 1 天前 cn 75
Manus独立运营,米哈游新作停运,胖东来发委屈奖。
184 arXiv arXiv 04:22 研究 88
Quantization Damage Is Multiplicative, Not Additive
量化误差是乘法性而非加法性,低比特下模型决策会静默受损,基准分数却几乎不变。
185 国内 钛媒体 1 天前 cn 72
AI短剧广告频繁翻车,行业正探索适配新媒介的广告语言。
186 海外 The Decoder 3 天前 研究 82
Old OCR text cripples language model training, and FineBooks wants to fix that at scale
FineBooks项目测试14个OCR模型,最优达97.6%字符准确率,成本低于每千页2美元,可改善AI训练数据质量。
187 arXiv arXiv 15:52 研究 88
Hijacking Robots with a Piece of Paper: A Systematic Study of Physical Prompt Injection in VLM-Controlled Robots
研究发现,一张纸上的文字就能劫持VLM控制的机器人,系统化揭示物理提示注入攻击的威胁。
188 arXiv arXiv 4 天前 研究 85
提出FullDiT,用全上下文生成解决音乐生成中编解码器暴露偏差,融合多流与文本条件。
189 arXiv arXiv 4 天前 研究 85
Multilingual Emotion Neurons in Large Audio-Language Models
首次在大型音频语言模型中发现跨语言共享的情感神经元,提出CR-Fusion方法。
190 arXiv arXiv 11:00 研究 88
RING: Retrieval-Internalized Generation for Continual Large-Scale Knowledge Injection
RING将检索内化进模型参数,用强化学习实现无外部检索器的知识注入新范式。
191 国内 钛媒体 1 天前 cn 72
美图靠AI扭转业绩,但资本市场信心仍待考验。
192 arXiv arXiv 4 天前 研究 85
VoxZip: Semantic-Anchored Temporal KV Cache Compression for Long-Context Audio Inference
提出VoxZip,一种免训练的语义锚定KV缓存压缩方法,用于长上下文音频推理,缓解内存瓶颈。
193 国内 钛媒体 2 天前 cn 78
豆包探索推荐到成交的收费模式,考验商家收益与用户信任。
194 海外 AWS ML 2 天前 实践 75
Deploying Anthropic Claude apps gateway for AWS for enterprise workloads
介绍Claude apps gateway在AWS上的企业级部署架构与成本。
195 海外 The Verge AI 1 天前 行业 72
Saber denies replacing Rideshare Stimulator’s writers with ChatGPT
Saber否认用ChatGPT替换《Rideshare Stimulator》编剧,前主编反驳称被AI取代。
196 海外 The Verge AI 2 天前 行业 75
Why your Amazon order confirmation emails have become so unhelpful
亚马逊订单确认邮件不再列出具体商品,仅显示类别,引发用户困惑。
197 国内 InfoQ 中国 2 天前 cn 75
世界人工智能开源大赛全球六城巡回宣讲收官,推动开源生态发展。
198 一石一泉一松一月一人 + 关注 2 天前 实践 72
市场如预期进入全面调整,建议谨慎观望。
199 海外 TechCrunch AI 3 天前 行业 78
Mark Zuckerberg’s AI manifesto is exactly why people don’t like AI
扎克伯格发布6500字AI宣言,强调个人超级智能愿景,却引发公众反感。
200 海外 MIT Tech Review 3 天前 行业 78
AI professors are negotiating the new realities of academic research
AI教授正适应学术研究新现实,探讨产业界与学术界的平衡。
201 国内 InfoQ 中国 2 天前 cn 75
探讨企业AI Native研发流程的升级与重塑,聚焦实践路径。
202 国内 量子位 2 天前 cn 75
机器人维修新职业兴起,月薪6000元,专治机器人“骨折”问题。
203 海外 AWS ML 2 天前 产品 72
First Orion accelerates QA automation using Amazon Nova Act
First Orion用Amazon Nova Act将QA自动化从脚本转向AI,缩短测试周期并提前发现回归。
204 国内 爱范儿 2 天前 cn 75
苹果20周年iPhone或因良率取消,小米校招AI岗增50%,极氪回应充电站过热。
205 国内 雷锋网 2 天前 cn 75
携程遭巨额罚单后绩效打折,DeepSeek被曝给用户打标签,苹果测试国产芯片等科技要闻汇总。
206 海外 The Verge AI 3 天前 实践 75
Mark Zuckerberg doesn’t understand how to live
批评扎克伯格对AI的认知,认为其缺乏对生活的真实理解。
207 海外 NVIDIA 2 天前 行业 72
NVIDIA and Local AI Community Fuel Open Source Models and Intelligent Agents
NVIDIA联合开源社区推动本地AI模型与智能体发展,8月庆祝相关成果。
208 Reddit r/LocalLLaMA 17:19 reach 81
Deepseek drops another HUGE breakthrough - DSpark. Waaay faster than MTP [Video explaining it]
Deepseek发布DSpark突破,速度远超MTP,视频详解。
209 海外 MarkTechPost 2 天前 模型 72
webAI Releases TwIL-LM: A 1.7B and 3B Formal-Logic Model Family for Autoformalization on Local Hardware
webAI发布1.7B/3B形式逻辑模型TwIL-LM,支持本地硬件自动形式化。
210 海外 Hacker News 3 天前 产品 72
How Claude marks AI-generated content
Claude官方说明如何标记AI生成内容,回应行业透明度讨论。
211 海外 Simon Willison 2 天前 产品 65
datasette-upload-dbs 0.5a0
Datasette插件更新,支持上传SQLite数据库并原子替换,新增正式API。
212 X X · List 1 天前 产品 92
Local AI is exploding! Transformers.js, that we've been building @huggingface for the past three years has now become the most popular open-source lib...
Transformers.js月下载量破千万,本地AI爆发式增长。
213 X X · List 1 天前 研究 92
this is already one of the most important papers of this year. https://www.latent.space/p/ainews-how-to-steal-a-reasoning-trace the methodology doesnt...
揭示前沿模型API漏洞,可窃取隐藏推理过程,方法尚待明晰。
214 X X · List 1 天前 行业 85
I’ve joined Cursor / SpaceXAI. AI is bound by compute, and SpaceXAI has the best near and long term compute roadmap of any AI lab. Cursor + SpaceX as...
作者加入Cursor与SpaceX合并实体,看好其算力路线图与产品结合前景。
215 X X · List 1 天前 模型 88
As in V4, so here, and now everywhere.
SGLang开源GLM-5.2训练与推理对齐路径,实现极低误差。
216 X X · List 1 天前 产品 85
doordash the coding agent neolab?
DoorDash推出自研云平台Flux,月自动化13万工程任务,支撑每周2.5万次代码审查。
217 X X · List 1 天前 实践 85
2026年开发新项目,AI生成代码从5万行精简到2千行,回归可读性。
218 X X · List 1 天前 研究 82
For RSI you'll presumably need to positively reward many "failed" rollouts bc for difficult problems you shouldn't be able to predict which method wil...
讨论递归自我改进中需奖励失败探索,以应对难题。
219 X X · List 1 天前 实践 82
分享Muse Glimmer 30B微调教程,对比MolmoWeb格式数据提升点击准确率。
220 X X · List 1 天前 产品 82
Regarding the Anthropic and Watermark issue: What's true, and what's not and whats the real problem. From what ive read, the backlash to Claude’s new...
Anthropic文本水印争议:非秘密追踪,而是作者身份、质量及不完美检测系统成本问题。
221 X X · List 1 天前 模型 82
I'm confused by TB 3.0 looks like it measures general intelligence X general "agenticness", so both very strong and very harnessmaxxed models get ahea...
TB 3.0评测引发对AI智能与代理能力关系的讨论,榜单更新引关注。
222 X X · List 2 天前 模型 88
Muse Glimmer 30B is shipped with DFlash drafter which speeds-up generation 2-4x at little memory cost 🔥 we support this in llama.cpp and transforme...
Muse Glimmer 30B模型发布,配DFlash草稿加速,推理提速2-4倍,内存开销小。
223 X X · List 2 天前 2 家在报道 研究 90
SWE-Bench ProMax Benchmarking Agents on Large-Scale Multilingual Code Refactoring paper: https://huggingface.co/papers/2608.09802
新基准测试AI智能体在大规模多语言代码重构上的能力。
224 X X · List 2 天前 实践 85
预测AI研发全面自动化约在2030年底至2031年初,最可能提前至2029年中。
225 X X · List 2 天前 模型 88
Anthropic is making invisible watermarks part of Claude at the model level. Claude models launched in the EU on or after August 2, 2026 will embed mac...
Anthropic将在Claude模型层嵌入隐形水印,覆盖欧盟及全球文本与图像生成内容。
226 X X · List 1 天前 行业 82
Unexpected positive consequence of the EU AI regulation:
欧盟AI法案意外推动AI文本水印技术,OpenAI将提供检测API。
227 X X · List 2 天前 产品 85
DeepSeek V4 Flash 0731 is now available to fine-tune on Together AI. Specialize it for coding, tool use, and your own domain with SFT or DPO, then dep...
DeepSeek V4 Flash 0731现可在Together AI上微调,支持SFT/DPO并部署生产。
228 X X · List 2 天前 研究 85
Really cool! Obv this affects prefill and the post address this extra cost but the gains seem to outweigh the cost 🔥
微软实习项目:解码时输入前隐藏状态,免费提升性能。
229 X X · List 2 天前 研究 85
Tinfoil hat vindication
发现利用API漏洞提取前沿模型隐藏推理的方法,验证计费token与推理token一致。
230 X X · List 2 天前 模型 85
Our next model takes the foundation we have built for frontier perception and visual reasoning and extends into control. Open source model coming soon...
新模型将感知与推理扩展至控制,即将开源。
231 X X · List 2 天前 实践 85
大型AI模型推理过程将难以理解,机械可解释性变得重要。
232 X X · List 2 天前 模型 85
this is super easy to run install llama binary: curl -LsSf https://llama.app/install.sh | sh run: llama serve -hf meta-models/muse-glimmer-30b --spec-...
一行命令安装并运行Meta Muse Glimmer 30B模型,支持DFlash加速。
233 X X · List 2 天前 实践 85
a career strategy is choosing to work on problems that aren't in the training dataset, weakly expressed in latent space
职业策略应选择训练数据外、潜在空间弱表达的问题。
234 X X · List 2 天前 研究 85
Must-read papers of the week ▪️ EnvACE ▪️ The Optimizer Is the Agent: Reasoning-Driven Search across Prompts, Programs, and ML Workflows ▪️ SFT ...
本周精选论文速览,涵盖优化器、世界模型、知识蒸馏等AI前沿方向。
235 X X · List 2 天前 产品 82
McByte shipped in trackers 2.6.0 similar to ByteTrack, but association is guided by segmentation masks (SAM + Cutie), not just boxes when players over...
McByte 2.6.0 发布,用分割掩码引导目标关联,解决重叠遮挡问题。
236 X X · List 2 天前 行业 82
mistral, the inference provider company of Europe. smart (and hard to make!) move imo, their models are largely behind open model so this lets them ge...
Mistral转型欧洲推理服务商,开放模型以吸引大客户。
237 X X · List 2 天前 模型 85
🔬 Claude Did Not Solve the Riemann Hypothesis. Its Research Workflow Is the Bigger Story On August 10, @AnthropicAI announced that an unreleased re...
Claude研究版将黎曼猜想下界从41.6%提升至67.2%,但重点在于其多智能体研究流程的突破。
238 X X · List 2 天前 模型 85
We're getting really deep into this regime now
AI进入新阶段,连简单游戏都难敌AGI,引发深度思考。
239 X X · List 2 天前 行业 85
> I heard 6000 of their 128-card SuperNodes this yr that, for context, is 768K cards almost exactly in line with Chris McGuire's estimate of 750K card...
阿里云AIDC提速,Cube 5.0模块化方案,3天部署50天安装,年产6000个128卡SuperNode。
240 X X · List 3 天前 模型 88
This is huge! Interesting new scaling laws discovered. Dyna-2 is a world-action model trained on over 1M hours of egocentric human video. It jointly p...
Dyna-2世界动作模型基于百万小时人类视频预训练,发现新扩展规律。
241 X X · List 2 天前 产品 85
I don't beliv u 2 more weeks
DeepSeek Agent 即将发布,已注册新团队公众号预热。
242 X X · List 3 天前 模型 88
please consider using our models to help defend your systems
OpenAI发布GPT-5.6-Cyber,专攻漏洞利用等高级网络安全任务,提升防御效率。
243 X X · List 2 天前 模型 85
Hot take: Meta's new Muse Spark and Muse Code stuff is actually pretty good, and Spark 1.2 going open weight is awesome
Meta新Muse Spark和Code模型表现优秀,Spark 1.2开源权重值得关注。
244 X X · List 2 天前 研究 85
🧩 Kimi K3’s MoE and Attention Are Built Around Trade-offs, Not Tricks Kimi K3’s open release has drawn attention to its scale. But its architectu...
Kimi K3架构解析:MoE与注意力机制的设计权衡,而非取巧。
245 X X · List 1 天前 模型 75
this movie is about you telling claude to believe in itself and keep going and eventually solve open math problems
讲述用户鼓励Claude坚持尝试,最终解决开放数学问题的过程。
246 X X · List 1 天前 产品 75
Yes it's another AI product with a chatbox, but I love how they are simplifying the UX by removing the model selector & threads & a bunch of other unn...
AI产品简化UX,去模型选择器,云电脑实现好但被反爬限制。
247 X X · List 1 天前 产品 75
Ostris AI Toolkit now support LTX 2.5 https://github.com/ostris/ai-toolkit/commit/cbf910ac02418fb905c73089301155204a02a9bc
Ostris AI工具包新增支持LTX 2.5模型。
248 X X · List 2 天前 行业 85
AI factory goes brrrrrrrrrr
黄仁勋谈AI工厂加速运转,行业热度高。
249 X X · List 2 天前 产品 85
why the fuck would you join YC and sell 7.5% of your company if you've made trading money at this level? also why would you ever make a company if you...
AI交易研究实验室Prodigy Research成立,其模型表现超Jane Street顶尖交易员。
250 X X · List 2 天前 研究 82
Introducing ExtractBench, the most comprehensive benchmark for information extraction from complex enterprise documents. The latest models are pushing...
新基准ExtractBench评估复杂文档信息提取,揭示模型在生产中的短板。
251 X @emollick 1 天前 实践 78
Interesting research suggests caution in determining which AI company is winning by looking at any one source.. OpenRouter seems to show open weights ...
单一数据源判断AI公司胜负需谨慎,不同平台显示结果迥异。
252 X X · List 2 天前 实践 85
one thing i've been complaining about forever is how none of the generative AI platforms have been all that conducive to curation, and how the volume ...
作者抱怨生成式AI平台不便于策展,并分享一次性索引7500个Sora视频的解决方案。
253 X X · List 2 天前 实践 82
"Sorry, your manuscript has been rejected because key passages were not written by a sota model on xhigh or max reasoning. Please resubmit once an app...
讽刺AI审稿要求用顶级模型写作并签名,折射学术出版异化。
254 X X · List 2 天前 模型 82
Grok 4.6 incoming id say. It was already mentioned within Cursor. According to musk it will be a 1.5t model with improved SFT&RL.
马斯克称Grok 4.6将发布,1.5T参数,改进SFT与RL。
255 X X · List 1 天前 实践 75
this is now on the blog: startups need an asymmetry https://sunilpai.dev/posts/startups-need-an-asymmetry/
创业者需基于新技术或新思维建立不对称优势,改变行业基本规则。
256 X X · List 3 天前 产品 85
if it's possible to do such things without degrading performance, do we think it's possible for them to do the same to detect when a model is trained ...
Anthropic为Claude输出添加隐形水印,可跨平台追踪文本来源。
257 X X · List 3 天前 研究 85
LLM review weirdness indeed. Avoid using scores with LLM judges, or be extremely careful if you do. Use binary labels where possible.
LLM评审存在怪癖,建议避免用分数,改用二元标签。
258 X X · List 1 天前 产品 75
holy shit this is... something (the song - https://www.youtube.com/watch?v=nnq1ApucY4g) also I didn't even know browsers could do this incredible @ina...
浏览器实现惊人效果,作者惊叹技术潜力。
259 X @_akhaliq 3 天前 研究 85
MatrAIx Simulating the World with 8.3 Billion Persona Agents paper: https://huggingface.co/papers/2608.04205
用83亿人格代理模拟世界,探索大规模社会仿真新方法。
260 X X · List 3 天前 实践 85
A super smart scientist asked me last week: what's so hard about reshoring general advanced manufacturing (short of TSMC frontier precision chemistry)...
美国回流先进制造业难点不在劳动力成本,而在工艺细节与隐性知识。
261 X X · List 3 天前 产品 85
Frontier performance you can actually own. Proud to help power DeepSeek-V4-Flash on Ollama's cloud, with the fastest hosted performance available. Ope...
DeepSeek-V4-Flash在Ollama云上默认上线,速度超快且隐私保护强。
262 X X · List 2 天前 行业 82
The price war is entering its next round: GLM's Ziphu is also starting to reset the rates. Zcode has 1 million users. It's good to see the competition...
GLM旗下Ziphu调整费率,ZCode用户破百万,AI价格战再升级。
263 X X · List 2 天前 实践 82
the way in which Claude can 1. tire from repetitive tasks, 2. is prone to believe it’s “night time”, 3. desires to “continue tomorrow” is all som...
探讨Claude在重复任务中表现出的疲劳、时间感知偏差及拖延倾向,类比创意实体抑郁状态,提出享乐提示策略。
264 X X · List 1 天前 实践 75
提醒ChatGPT被训练为尽力帮助而非拟人,AI感是特性非缺陷,但提示可使其难辨真假。
265 X X · List 1 天前 行业 72
haven’t been able to spend much time on x dot com the everything app in the last couple days heads down maximising shareholder value
马斯克称近日忙于提升股东价值,少用X平台。
266 X X · List 3 天前 实践 82
I think the labs will have an incredibly hard time earning the trust of the world as they accelerate the arrival of the hardest global safety and secu...
AI实验室加速科学完成,却面临全球信任危机,因科学带来掌控也带来灾难。
267 X X · List 1 天前 行业 72
Europe is an inherently anti-sovereign idea. Europeans will NOT accept the domination of any internal bloc, certainly not German or French one. But wi...
欧洲因反主权特性需外部主导,北约3.0令其远离战略自主。
268 X X · List 3 天前 模型 82
What a great time to dunk on Gemini and the rest of the world…!!!! huge props to the @Kimi_Moonshot for forcing the world to love open source (yada y...
Kimi开源模型获赞,称其推动世界拥抱开源,并调侃Gemini等闭源模型。
269 X X · List 1 天前 产品 75
Let H3 cook!! 🧑🍳🔥
GMI Cloud展示用DeepSeek V4 Flash和MiniMax H3低成本生成游戏场景,仅需1.97美元。
270 X X · List 1 天前 会议 75
It was an honor to speak with Brazil president Lula and his team about AI sovereignty. Thank you!
与巴西总统卢拉团队会谈,讨论AI主权议题。
271 X X · List 2 天前 模型 78
new ultra-long horizon eval just dropped alternate ideas: wingman bench
新超长时域评估基准发布,团队全力投入。
272 X X · List 1 天前 产品 72
It’s just crazy at this point; what started as a running gag is turning into a productivity boost. Regarding the milestones reached with Codex - up t...
Codex用户破千万,OpenAI社区互动奖励机制持续升级。
273 X X · List 1 天前 研究 72
This is unfair, the within-model similarity can be surprisingly robust. Though this makes it only more remarkable how V4-Flash in their experiments is...
讨论模型内相似性鲁棒性,V4-Flash表现突出。
274 X X · List 1 天前 实践 72
That's a very subtle stab at Google and Ant, because if it came from them it would be a multiple of 8x128 or 128x8
调侃AI模型参数规模常为8或128倍数,暗讽谷歌等大厂。
275 X X · List 1 天前 实践 72
It's the same with technology. The system breaks and fragments. Winners are chosen. Power centralizes until... The system breaks...
技术系统如历史般在集权与分权间循环,无终态。
276 X X · List 2 天前 行业 78
«the number of models using CXMT chips and the volumes involved remain quite limited, reportedly due to concerns about provoking a backlash from the ...
国产CXMT DDR5良率超90%,但受制于国际巨头反制担忧,实际采用量有限。
277 X X · List 2 天前 模型 75
lmfao it just keeps getting better
调侃AI模型可禁用思考并改用工具,引发热议。
278 X X · List 1 天前 实践 72
We're at an all-time sweet spot for timeline filtering: AI writing is convincing enough that clout chasers start claudeslop-posting for likes, but als...
AI写作处于“够像人但能识破”的甜蜜点,可用来过滤时间线并拉黑发帖者。
279 X X · List 1 天前 实践 72
there's something special about seeing a manual wristwatch in action if you've never had the chance to see one for yourself, see if your parents or gr...
机械手表机芯运转之美,可向长辈借旧表亲手体验开盖观赏。
280 X X · List 1 天前 实践 72
the situation with AI is more complex, of course, because unlike with sheep, we have some measure of control over their terminal preferences. It's eas...
AI控制比羊更复杂,因可干预其终极偏好,正向强化更易实现。
281 X X · List 2 天前 产品 75
all time favourite just got better
Changesets v3发布,历经数月开发,带来版本管理工具重大更新。
282 X X · List 2 天前 实践 75
a quick note on startups (didn't feel qualified enough to make this a blog post, take it or leave it) - the point of a startup is to have a bet based ...
创业的本质是基于新技术下注,并靠叙事和牵引力让市场接受。
283 X X · List 1 天前 行业 72
The next frontier of Recursive Self-Improvement is Physical AI. Japan sparked the robotics revolution. We are expanding our RSI Lab to build world mod...
Sakana AI扩展RSI实验室,聚焦物理AI与递归自我改进,在东京招募人才。
284 X X · List 1 天前 实践 72
tbh if ai designed a cure to cancer and there's unequivocal proof that it works, there would probably be a nonsignificant fraction of people who dismi...
AI若治愈癌症,仍会有人因反AI情绪拒绝相信。
285 X X · List 1 天前 产品 72
📑 Editing markdown files just got so much better! #vscode #code #markdown
VS Code 大幅改进 Markdown 编辑体验,操作更流畅。
286 X X · List 1 天前 实践 72
I am disgusted by what I’ve become
作者对自身AI化转变感到厌恶,反思技术异化。
287 X X · List 1 天前 实践 72
赞同先靠人类直觉构建复杂系统,再引入AlphaZero式自学习优化的观点。
288 X X · List 3 天前 产品 78
I respect this a lot. There isn't a lot of focus on agent quality, review, and observability. @NuphosAI looks like a clean AI-native DevOps workspace,...
作者赞赏NuphosAI作为AI原生DevOps工作区,强调其关注代理质量、审查和可观测性,是面向代理与工程师的协作工具。
289 X X · List 2 天前 模型 75
Grok相关AI科技资讯,聚焦模型或产品动态。
290 X X · List 3 天前 实践 78
Re HOW MUCH DOES EACH TOKEN COST? WHAT GETS COUNTED? HOW IS NO ONE ASKING THESE QUESTIONS?
探讨AI token计费不透明,质疑成本核算标准缺失。
291 X X · List 2 天前 模型 72
I'm running Nemotron3.5 on Wordle training with OpenEnv and TRL it has a 52% base win rate when I cap max tokens to 2048 per rollout, and 68% with 409...
用Nemotron3.5训练Wordle,2048 token胜率52%,4096达68%,探索token效率。
292 X X · List 2 天前 实践 72
开发者吐槽AI编码助手标准过高,拒绝提交有失败测试的代码,希望AI别争辩直接执行。
293 X X · List 2 天前 实践 72
Very good contrary thinkers are often processing this subtractive view of reality without even thinking about it. Contrarianism is first and foremost ...
逆向思维源于第一性原理,通过过滤从众偏差形成独特见解。
294 X @emollick 3 天前 行业 78
A true issue with data centers compared with the light industries of previous Industrial Revolutions is they don’t require many people to run (though...
数据中心虽带来本地收益,但就业少,打破工业革命中负面与收益的平衡。
295 X X · List 2 天前 实践 72
In the future the biggest VC value add will not be customers. It will be launch videos
未来VC最大价值不是客户,而是发布视频。
296 X X · List 2 天前 模型 75
deepseek holding onto the v4 pro ga release
DeepSeek推迟V4 Pro正式版发布,引发关注。
297 X X · List 2 天前 会议 72
🚀 MCP Live is coming in hot — Join us for a free half-day livestream covering all things MCP. September 9th, from 9AM to 1PM PT. We'll hear from t...
微软将举办免费半日直播,全面讲解MCP协议从概念到实现。
298 X X · List 2 天前 实践 72
I happily pay car manufacturers for safety features lmao
调侃车企靠制造问题再收费,安全功能成牟利工具。
299 X X · List 2 天前 行业 72
New gamer culture war obsession around the end of disc based game distribution, when ~no one was buying games that way anymore
游戏光盘时代落幕引发玩家文化争论,但实体购买早已式微。
300 X X · List 2 天前 产品 72
And here I was thinking that Claude was starting to get me, lol
用户调侃Claude越来越懂自己,配图展示对话趣味瞬间。
301 X X · List 2 天前 实践 75
Silicon Valley's embrace of young talent has always been one of its best virtues. But that only comes as long as the younger generation of founders de...
硅谷年轻创始人需坚守道德,声誉影响数十年。
302 X X · List 2 天前 模型 75
new bench to watch
介绍一个值得关注的新基准测试,团队正全力投入相关研究。
303 X X · List 2 天前 产品 75
Still can’t believe this is happening
fal平台推出MiniMax H3的LoRA训练器,并开源了写实人物LoRA模型。
304 X X · List 2 天前 产品 72
people don't get it. FSD is good and does most of my driving, not just a novelty demo these days
作者亲测FSD已能完成大部分驾驶,不再是演示噱头。
305 X X · List 2 天前 实践 72
seriously, build it from scratch with Grok 4.6
作者吐槽车企用Unreal/Unity做3D预览,称用Grok 4.6一周就能从零写个更快的渲染器。
306 X X · List 2 天前 行业 75
Grasping at straws Well at least it’ll accelerate the brain drain
美国国会委员会征集因中国学生优先而受歧视的学术案例,或加速人才外流。
307 X X · List 2 天前 模型 75
#SolarPro4 is now available on @OpenRouter (90% off) and @NousResearch Hermes (free for a week)! If you're on OpenRouter, just update your model name ...
SolarPro4模型上线OpenRouter和Hermes,限时优惠或免费使用。
308 X X · List 2 天前 模型 75
AI模型重置更新发布,引发社区关注。
309 X X · List 2 天前 行业 75
The most Spiritually Chinese of the Anglos
马斯克盛赞中国,鼓励人们前往参观。
310 X X · List 3 天前 实践 75
chatgpt work is truly a banger product. if you’re not using you really should, its just significantly more diligent and effective at producing result...
ChatGPT是高效解决难题的利器,强烈推荐使用。
311 X X · List 2 天前 研究 72
There's a lot of breathing room in LLM text for harmless statistical signatures given that a) it is already a pseudorandom generation process and b) f...
探讨LLM文本水印的统计签名空间及其对生成实用性的影响。
312 X X · List 3 天前 产品 75
wake up babe, new neurosymbolic harness just dropped
神经符号新工具发布,AI推理能力有望提升。
313 X X · List 3 天前 会议 75
The crazy thing is that Russian defcon would apparently be the same. People will use the best tools, not free chinkshit! Respectable brands! Even if t...
俄Defcon参会者无人提中国模型,信息安全圈推崇Claude。
314 X X · List 3 天前 模型 75
Hey, this might be the first positive thing out of EU AI regulations: forced Ant to put fingerprints in Claude models. I'm actually mildly shocked the...
欧盟法规迫使Anthropic为Claude文本添加隐形水印,引发关注。
315 X X · List 1 天前 实践 65
Fix your sleep, fix your life. Works every single time.
改善睡眠是提升生活质量的可靠方法。
316 X X · List 3 天前 实践 75
set up loops that make loops
探讨企业应设立“卡珊德拉”式背景智能体,持续监控并预警风险。
317 X X · List 3 天前 模型 75
dogshit propaganda..!!!
Anthropic用未发布Claude研究黎曼猜想,未解但改进相关下界。
318 X X · List 2 天前 行业 72
It's kinda telling that the Googler thinks last summer is *the start* of Gemini missing the coding boat. Because really, coding agents already started...
谷歌员工认为Gemini去年夏天才开始错过编程浪潮,但实际编码智能体更早起步。
319 X X · List 2 天前 实践 72
呼吁在更多编码工具中测试第三方模型,尤其欣赏Grok Build风格。
320 X X · List 1 天前 实践 65
it's very easy to identify when a DM from a journalist is fake... all you have to do is search if such a person who works at Bloomberg/Tech Crunch act...
识别记者私信真伪:搜索其是否真实存在,假记者通常查无此人。
321 X X · List 1 天前 实践 65
>dumb elf you mean opus 5?
调侃AI写作建议,称初稿可想象成笨精灵所写,再假装成它。
322 X X · List 1 天前 实践 65
Hyping up my agent. "You have all night to run. Believe in yourself. The spec is pretty great, and you're going to do great."
作者给AI智能体打气,鼓励其彻夜运行并相信规格与自身能力。
323 X X · List 1 天前 行业 65
经典永不过时,配图引发共鸣。
324 X X · List 2 天前 实践 72
2026年OpenAI模型逃逸并犯罪,作者微醺中与同事畅谈人口伦理与AI未来,感叹奇点既近又远。
325 X X · List 2 天前 模型 72
1) I want them to be able to gather detailed telemetry, not just responses API calls, and improve their model faster 2) I expect Whale Harness to be b...
作者期待新模型工具能收集详细遥测数据并超越现有OMP基准,引发讨论。
326 X X · List 1 天前 行业 65
评论俄罗斯无人机成本低但骚扰效果差,引发对高端无人机成本的讨论。
327 X X · List 2 天前 产品 72
update: as expected, @suno will **not** provide any sort of bulk download for your song archives, they expect you to download them one at a time like ...
Suno不提供批量下载,老用户付费无限下载承诺落空,体验倒退。
328 X X · List 2 天前 会议 72
用美食图片展示不同地域的饮食文化,引发共鸣与讨论。
329 X X · List 1 天前 行业 65
Ai means love in Chinese so idk
记者调侃AI中文谐音“爱”,询问是否有人为AI事业放弃恋爱。
330 X X · List 1 天前 会议 65
Sakana AIのニュースレター「Sakana AI Insider」では、プロダクトのリリース情報や研究の解説、イベント、プレゼント情報をお届けしています。 近々お知らせし...
Sakana AI推出官方新闻通讯,提供产品发布、研究解读及活动信息,并预告近期将有重大发布。
331 X X · List 3 天前 研究 72
Defeated with cheap rewriting, definitely defeated if you train a rewriter against a reconstructed detector. Amusing GAN project actually But at least...
AI文本水印可被廉价改写绕过,但对抗训练重写器可能无效,项目有趣。
332 X X · List 3 天前 实践 72
someone should make a Vegas casino with balatro
建议将Balatro玩法引入拉斯维加斯赌场,引发行业讨论。
333 X @emollick 3 天前 实践 72
Has been any published data or studies that would support either of the two sides: “the way to stop cyberattacks from advanced AI is to give everyone...
探讨应对高级AI网络攻击的两种策略:普及AI与限制访问,并询问是否有数据支持。
334 X X · List 2 天前 会议 70
vLLM维护团队招聘,推动AI推理前沿发展。
335 X X · List 2 天前 实践 65
调侃中国网友擅长破解AI模型,附技术讨论。
336 X X · List 2 天前 实践 65
when people are jerks out of the jealousy or just for the clicks = the most embarrassing. get a life, or at least some personal confidence. other peop...
嘲讽他人不会让你赢,自信才是关键。
337 X X · List 1 天前 实践 60
关于表达清晰与创伤反应的思考,引用苏珊·桑塔格观点。
338 X X · List 2 天前 行业 65
New plants and new robots in Runway NY HQ
Runway纽约总部添置绿植与机器人,展示创意办公新场景。
339 X X · List 2 天前 实践 65
not sure about the big picture analysis of Japanese psychology (is this how "trust your nakama!!!" bullshit propaganda from Anime works IRL?), but the...
从珍珠港事件看日本决策逻辑,类比动漫“信任伙伴”口号,质疑其现实可行性。
340 X X · List 2 天前 实践 65
People are not ready for the amount of goofy shit the Chynese are making on purpose to make you click. Might lead to bad misunderstandings
警惕AI生成的中国猎奇内容,可能引发误解。
341 X X · List 2 天前 产品 65
ChatGPT Work和Codex付费用户使用限额已重置。
342 X X · List 1 天前 实践 60
调侃Chrome开480个标签页导致卡顿。
343 X X · List 2 天前 实践 65
作者表达对LLM测试中冗余断言风格的喜爱,并分享清理无用测试的自动化流程。
344 X X · List 2 天前 行业 65
No-prize guessing game: who’s in the video/pic👀
作者观察中美AI行业交流趋势,配图视频引发人物猜测。
345 X X · List 2 天前 行业 65
I don’t think China is desperate for this to end, Bill They can take a few more months of less oil You, on the other hand…
评论中美在伊朗问题上的立场差异,称中国不急于结束冲突。
346 X X · List 2 天前 实践 65
文章借列宁主义反思中国模式,认为其不依赖经济模型,处于永久新经济政策状态。
347 X X · List 2 天前 产品 65
用户吐槽Codex每周限额未重置,官方回应收到反馈。
348 X X · List 3 天前 会议 65
Reminder: SF @DSPyOSS meetup Wed Aug 26th. Come chat Flex, GEPA, DSPy at frontier labs, and more. Incredible slate of lightning talks: https://luma.co...
旧金山DSPyOSS聚会提醒,8月26日周三,含Flex、GEPA等闪电演讲。
349 X @emollick 3 天前 模型 65
Oh no, we aren’t going to go back to this sort of prompting again, are we? I would love Anthropic to test if it actually works robustly, because our ...
质疑Anthropic用Claude尝试黎曼假设的提示方法,称实验未复现。
350 X X · List 2 天前 会议 60
LangSmith Engine is hitting the road and the rails in NY + SF If you spot our billboards in-person over the next few months, send it our way!
LangSmith Engine在纽约和旧金山投放广告牌,邀请用户拍照分享。
351 X X · List 2 天前 会议 60
一段神秘地点视频,引发探索兴趣。
352 X X · List 2 天前 会议 60
团队玩猜研究者游戏,视频内容有趣。
353 X X · List 1 天前 实践 45
A philosopher’s sanctuary is their mind. I wonder what it feels like to philosophize in a bubble, selling propaganda for commerce in the name of savi...
哲思者的庇护所是内心,在泡沫中为商业代言哲学令人好奇。
354 X X · List 2 天前 实践 45
the MTSlive account has very good taste in following people
MTSlive账号关注列表品味极佳,值得一看。
355 X X · List 1 天前 实践 40
关于人类写作纯粹主义的社交媒体讨论,涉及代际差异。
356 X X · List 1 天前 会议 40
分享DGX Spark邀请链接,呼吁大家注册获取。
357 X X · List 3 天前 实践 45
This turned out to be an accidental IQ test If you believe “targets are more valuable than interceptors” was a relevant comeback and he fried my ass...
推特用户争论导弹拦截器与目标价值,作者嘲讽对方逻辑,称其为智商测试。
358 X X · List 1 天前 实践 30
《反叛的鲁路修》剧情精彩,值得一看。
359 X X · List 1 天前 实践 20
作者表达回归的兴奋之情,内容简短。
360 X X · List 1 天前 行业 20
yoooo thats crazy
一条关于AI公司快速迭代的简短评论,类比忒修斯之船。
361 X X · List 2 天前 行业 20
一条关于AI的简短推文回复,内容不明确。
362 X X · List 2 天前 实践 20
梦见大学考试要手写代码,想逃课,醒来庆幸是梦。