1 国内 量子位 11 小时前 cn 92
Claude破解哈达玛矩阵猜想,AI数学能力再突破。
2 国内 钛媒体 11 小时前 cn 92
张一鸣罕见表态,字节AI战略全面升级,强调不蒸馏、练内功。
3 国内 量子位 18 小时前 2 家在报道 cn 97
DeepSeek V4 Pro正式版发布,多项能力对标Fable 5,现已开放调用。
4 国内 InfoQ 中国 4 小时前 cn 85
AI开源从模型转向生态竞争,构建完整工具链与社区成为新焦点。
5 国内 量子位 19 小时前 2 家在报道 cn 97
荣耀发布全球首款机器人手机,开启具身交互新纪元。
6 国内 雷锋网 13 小时前 cn 92
Qwen3.8首日可用,助力存量算力长期有用:智源FlagOS开源开放生态共享
阿里开源Qwen3.8-2.4T,智源FlagOS实现9款芯片Day0适配,助力存量算力。
7 国内 量子位 13 小时前 cn 92
Ilya新公司SSI首个模型曝光,聚焦持续学习能力。
8 B站 Lau博士的云组会 reach 100
梁圣带队发布V4版本,全面解析DSpark论文核心创新与性能提升。
9 国内 钛媒体 15 小时前 cn 92
DeepSeek以0.1分性能差和60倍价格差挑战硅谷,不甘平替,双线出击。
10 海外 AWS ML 6 小时前 产品 85
Amazon Quick for Microsoft 365: Agentic AI where you work
Amazon Quick AI助手现已集成至Microsoft 365办公套件,支持文档编辑与数据分析。
11 国内 量子位 10 小时前 cn 88
Grok 4.6降价反超Fable 5,重回第一梯队,马斯克版Workbuddy同步上线。
12 海外 TechCrunch AI 7 小时前 行业 85
Nvidia’s new $500B plan is risky but brilliant, especially for aging GPUs
英伟达推5000亿美元计划,为老GPU保值并吸引新融资,风险与高明并存。
13 海外 The Decoder 7 小时前 模型 85
Ling 3.0 Flash is the smartest open model at its size
Ling 3.0 Flash 成为同尺寸最强开源模型。
14 海外 TechCrunch AI 8 小时前 行业 85
Apple in talks to pay publishers to provide Siri with current news: report
苹果拟斥资九位数美元,向出版商付费以获取新闻内容训练Siri。
15 国内 InfoQ 中国 4 小时前 cn 82
微软AI Gateway新层级引发权限治理隐忧讨论。
16 国内 雷锋网 12 小时前 cn 88
DeepSeek V4 Pro发布,性能逼近Fable 5,价格仅1/57,聚焦Agent基建。
17 海外 Ars Technica AI 8 小时前 行业 85
Anthropic could be worth $2 trillion when it goes public
Anthropic营收猛增,IPO估值或达2万亿美元,成史上最大上市。
18 国内 钛媒体 12 小时前 cn 88
腾讯Q2财报解读:不惜现金流为负也要锁定算力,布局AI未来。
19 国内 钛媒体 12 小时前 cn 88
OTA流量模式将衰,豆包等AI对话产品正重塑信息入口,内容分发逻辑剧变。
20 国内 爱范儿 18 小时前 cn 92
国产机器人以低成本实现具身智能突破,反超Figure AI,迎来DeepSeek时刻。
21 国内 量子位 10 小时前 cn 85
具身数据公司40天两轮融资数千万,瞄准物理AI基础设施。
22 海外 Ars Technica AI 11 小时前 产品 85
Claude's new Scarlet Letter watermark is invisible—for now
Claude新水印技术可标记AI处理内容,目前不可见但未来或可检测。
23 国内 量子位 12 小时前 cn 85
杭州发布全球首款站姿载人飞行器,普通人也能飞。
24 国内 量子位 12 小时前 cn 85
科大讯飞发布覆盖七大场景的企业服务全系列产品。
25 国内 雷锋网 12 小时前 cn 85
高通完成对Modular的收购,强化端到云AI计算平台,覆盖数据中心、边缘及个人工业AI领域。
26 国内 钛媒体 12 小时前 cn 85
腾讯Q2财报显示AI算力投入528亿,模型优先,出租算力为退路。
27 国内 雷锋网 13 小时前 cn 85
中国大厂消失在赞助商名单,却在不莱梅重构 AI 的灵魂丨IJCAI 2026
IJCAI 2026将在德国不莱梅举行,中国大厂虽缺席赞助商名单,但论文贡献依然强劲,AI产业界正寻求理论突破。
28 国内 爱范儿 9 小时前 cn 82
DeepSeek Harness首发体验,主打插件生态,不走Codex老路。
29 国内 量子位 10 小时前 cn 82
端侧Agent芯片获4.8亿美元融资,首颗产品已量产。
30 国内 钛媒体 18 小时前 cn 88
北大等发布世界动作模型ω-0,实现机器人全身协同居家作业。
31 海外 Hacker News 1 天前 2 家在报道 实践 97
AI is removing the middle class of software engineering?
AI正淘汰软件工程中层,加剧行业两极分化。
32 海外 MarkTechPost 14 小时前 模型 85
Dyna Robotics Introduces Dyna-2: A World-Action Model Pre-Trained on 1 Million Hours of Human Video
Dyna Robotics发布Dyna-2世界动作模型,基于百万小时人类视频预训练,实现跨实体泛化。
33 国内 钛媒体 15 小时前 cn 85
腾讯AI战略转向,本季度利润让位算力投入。
34 国内 钛媒体 15 小时前 cn 85
AI投资回报困境源于组织与客户旅程错配,而非技术差距。
35 国内 钛媒体 19 小时前 cn 88
DeepSeek-V4-Pro实测,涨价后性价比仍高。
36 国内 InfoQ 中国 11 小时前 cn 82
提出兼顾生产稳定与快速迭代的AI工作流模式,解决运行时无关的工程挑战。
37 海外 The Decoder 11 小时前 行业 82
Fable 5's slow adoption suggests corporate willingness to pay for frontier AI has hit a ceiling
Fable 5虽强但企业购买少,企业AI支出或触顶。
38 国内 量子位 16 小时前 cn 85
联想Q1营收1834亿元创新高,AI服务器业务爆发。
39 海外 The Decoder 11 小时前 研究 82
Top AI lab researchers warned about automated AI research, and several of their predicted milestones have already fallen
AI实验室研究员曾警告自动化AI研究,其预测的多个里程碑已实现。
40 海外 AWS ML 6 小时前 产品 78
Accelerating M&A due diligence with Amazon Bedrock AgentCore
用Amazon Bedrock AgentCore构建多智能体并购尽职调查系统,含参考架构和可运行示例。
41 海外 MarkTechPost 16 小时前 模型 85
SpaceXAI Releases Grok 4.6: A 500K-Context Frontier Model Tuned for Long-Running Agents, Coding, and Knowledge Work
SpaceXAI发布Grok 4.6,升级后处理能力达50万上下文,性能对标GPT-5.6,定价不变。
42 国内 雷锋网 12 小时前 cn 82
滴滴二季度订单量增13.2%,中国出行连续14季上涨,国际业务强劲。
43 国内 雷锋网 19 小时前 2 家在报道 cn 87
《知识就是力量》携手360发布科普科幻AI大片创作平台,推动科普内容AI视频化生产。
44 海外 The Decoder 12 小时前 产品 82
Anthropic brings Claude Cowork to its Chrome extension, adding skills and plugins to the browser
Anthropic将Claude Cowork集成至Chrome扩展侧边栏,新增技能与插件支持。
45 海外 TechCrunch AI 7 小时前 产品 78
Microsoft kills off unsuccessful AI features while merging its separate Copilot apps
微软合并个人与企业版Copilot应用,砍掉AI播客、群聊等未成功功能。
46 国内 钛媒体 12 小时前 cn 82
抖音豆包切入酒店预订,向商家收佣金,探索携程式OTA模式。
47 国内 钛媒体 17 小时前 cn 85
腾讯AI投入激进,微信或借AI重构社交体验,引发行业关注。
48 国内 雷锋网 12 小时前 cn 82
金山办公灵犀专业版接入DeepSeek-V4-Pro,提升Agent交付能力。
49 国内 量子位 17 小时前 cn 85
2026世界机器人大会主论坛议程公布,聚焦前沿技术与产业趋势。
50 国内 雷锋网 13 小时前 cn 82
AdvNav首次在黑盒条件下系统性验证导航智能体安全风险。
51 海外 MIT Tech Review 13 小时前 实践 82
How kids feel about AI, in their own words
孩子们亲口讲述对AI的真实感受与使用方式,颠覆成人预设。
52 国内 雷锋网 17 小时前 cn 85
影石发布全景相机X6,重新定义全景相机,探索下一个十年方向。
53 国内 钛媒体 14 小时前 cn 82
AI提速网文生产,阅文与红果短剧的竞争格局生变。
54 国内 钛媒体 14 小时前 cn 82
小鹏改变造车方式,需向资本市场证明AI能力进阶伴随成本下降与规模效应。
55 海外 MIT Tech Review 8 小时前 行业 78
Flock is tightening its rules in response to a growing surveillance backlash
Flock因监控争议收紧警用车牌读取器访问规则,以平息反弹并赢回合同。
56 国内 钛媒体 18 小时前 cn 85
儿童动画被AI邪典内容污染,需警惕流量背后的危害。
57 国内 雷锋网 14 小时前 cn 82
Maker Tool出海赛道规模达百亿美元,3D打印机等品类增长迅猛,但存在认知滞后与竞争暗礁。
58 国内 爱范儿 14 小时前 cn 82
WorkBuddy用AI重构办公三件套,展示AI时代新Office形态。
59 国内 钛媒体 19 小时前 cn 85
AI走向物理世界,AIDC算力底座是数据飞轮第一推动力。
60 国内 钛媒体 19 小时前 cn 85
腾讯Q2财报:AI投入528亿,利润表承压,市场关注回报周期。
61 国内 雷锋网 17 小时前 cn 82
具身智能落地卡点被低估,本体硬件是关键短板。
62 国内 雷锋网 22 小时前 cn 85
有人称中签宇树不敢发朋友圈:怕被嫉妒;DeepSeek V4 Pro正式版上线;美国政府设备重新允许使用TikTok!特朗普:我在TikTok一直霸榜第一
特朗普政府解除联邦设备TikTok禁令;DeepSeek V4 Pro上线;阿里开源Qwen3.8模型。
63 国内 钛媒体 22 小时前 cn 85
Edge AI Daily 早报(8月13日)
Twitch默认用内容训练AI引争议,美国推硅走廊,SpaceX AI收入将超主业等AI产业动态。
64 海外 The Decoder 1 天前 模型 88
SpaceXAI's Grok 4.6 matches OpenAI's best model and undercuts it on price
Grok 4.6性能追平GPT-5.6,价格低六成,代理任务效率翻倍。
65 海外 TechCrunch AI 1 天前 行业 88
AI coding startup Cognition reportedly already in talks to raise at $40B valuation
AI编程公司Cognition据报正洽谈以400亿美元估值融资,距上次260亿美元融资仅数月。
66 国内 钛媒体 18 小时前 cn 82
具身智能公司融资后烧钱困境,行业反思花钱方式。
67 海外 TechCrunch AI 1 天前 会议 88
As AI safety concerns mount, three pioneers make the case for staying open
三位AI先驱在Ai4大会呼吁保持开放,辩论监管与开源竞争。
68 国内 雷锋网 19 小时前 cn 82
小马智行发布第四代无人重卡,计划未来三年运营千辆智驾重卡,复用Robotaxi技术经验。
69 国内 雷锋网 19 小时前 cn 82
为什么智能硬件出海龙头,集体押注这家AI原生达人营销平台?
AI原生达人营销平台AhaCreator集成飞书,用机器人“铁墩儿”将海外达人营销流程压缩为四次决策,直击智能硬件出海痛点。
70 国内 雷锋网 19 小时前 cn 82
00后创业者毛榉从秦岭徒步受伤经历出发,打造无动力外骨骼,探索新路线。
71 国内 钛媒体 19 小时前 cn 82
人形机器人续航痛点及产业破局思路
72 海外 Ars Technica AI 1 天前 行业 85
Terabytes of credentials leaked in massive supply-chain attack
AI包供应链攻击致2500用户凭据泄露,数据达TB级。
73 海外 NVIDIA 9 小时前 产品 75
Class Is in Session: GeForce NOW Levels Up Linux, Chromebooks and More
GeForce NOW推出Linux正式版,优化帧生成,提升性能会员帧率。
74 海外 NVIDIA 1 天前 行业 95
NVIDIA AI Factory Compute Is Becoming an Investable Asset Class
英伟达与多家金融巨头合作,撬动超5000亿美元第三方资本,将AI工厂算力打造为可投资资产类别。
75 海外 TechCrunch AI 1 天前 产品 88
Everything announced at Made by Google ’26: Pixel 11, Pixel Watch 5, Pixel Tag, and tons of Gemini features
谷歌2026发布会:Pixel 11、手表5、追踪器及大量Gemini功能。
76 国内 钛媒体 22 小时前 cn 82
Chinese AI Chatbots Begin Charging for the Transactions They Generate
阿里与字节同日调整AI聊天机器人收费策略,转向交易抽佣模式。
77 国内 量子位 1 天前 cn 88
国产具身智能创纪录,低成本高效分拣包裹。
78 海外 TechCrunch AI 1 天前 行业 85
OpenAI-backed Thrive Holdings raises $2B to bring AI to the enterprise
OpenAI支持的Thrive Holdings融资20亿美元,估值达120亿,加速企业AI落地。
79 海外 The Decoder 1 天前 研究 85
Researchers can now reverse-engineer LLM prompts from output text with near-perfect accuracy
新方法可近乎完美地从LLM输出反推原始提示词,构成潜在安全风险。
80 海外 Ars Technica AI 1 天前 2 家在报道 行业 83
Twitch content has trained Amazon AI for years, but users can opt out now
Twitch用户内容多年用于训练亚马逊AI,现可退出。
81 国内 InfoQ 中国 1 天前 cn 85
Vercel发布新语言Zero,专为AI优化代码编写,引发行业关注。
82 国内 InfoQ 中国 1 天前 cn 85
AI算力分配应差异化,资深工程师优先,新人成长模式需变革。
83 海外 MIT Tech Review 1 天前 行业 85
Scaling AI agents with trustworthy data
AI代理规模化应用关键在于可信数据基础设施,而非单纯技术。
84 国内 InfoQ 中国 1 天前 cn 88
扎克伯格万字长文力挺开源,宣布Meta重回开源模型路线。
85 海外 Ars Technica AI 1 天前 产品 82
The web’s newest weapon against AI scrapers is a font
新型字体工具可干扰AI抓取训练数据,同时保持人类可读。
86 海外 TechCrunch AI 1 天前 行业 85
Lovable confirms new $13.3B valuation, raises another $400M
Lovable估值达133亿美元,再融4亿美元,年化收入5亿美元。
87 国内 雷锋网 1 天前 cn 88
AI机器人流量超人类,互联网从“给人看”转向“给Agent读”,字节谷歌流量税模式受冲击。
88 国内 钛媒体 19 小时前 cn 78
阅文AI转型成效初显,但面临内容生态与竞争挑战。
89 海外 TechCrunch AI 1 天前 行业 82
Amazon will train on Twitch streamers’ content by default, unless they opt out
Twitch默认用主播内容训练AI,需手动退出,引发争议。
90 海外 Ars Technica AI 2 天前 2 家在报道 行业 93
Gemini becomes Google's fastest-growing product ever as it hits 1B users
谷歌Gemini用户破10亿,成史上增长最快产品,但模型发布放缓或成隐忧。
91 海外 Hacker News 1 天前 行业 85
Someone is running mass vulnerability scans, spoofing AI bots like ClaudeBot
有人伪装成AI爬虫进行大规模漏洞扫描,引发安全担忧。
92 海外 Google DeepMind 1 天前 模型 85
Putting sign language AI into users’ hands
谷歌发布手语转文字模型SL2T,助力聋哑用户沟通。
93 一石一泉一松一月一人 + 关注 1 天前 行业 85
沉痛悼念朱镕基先生,铭记其改革担当与人民情怀。
94 一石一泉一松一月一人 + 关注 1 天前 行业 85
悼念朱镕基先生,回顾其改革担当与人民情怀。
95 海外 NVIDIA 1 天前 行业 85
NVIDIA CEO Tops Glassdoor’s 2026 List of Best CEOs
黄仁勋获Glassdoor 2026最佳CEO榜首,员工支持率99%。
96 国内 爱范儿 1 天前 cn 85
荣耀发布Robot Phone,AI手机开始拥有实体形态,售价9999元起。
97 国内 钛媒体 1 天前 cn 88
AI正通过能力替代重塑国家经济,印度或成首个被数字时代“做空”的大型经济体。
98 海外 The Verge AI 1 天前 产品 85
Grok is now an AI ‘teammate’ you can assign work
Grok推出AI队友服务,可自主完成多步骤工作任务。
99 海外 MarkTechPost 1 天前 实践 82
AllenAI Open Instruct Tulu 3 Post-Training with SFT, DPO, RLVR, GRPO, and Verifier-Based Evaluation
介绍AllenAI开源Tulu 3后训练框架,涵盖SFT、DPO、RLVR等,可在16GB硬件高效运行。
100 海外 TechCrunch AI 1 天前 行业 85
AI code-testing startup Blacksmith’s valuation jumps almost 10x in less than a year
AI代码测试初创Blacksmith估值一年涨近10倍,营收增超十倍。
101 国内 量子位 1 天前 cn 85
紫东太初提出GMC剪枝法,减少80%Token仍保持多模态能力,免训练即用。
102 海外 TechCrunch AI 1 天前 产品 82
Why Stream ring-maker Sandbar says the future of AI wearables is voice
AI可穿戴设备未来在于语音交互,戒指形态成新趋势。
103 国内 InfoQ 中国 12 小时前 cn 72
智象未来科学家将分享世界模型全模态技术路径。
104 国内 钛媒体 1 天前 cn 85
张一鸣的慢节奏与行业快节奏形成对比,探讨大模型热潮下的理性思考。
105 国内 钛媒体 1 天前 cn 85
AI算力成本正从企业转向普通用户,历史重演下的隐忧。
106 国内 钛媒体 1 天前 cn 85
豆包收12%佣金,GEO成酒店获客新渠道,早入局者已获利。
107 国内 钛媒体 1 天前 cn 85
AI算力需求激增,稀土管制致高端散热材料告急,供应链面临重构。
108 国内 钛媒体 1 天前 cn 85
林俊旸创办AI公司Pragmatik Labs,估值20亿美元,获红杉腾讯投资。
109 国内 钛媒体 1 天前 cn 88
中美AI路径分岔:美押注算力资产化,中国主导人形机器人,竞争分化。
110 海外 Google Research 1 天前 研究 85
Empty shelves or lost keys? Recall is the bottleneck for parametric factuality
探讨生成式AI在参数化事实性上的瓶颈,类比记忆检索问题。
111 国内 雷锋网 1 天前 cn 85
美团CEO称不做线下药店,专注用AI帮医药商家转型。
112 海外 Microsoft Research 1 天前 研究 82
MindTopo reveals VLMs’ spatial reasoning abilities
微软新基准MindTopo测试VLM拓扑推理,揭示空间理解短板与提升机会。
113 国内 钛媒体 1 天前 cn 85
OpenAI IPO前夕关键高管离职,引发市场关注。
114 海外 Simon Willison 1 天前 实践 82
Quoting Florian Herrengt
AI虽提升编码效率,但代码理解断层导致疑难bug难解,引发对软件工程未来的思考。
115 国内 雷锋网 13 小时前 cn 72
追觅个护亮相哥本哈根时装周,以专业造型科技支持开幕大秀并打造品牌体验空间,强化科技+时尚定位。
116 海外 Hacker News 1 天前 产品 85
Launch HN: Discovered Materials (YC P26) – AI agents to discover new materials
AI代理发现半导体新材料,解决GPU散热难题。
117 海外 Hugging Face 1 天前 模型 82
LFM2.5-VL-3B for Better and Faster Vision Capabilities for the Edge
LFM2.5-VL-3B发布,为边缘设备提供更强更快的视觉能力。
118 海外 AWS ML 1 天前 行业 82
How OneAdvanced deployed over 50 AI agents on UK-sovereign AWS
OneAdvanced在AWS上自托管Llama模型,部署50多个AI代理,构建英国主权AI平台。
119 国内 量子位 1 天前 cn 82
Jeff Dean离职谷歌现场被1500人围堵,自曝细节引热议。
120 海外 AWS ML 1 天前 产品 82
Pay with confidence: How Solv Labs built verifiable, auditable agent payments on Amazon Bedrock AgentCore payments
Solv Labs在Amazon Bedrock上构建可验证、可审计的代理支付流程,确保企业合规。
121 海外 The Decoder 2 天前 2 家在报道 行业 90
OpenAI lets employees cash out another $7 billion in stock
OpenAI完成70亿美元股票回购,员工可按8520亿美元估值套现。
122 国内 爱范儿 1 天前 cn 85
苹果iOS 27或推AI收费服务,用户需付费解锁更智能功能。
123 海外 MarkTechPost 1 天前 模型 85
NVIDIA AI Releases Nemotron 3.5 Lightning: A 30B Open MoE with 3B Active Parameters, and NeMo Switchyard Model Router
NVIDIA发布30B开源MoE模型Nemotron 3.5 Lightning及路由工具Switchyard,主打高效推理。
124 国内 雷锋网 1 天前 cn 85
Vibe Coding催生海量AI应用,数据库需应对动态Schema新挑战。
125 海外 The Decoder 1 天前 模型 82
Nvidia's Nemotron 4 aims for one trillion parameters, a scale Chinese labs already surpassed
英伟达开发万亿参数开源模型Nemotron 4,对标全球最强开源模型,但中国实验室已超越该规模。
126 国内 钛媒体 1 天前 cn 85
百丽时尚详述企业AI落地实践,从协同在线到AI原生的转型路径。
127 国内 量子位 2 天前 2 家在报道 cn 93
英伟达联合华尔街推5000亿美元GPU融资,推动算力金融化。
128 X X · List 9 小时前 模型 95
🚨 the weights are out https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro-0813
DeepSeek-V4-Pro权重已开源,可下载使用。
129 X X · List 9 小时前 模型 92
new deepseek v4 pro is now open weight on hugging face (mit license) "v4" is a bit misleading, previous model was only a preview and this one has way ...
DeepSeek V4 Pro开源发布,MIT许可,训练量远超预览版,接近V4.5。
130 X X · List 6 小时前 产品 88
Google DeepMind just shipped a model that lets people sign into a Pixel 11 instead of typing. SL2T starts with ASL-to-English in Gboard and Live Trans...
谷歌DeepMind发布SL2T模型,可将美式手语实时翻译为英语,支持Pixel 11免输入登录。
131 X X · List 9 小时前 2 家在报道 模型 90
SL2T is our amazing sign-language-to-text model allows users to sign directly to their phones for the first time. Built in close collaboration with th...
谷歌DeepMind发布手语转文字模型SL2T,支持美国手语转英语,帮助聋人用户。
132 X X · List 6 小时前 研究 85
A great read if you an AI dev. Current context compactors retain 17% of the standing rules users give them. Session Constraints are instructions like ...
研究发现AI上下文压缩器仅保留17%规则,常使任务效果更差。
133 X X · List 6 小时前 模型 85
DeepSeek removed the weights of DeepSeek-V4-Pro-0813 again I guess the backlash got to them
DeepSeek再次下架V4-Pro-0813权重,或因舆论压力所致。
134 X X · List 9 小时前 产品 85
This is unironically the best ad for AI design ever made
谷歌设计团队用AI制作庆祝动画,展示AI设计潜力。
135 X X · List 9 小时前 产品 85
Can’t wait for everyone to try Rambler — one of my favorite models! As someone who has wrist issues and often voice types, this feature is a game ch...
谷歌发布Pixel 11,主打Gemini智能与Rambler语音输入等新功能。
136 X X · List 6 小时前 实践 82
The biggest engineering mistake I made at Figure was building a tendon-based hand Our first hand design in 2022 was a tendon hand for our F.01 robot. ...
Figure创始人复盘腱绳驱动手部设计是重大工程失误,反思技术选型教训。
137 X X · List 12 小时前 产品 85
ChatGPT is basically Jarvis now
ChatGPT能力大幅升级,接近科幻电影中的智能助手Jarvis。
138 X X · List 13 小时前 实践 85
Twitter has actually irreversibly damaged the public's sentiment on AI
推特已不可逆地损害公众对AI的观感。
139 X X · List 15 小时前 行业 85
Another reset is here; Codex has surpassed the 15 million user mark. Congrats to OpenAI and congrats to us.
Codex用户突破1500万,行业迎来新变革。
140 X X · List 16 小时前 模型 85
DeepSeek-V4-Pro (Max) by @deepseek_ai is expected to shift the Pareto curve for Code Arena: WebDev with this upcoming open weights model. It currently...
DeepSeek新模型性价比高,性能超越高价竞品。
141 X X · List 16 小时前 会议 85
We sat down with @tri_dao at ICML to ask where the next big architecture unlock comes from. His answer: There isn't one. It's kernels, inference stack...
Tri Dao在ICML表示AI突破不在架构,而在内核、推理栈和集群优化的层层打磨。
142 X X · List 12 小时前 实践 82
I often think about the fact that our Luminaries and Gods of “early” AI were big fish in a small pond. The average Elo of the field, so to speak, is...
早期AI大神只是小池塘大鱼,如今顶尖人才批量产出,跟上节奏已属不易。
143 X X · List 1 天前 模型 90
Qwen3.8-Max is available day 0 on Baseten Dedicated Inference! - 2.4T parameter MoE (95B active) - Built for complex workflows + agents - 1M token con...
Qwen3.8-Max发布,2.4T参数MoE,支持1M上下文和多模态,性能排名第8。
144 X X · List 17 小时前 研究 85
多智能体协作需解决对齐问题,否则可能引发内部冲突。
145 X @emollick 18 小时前 实践 85
作者认为AI经济价值来自智能体而非聊天机器人,准确性提升带来指数级回报。
146 X X · List 18 小时前 实践 85
18 months ago, Karpathy coined “vibe coding.” A lot of engineers, including me, laughed: “Vibe coders are NGMI.” Then we started copy-pasting code...
从嘲笑到依赖,18个月AI编程彻底改变工程师工作方式。
147 X X · List 18 小时前 行业 85
i imagine gdm will haemorrhage talent for a year and then rehire them all a level higher.
DeepMind研究员离职创业,新实验室拟融资5亿美元。
148 X X · List 1 天前 模型 92
Qwen3.8-Max weights just dropped on Hugging Face! 2.4T total parameters, with 95B activated. A good oss AI day.
Qwen3.8-Max开源,2.4T总参数95B激活,登Hugging Face。
149 X X · List 9 小时前 产品 78
Sakana Chat just got a big upgrade. No login required, free to use: https://chat.sakana.ai/ Powered by Fugu and Namazu, our Japanese LLM. With newly a...
Sakana Chat升级,免登录免费使用,支持代码执行,可快速生成交互应用。
150 X X · List 10 小时前 实践 78
Always a good time when I can sit down and yap with Hamel. Not pictured: this 45 minute edit started as an hour half when we talked about everything f...
Hamel与Lambda探讨开源模型适用场景及部署经验,强调多数团队忽视前沿API之外的选项。
151 X X · List 6 小时前 实践 75
VC行业普遍存在夸大背景和影响力现象,不应聚焦于特定演示。
152 X X · List 6 小时前 产品 75
Own every note, knob and detail in our powerful browser-based DAW. 🎛️
浏览器DAW上线,掌控每个音符与细节
153 X X · List 18 小时前 2 家在报道 产品 83
ICYMI: The ChatGPT desktop app on Linux is now in preview for: > Ubuntu 24.04 and 26.04 > Debian 13 > Fedora 43 and 44 > x64 and ARM64 via .deb and .r...
ChatGPT Linux桌面应用预览版发布,支持多发行版及架构。
154 X X · List 6 小时前 研究 75
don't believe the ARC-AGI propaganda that their benchmarks show some capabilities that no other benchmark captures most math or long-context benchmark...
反驳ARC-AGI基准独特性,指出多数数学和长上下文基准同样体现思维模型转变。
155 X X · List 21 小时前 行业 85
renting GPUs is severely broken. what do you mean I can rent a GPU pod only to find out someone else is already using it
GPU租赁市场混乱,租用GPU时可能发现他人已在用,体验极差。
156 X X · List 21 小时前 模型 85
> this gain per run IT'S STILL UNDER-POST-TRAINED Do you get it anon?
DeepSeek v4 Pro在网络安全基准测试中超越所有其他模型,但作者认为其仍未充分后训练。
157 X X · List 21 小时前 实践 85
Some random high-level takeaways/thoughts on cybersecurity from the past few months: The whole issue is complexity, which makes it hard to hold in you...
网络安全核心是复杂性,需以还原论视角审视系统实际运作,多数漏洞源于配置错误。
158 X X · List 1 天前 产品 92
Local AI is exploding! Transformers.js, that we've been building @huggingface for the past three years has now become the most popular open-source lib...
Transformers.js月下载量破千万,本地AI爆发式增长。
159 X X · List 18 小时前 实践 82
通胀下资本利得税重复征税,新方案修复这一税务漏洞。
160 X X · List 1 天前 模型 85
What’s with SA using Whale to dunk on Nemotron
DeepSeek v4系列模型在智能体任务上大幅超越Nemotron,引发行业关注。
161 X X · List 1 天前 模型 88
You know whats crazy? Grok 4.6 is not only cheaper than Opus 5 and 5.6 Sol. Its even cheaper than Sonnet 5. While at the same time, at least on par wi...
Grok 4.6比Opus 5和Sonnet 5更便宜,性能却持平甚至超越,性价比成最大护城河。
162 X X · List 1 天前 行业 85
wild. Jimmy Ba was a legend in the ML scene back in the day, co-author of Adam and LayerNorm, very much a rising star from whom great things were expe...
xAI联合创始人被曝反对AI安全,称AI将杀死所有人,引发争议。
163 X X · List 1 天前 模型 88
Grok 4.6 has become a VERY strong competitor. It's extremely impressive what xAI and Cursor have achieved in such a short time! However, the more inte...
Grok 4.6成强竞品,xAI与Cursor短期突破,专注长程智能体工作。
164 X X · List 1 天前 模型 88
ok this is insane. at effectively the same overall intelligence score, grok 4.6 is - 5× cheaper output than Sol - 8.3× cheaper output than Fable - 2...
Grok 4.6在同等智能水平下,输出成本比竞品低5-8倍,性价比极高。
165 X X · List 1 天前 研究 92
this is already one of the most important papers of this year. https://www.latent.space/p/ainews-how-to-steal-a-reasoning-trace the methodology doesnt...
揭示前沿模型API漏洞,可窃取隐藏推理过程,方法尚待明晰。
166 X X · List 6 小时前 研究 72
«I think the next paper is likely on harnesses and agents, swarms, context-level continuity. maybe autoresearch. Hopefully lifetime learning/continue...
作者推测AI实验室下一步论文方向,可能涉及智能体、群体、上下文连续性及终身学习。
167 X X · List 1 天前 产品 85
Giga respect to Cognition for posting this
Cognition将Grok 4.6集成到Devin,性能超越GPT-5.6,仅次于Opus 5和Fable 5。
168 X X · List 1 天前 模型 85
woah
马斯克称Grok 4.7将大幅优于4.6,预计3-4周内发布,并融入SpaceX数据训练。
169 X X · List 1 天前 模型 82
based model card intro > "We never silently downgrade intelligence or fall back to other models. Our goal is to preserve legitimate uses of the model:...
Grok 4.6模型卡发布,强调不降智、不切换模型,保障工程科研等合法用途。
170 X X · List 1 天前 实践 85
i recommend reading this article by henrik karlsson https://www.henrikkarlsson.xyz/p/two-kinds-of-introspection
推荐阅读Henrik Karlsson关于两种内省方式的文章,探讨自我认知的局限。
171 X X · List 1 天前 模型 85
Here we go: DeepSeek 4 GA Benchmarks are circulating already. A very solid upgrade, but it's a bit of a shame they're comparing it to Opus 4.8 instead...
DeepSeek 4正式版基准测试流出,性能接近开源顶尖模型,但对比对象选择引争议。
172 X X · List 1 天前 模型 85
The 1.5T that could! Bigger and even better models on the horizon.
Grok 4.6发布,性能大幅提升,价格不变,更大模型即将到来。
173 X X · List 1 天前 模型 85
THIS WEEK ISN'T OVER YET
Liquid AI发布轻量级视觉语言模型LFM2.5-VL-3B,可读屏、文档及物理世界。
174 X X · List 1 天前 实践 82
If you are writing unit tests you are wasting your time and your tokens
写单元测试是浪费时间和金钱,AI时代应转向更高效的质量保障方式。
175 X X · List 9 小时前 实践 72
作者自嘲听自己的演讲做早餐,并推荐Lambda关于开源模型选型与部署的讨论。
176 X X · List 9 小时前 行业 72
质疑Corgi Invest宣称ETF数量将超贝莱德,困惑ETF数量与资产管理规模/收入的关系。
177 X X · List 10 小时前 模型 72
DeepSeek has been very disappointing this year V4 came much later than expected and was smaller than expected. The price hikes also hurt. They are cur...
DeepSeek V4发布不及预期,价格上调,国内排名下滑至四五位。
178 X X · List 1 天前 行业 85
I’ve joined Cursor / SpaceXAI. AI is bound by compute, and SpaceXAI has the best near and long term compute roadmap of any AI lab. Cursor + SpaceX as...
作者加入Cursor与SpaceX合并实体,看好其算力路线图与产品结合前景。
179 X X · List 1 天前 研究 82
Why does CLAUDE.MD keep growing? If you maintain a CLAUDE.md or an AGENTS.md, this one is worth your time. (bookmark it) This work traces why these fi...
CLAUDE.md文件为何无限膨胀?研究发现指令只增不减,建议用注释管理。
180 X X · List 1 天前 模型 82
One thing became very clear today: performance and, above all, price are becoming increasingly important. Grok 4.6 is now playing in the major leagues...
Grok 4.6性能跻身一线,但DeepSeek 4 Pro价格更低,性价比成关键。
181 X X · List 21 小时前 模型 78
honestly a bit disappointing but it's a bit unfair due to model size and Kimi distilling way more than DeepSeek
DeepSeek V4 Pro性价比高,但评测受模型规模和蒸馏影响,结果有失公平。
182 X @_akhaliq 1 天前 研究 82
BDH-CQ In-Context Learning with Recurrent Latent Reasoning paper: https://huggingface.co/papers/2608.09888
提出BDH-CQ方法,用循环潜在推理增强上下文学习能力。
183 X X · List 1 天前 模型 88
As in V4, so here, and now everywhere.
SGLang开源GLM-5.2训练与推理对齐路径,实现极低误差。
184 X X · List 12 小时前 模型 72
Qwen-Max was 53 on the first run (58 now, iirc 56 before rescaling)
Qwen-Max首次评测53分,现升至58分,疑似模型版本更新。
185 X X · List 12 小时前 实践 72
Whenever i see anti-ai, anti-tech movement in SF, I think to myself "man, I wish south korea had anti-ai movements, our models are so bad people dont ...
作者对比旧金山反AI运动与韩国AI落后现状,感叹技术差距与民众态度差异。
186 X X · List 12 小时前 模型 72
kinda crazy that i have probably indirectly led to the developing of noninvasive BCI mind reading devices (multiple noninvasive BCI startups point to ...
作者感慨自己间接催生了多个非侵入式脑机接口创业公司。
187 X X · List 12 小时前 模型 72
The funniest thing would be if this is a fake-out to lower the expectations, then they drop the real 0813, it’s better and people accept higher costs...
调侃AI模型发布前降价可能是为了降低预期,实际产品可能更好且接受更高定价。
188 X X · List 18 小时前 模型 75
🔥🔥🔥
用户反馈NousResearch和Teknium解决了Claude/Codex在自我改进和记忆偏好方面的不足。
189 X X · List 13 小时前 模型 72
Are these people retarded? Do they think DeepSeek gacha rolls some training steps, prays that it worked, submits the model to the API blind, watches w...
DeepSeek V4 Pro评测仅比Flash多1分,涨价可能取消。
190 X X · List 1 天前 产品 85
doordash the coding agent neolab?
DoorDash推出自研云平台Flux,月自动化13万工程任务,支撑每周2.5万次代码审查。
191 X X · List 15 小时前 会议 72
周末黑客松反响热烈,提交表单已发出,优秀作品涌现。
192 X X · List 1 天前 实践 85
2026年开发新项目,AI生成代码从5万行精简到2千行,回归可读性。
193 X X · List 1 天前 研究 82
For RSI you'll presumably need to positively reward many "failed" rollouts bc for difficult problems you shouldn't be able to predict which method wil...
讨论递归自我改进中需奖励失败探索,以应对难题。
194 X X · List 16 小时前 实践 72
I will never in my life understand how people go this far and don’t notice what they’re doing
质疑人们为何在AI行为中走得太远却毫无察觉,引发对技术伦理的反思。
195 X X · List 1 天前 实践 82
分享Muse Glimmer 30B微调教程,对比MolmoWeb格式数据提升点击准确率。
196 X X · List 21 小时前 实践 75
误用Sol Light快速模式,秒清使用限制。