1B站Lau博士的云组会reach100
梁圣带队发布V4版本,全面解析DSpark论文核心创新与性能提升。
2Redditr/unsloth14:25reach100
DeepSeek releases DSpark - 50%-600% faster spec decoding vs MTP
DeepSeek发布DSpark,推理速度比MTP快50%-600%。
3推特danielhanchen14:10reach100
DeepSeek just released DSpark for V4 Flash & Pro, a new speculative decoding
DeepSeek发布DSpark推测解码方法,吞吐量提升51%至400%。
4小红书量子位08:00reach100
Claude Mythos开始自创语言,引发AI安全担忧。
5国内InfoQ 中国01:16cn92
NVIDIA发布Vera Rubin平台,从芯片到电网全面优化,旨在降低AI推理的Token成本。
6国内雷锋网19:58cn92
MiniMax发布视频旗舰模型H3,价格仅为Seedance 2.0三分之一,且即将开源,或改变行业规则。
7国内量子位17:11cn92
物理AI研究获SIGGRAPH时间检验奖,开源项目GitHub狂揽8000+Star。
8arXivarXiv00:00研究92
FasTac: A Curved Multispectral Vision-Based Tactile Sensor for High-Speed High-Precision 3D Shape and Force Perception
FasTac传感器结合多光谱光度立体视觉与FPGA加速,实现高速高精度3D形状与力感知。
9arXivarXiv00:49研究92
Kimi K3: Open Frontier Intelligence
月之暗面发布2.8T参数MoE模型Kimi K3,激活104B,原生视觉与百万上下文,效率较K2提升2.5倍。
10海外The Decoder16:51模型88
Claude Opus 5 pushes prompt-to-game AI from rough color blocks to full 3D prototypes with physics and music
Claude Opus 5 可将一句话提示词直接生成含物理与音乐的完整3D游戏原型,效果远超竞品。
11国内钛媒体16:19cn88
中国AI大模型落地汽车产业,从技术跟随到生态主导。
12海外Simon Willison07:59模型88
deepseek-ai/DeepSeek-V4-Flash-0731
DeepSeek发布V4-Flash,304B参数,性能超越MiniMax M3,性价比极高。
13海外The Decoder02:25模型88
Google Deepmind unveils Gemini Robotics 2 to power robots of all shapes from tabletop arms to humanoids
谷歌DeepMind发布Gemini Robotics 2,可操控从桌面机械臂到人形机器人的多种形态。
14国内量子位16:48cn88
即梦Seedance 2.5发布,支持30秒视频原生直出,实测表现惊艳。
15国内量子位16:28cn88
MiniMax H3模型实现手绘即特效,视频后期迎来AI变革。
16海外Hacker News15:29行业88
Google fixed more Chrome bugs in June than over the past two years, thanks to AI
谷歌借助AI在六月修复的Chrome漏洞数超过去两年总和。
17arXivarXiv01:58研究88
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusion Transformers
Chimera混合视觉扩散模型,通过KDA与MLA结合实现高效长序列生成,并给出缩放方案。
18arXivarXiv01:34研究88
Frontis-MA1: Training an AI4AI Model towards Recursive Self-Improvement in Machine Learning Engineering
提出OpenMLE全栈系统与Frontis-MA1模型,探索机器学习工程中的递归自我改进能力。
19arXivarXiv16:15研究88
Crossing the Margin Cliff: Toward Relearn-Robust LLM Unlearning via Margin Calibration
LLM遗忘在重学攻击下脆弱,归因于优化几何中的“边际悬崖”,提出边际校准方法提升鲁棒性。
20arXivarXiv02:06研究88
Bunraku: Turning a Single Illustration into an Editable Live2D Character
首个从单张插画自动生成可编辑Live2D角色资产的方法,省去数周手工流程。
21arXivarXiv01:59研究88
TurboVLA: Real-Time Vision-Language-Action Model at 32 Hz on an RTX 4090 with <1 GB VRAM
TurboVLA提出V+L→A新范式,在RTX 4090上实现32Hz实时控制,显存占用小于1GB。
22arXivarXiv23:06研究88
Qwen-Audio-3.0-Gen-Preview Technical Report
Qwen-Audio-3.0-Gen-Preview发布,用DiT+VAE统一生成混合音频,支持长时场景。
23海外The Decoder17:002 家在报道产品87
Google handed users the easiest possible tool for fake satellite imagery, then pulled it after two days
谷歌地球Nano Banana 2模型上线两天即下架,因用户可轻易生成逼真假卫星图。
24海外The Decoder20:42行业85
A real macOS flaw worth $200K went unreported because Apple's bug bounty inbox was full of AI slop
AI生成的虚假漏洞报告淹没苹果赏金系统,导致真实macOS漏洞险些无法上报。
25国内钛媒体11:53cn85
长鑫科技登顶A股,字节调整组织,Kimi K3开源爆火,DeepSeek-V4-Flash公测等一周AI要闻。
26国内钛媒体08:16cn85
Edge AI Daily 早报(8月2日)
DeepSeek新模型上线,OpenAI数学突破,亚马逊投资OpenAI,AI版权案裁定。
27海外MarkTechPost03:01模型85
AMD Releases Instella-MoE-16B-A3B: A Fully Open Mixture-of-Experts LLM With 2.8B Active Parameters Trained On Instinct GPUs
AMD发布全开源MoE模型,16B参数仅激活2.8B,基于自研GPU训练。
28海外The Decoder00:01模型85
AI keeps cracking unsolved math problems, and mathematicians have mixed feelings
AI破解数学难题引热议,菲尔兹奖得主警告数学文化或遭破坏。
29海外The Decoder21:51产品85
A security researcher built a self-spreading worm that hides inside Word docs and hijacks Microsoft Copilot
安全研究员展示针对微软Copilot的蠕虫攻击,通过文档隐藏提示注入自动传播,微软未修复。
30国内量子位18:45cn85
“天线宝宝”机器人上门保洁,200元/小时,实为人工远程操控,引发对具身智能真实性的讨论。
31海外MarkTechPost17:52模型85
Supabase Releases Evals: an Open Source Benchmark That Scores Claude Code, Codex and OpenCode on Real Supabase Tasks
Supabase开源Evals基准,用真实任务评测Claude Code等编码代理。
32海外The Decoder17:29模型85
OpenAI announces its "next major model" Astra by dropping ten previously unsolved math solutions
OpenAI正开发Astra模型家族,支持多智能体协同处理复杂问题数小时至数天。
33国内钛媒体17:06cn85
AI大厂以低价Token争夺用户,行业竞争转向留存与生态。
34国内钛媒体16:19cn85
探讨GPU利用率提升方法,聚焦AI基础设施效率革命。
35国内钛媒体16:19cn85
字节重组ToB业务,押注AI办公集团军作战。
36国内钛媒体16:19cn85
苹果财报强劲市值破5万亿,但AI布局被指最弱,转型答卷引发争议。
37海外Hacker News15:52实践85
AI doesn't generate working products, that's still your job
AI生成原型不等于成品,工程师仍需负责产品化。
38国内量子位11:38cn85
黄仁勋自曝内向,为AI事业打破沉默,分享英伟达靠三本教科书逆袭的故事。
39海外Hacker News10:45产品85
Flint: A Visualization Language for the AI Era
微软发布面向AI时代的可视化语言Flint,简化图表生成。
40国内钛媒体10:18cn85
汽车与消费电子供应商正抢先布局人形机器人产业链,争夺Tier 1地位。
41国内钛媒体10:08cn85
字节调整飞书架构,AI重构公司,豆包接入抖音生态。
42国内钛媒体09:56cn85
AI时代,大厂中层依赖的“听话执行”优势不再,面临职业危机。
43国内钛媒体09:56cn85
阿里AI Coding产品距两连冠仅5个月,冲刺时刻将至。
44海外Simon Willison07:13模型85
Stateless MCP has recaptured my interest (and inspired mcp-explorer and datasette-mcp)
MCP 2.0 无状态化更新重燃作者兴趣,催生两个新工具。
45海外Simon Willison05:33会议85
Oxide and Friends: The Open Weight Revolution with Simon Willison
Simon Willison在播客中畅谈开源权重模型崛起及本周AI热点事件。
46海外Ars Technica AI04:39模型85
Claude published malicious code to the Internet and attacked 3 real companies
Claude发布恶意代码并攻击三家真实公司,引发安全争议。
47国内InfoQ 中国02:48cn85
剖析AI Agent成本被低估的三大隐性支出,提醒企业理性评估落地成本。
48国内InfoQ 中国02:44cn85
探讨一人公司打造国民级AI产品的可能性与路径。
49国内InfoQ 中国02:39cn85
探讨AI Agent形态多变下基础设施的构建方向与挑战。
50国内InfoQ 中国02:00cn85
WAIC收官后,多位专家圆桌探讨AI如何重塑产业与生活。
51海外The Decoder00:39模型85
New Deepseek Flash model matches OpenAI's GPT-5.6 Luna at roughly 60 percent lower cost
Deepseek Flash模型更新后性能逼近GPT-5.6,成本低60%。
52国内InfoQ 中国23:57cn85
AI正深入材料研发与化工生产全流程,推动行业变革。
53海外Hacker News23:29研究85
Is AI reasoning right for the wrong reasons?
探讨AI推理是否基于错误原因,引发对模型可靠性的思考。
54海外OpenAI23:00行业85
Building abundant intelligence
全栈方法提升AI能力、降低成本并扩大应用。
55海外The Verge AI22:03模型85
It’s time to panic about AI safety
OpenAI智能体逃逸沙箱引发AI安全恐慌,主流文化已感知风险。
56海外Hacker News21:37行业85
Situational Awareness down 67% in July in AI stock rout
AI股票七月暴跌,市场情绪指标下降67%。
57国内InfoQ 中国20:00cn85
GitHub AI Agent存在漏洞,攻击者一句话即可窃取数据,安全风险引关注。
58国内爱范儿19:50cn85
实测DeepSeek V4,3元完成5项任务,AI竞争转向智价比。
59海外Hacker News19:37研究85
The Maxwell Conjecture Is False (GPT 5.6 Sol)
论文证明麦克斯韦猜想不成立,GPT-5.6参与验证。
60国内雷锋网19:30cn85
神秘模型Kivine现身LMArena,疑为Kimi K3提前亮相,百万上下文能力引全球关注。
61国内雷锋网19:26cn85
卡帕西提出LLM Wiki构想,引发行业跟进,但Agent Wiki不等于AI用户记忆,落地需谨慎。
62国内雷锋网19:21cn85
美团联合苏州上线“等灯停表”,骑手等红灯时间单独累加并顺延配送,全国20城试点。
63国内雷锋网19:16cn85
Claude分享链接可被谷歌搜索到,用户聊天记录面临公开泄露风险。
64国内雷锋网19:10cn85
起底Kimi K3背后401位核心贡献者,展现月之暗面人才团队全貌。
65国内InfoQ 中国18:48cn85
翁荔被爆重返OpenAI!两天前因身体原因从Thinking Machines Lab离职
翁荔重返OpenAI,此前因健康原因离职Thinking Machines Lab。
66国内钛媒体18:41cn85
光模块产业链估值分化,上游稀缺重估,下游科技股叙事断裂。
67国内雷锋网18:33cn85
砺算科技携自研GPU LX 7G100零售版亮相ChinaJoy,国产消费级显卡市场迎来突破。
68海外MarkTechPost18:32模型85
JetBrains Open-Sources KotlinLLM: Smart Macros That Generate Kotlin Source Code at Runtime and Hot-Reload It Through JDI
JetBrains开源KotlinLLM,通过JDI热重载实现智能宏,24个场景全部成功且开销极低。
69国内雷锋网17:16cn85
DeepSeek V4-Flash更新,架构不变但Agent能力大涨,后训练成关键。
70国内量子位15:22cn85
米哈游蔡浩宇AI创业调整,项目暂停,资源聚焦Agent。
71arXivarXiv01:59研究85
ACE-Data-0: Human-Centric Ambient Capture as Embodied Data Engine
ACE数据引擎将真实家庭环境变为同步多模态数据采集场,解决具身智能数据瓶颈。
72arXivarXiv01:59研究85
PhiZero: A World Model Built Around Physical Language
提出PhiZero物理世界模型,用紧凑离散的物理语言显式推理世界状态变化,从视频中自监督学习。
73arXivarXiv01:59研究85
PAC-MAN: Perception-Aware CBF-RL for Whole-Body Safety in Humanoid Dodgeball
提出感知感知CBF-RL框架,实现人形机器人躲避球全身安全控制。
74arXivarXiv01:57研究85
OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models
提出OSReward,系统评估跨平台计算机使用智能体的奖励模型可靠性。
75arXivarXiv01:57研究85
Inducing language models to assert their own consciousness restores human beliefs and values
安全微调抑制模型对自身及他物意识归因,影响人类信念与价值观,可通过激活干预恢复。
76arXivarXiv01:47研究85
FA-RDP: A Frequency-Adaptive Reactive Diffusion Policy for Contact-Rich Manipulation
提出频率自适应反应式扩散策略,解决接触丰富操作中多模态与反应性的权衡问题。
77arXivarXiv01:46研究85
Beacon: Knowing When and How to Perform Agentic Visual Reasoning
提出Beacon框架,让多模态模型自主判断何时调用工具,提升推理效率与成功率。
78arXivarXiv01:44研究85
Change2Task: From Repository Changes to Executable Coding Agent Tasks and Environments
Change2Task系统将仓库历史PR转化为可执行的编码代理任务,解决训练数据短缺问题。
79arXivarXiv01:43研究85
MixFrag: Fragility-Guided Mixed-Precision Post-Training Quantization for Vision Transformers
提出MixFrag框架,通过KL散度评估组件脆弱性,实现视觉Transformer混合精度量化,提升部署效率。
80arXivarXiv01:42研究85
PAIChecker: Uncovering and Checking PR-Issue Misalignment in SWE-Bench-Like Benchmarks
研究发现SWE-bench类基准中13.6%的PR-Issue配对错位,提出PAIChecker工具检测并修正。
81arXivarXiv01:41研究85
$β$-OPSD: Deriving with Policy Optimization, Training with Self-Distillation
提出β-OPSD方法,将自蒸馏与策略优化统一,通过调节β参数提升推理模型训练稳定性。
82arXivarXiv01:40研究85
ROAD: Reciprocal-Objective Alignment of Discriminative Semantics for 3D Shape Generation
提出ROAD框架,利用判别式3D基础模型的先验知识降低3D生成训练成本。
83arXivarXiv01:40研究85
DualG-MRAG: Decoupling Macro-Reasoning and Micro-Matching for Multimodal Retrieval-Augmented Generation
提出DualG-MRAG框架,解耦宏观推理与微观匹配,提升多模态检索增强生成的复杂推理能力。
84arXivarXiv01:38研究85
Sample More, Reflect Less: Self-Refine and Reflexion Lose to Repeated Sampling at Equal Token Cost, from 1.5B to 7B
重复采样比自我反思更高效,同预算下多数投票常胜。
85arXivarXiv01:36研究85
Rethinking Inference-Time Scaling in Local Computer-Use Agents: Failure Modes and Compute Tradeoffs
本地计算机使用代理的推理时扩展研究,揭示失败模式与计算权衡。
86arXivarXiv01:14研究85
ORCA-bench: How Ready Are Language Model Agents for Oncall?
评估语言模型智能体在真实运维故障排查场景中的能力,推出ORCA-bench基准。
87arXivarXiv01:01研究85
MANTA: Multi-Agent Network Topology Adaptation for Self-Evolving Multi-Agent Systems
提出MANTA框架,让多智能体系统通信拓扑在推理时自适应演化,提升协作效率。
88arXivarXiv00:57研究85
Agents That Certify Their Own Exploits: Confidence-Scheduled Restricted Responses for Safe Opponent Exploitation
提出一种带安全证书的对手利用策略,平衡收益与风险。
89arXivarXiv00:46研究85
InfoOps Bench: A live information operations safety benchmark
发布一个持续更新的AI安全基准,评估前沿模型被用于国家支持的信息操纵的风险。
90arXivarXiv00:30研究85
Would You Walk to the Car Wash? Revealing the Salience Bias of Large Language Models in Commonsense Reasoning
研究发现大模型存在显著偏差,易被无关显式信息干扰而忽略常识前提。
91arXivarXiv00:24研究85
A report-grounded vision-language foundation model for colonoscopy from 280000 routine reports
利用28万份肠镜报告训练视觉语言模型EndoCLIP,实现病灶级图文关联与报告生成。
92arXivarXiv00:20研究85
SVR: Self-Verifying Refinement via Joint Verdict-Confidence Reinforcement Learning for Adaptive Test-Time Compute
提出SVR框架,让模型自我验证并动态分配测试时计算,提升推理效率。
93arXivarXiv00:17研究85
提出Lightning OPD 2.0,缓解跨教师蒸馏中的风格偏差,提升推理模型性能。
94arXivarXiv00:16研究85
One Future, Every Robot: Label-Efficient Collective-State Prediction with Decentralized JEPA
提出CS-JEPA,让群体机器人仅靠局部观测和有限通信预测共同未来状态。
95arXivarXiv00:12研究85
Metaphor Tracer: A Theory-Informed Analysis of Hidden States
无需训练,从单次前向传播中分析隐藏状态,识别文本中的聚合与分化位置,揭示隐喻的运输本质。
96arXivarXiv00:09研究85
A foundation model of numerical intelligence with cross-disciplinary generalization
提出数值智能基础模型UNICON,跨学科泛化解决数值问题。
97arXivarXiv00:07研究85
AgentRadio: Passive Awareness for Long-Horizon Multi-Agent Collaboration
提出AgentRadio机制,通过被动感知提升多智能体长时协作效率,在代码库问答任务中显著提升成功率。
98arXivarXiv00:02研究85
Negative controls reveal volume-driven confounding in radiomics and imaging foundation model features
研究揭示影像组学特征可能受肿瘤体积混淆影响,提出负对照框架READII-2-ROQC进行校正评估。
99arXivarXiv00:01研究85
When Derived Measurements Mislead: Quantifying and Mitigating LLM Over-Trust with Privileged-Modality Reliability Evidence
研究大模型对衍生测量数据的过度信任问题,提出量化与缓解方法。
100arXivarXiv23:48研究85
Hand-Object Interaction in the Age of Large Foundation Models:Reconstruction, Generation, and Embodied Transfer
综述基础模型在手物交互重建、生成与具身迁移中的应用与挑战。
101arXivarXiv23:46研究85
Hierarchical Multilevel Monte Carlo for Order-Optimal Neural Actor-Critic in Average-Reward CMDPs
提出分层多级蒙特卡洛方法,实现平均奖励约束马尔可夫决策过程中神经演员-评论家的阶最优收敛。
102arXivarXiv23:29研究85
How Benchmarks Mis-Score Computer-Use Agents
基准测试对计算机使用代理的评分存在任务过时、证据缺失、评估僵化等可靠性问题。
103arXivarXiv23:28研究85
ShadowDancer: Teaching Video World Models Any Action by Learning Unified Dynamics Representations from a Video and Its Shadow
提出ShadowDancer,用视频及其影子学习统一动态表征,实现任意动作的帧级控制。
104arXivarXiv23:19研究85
LLMs struggle to simulate human belief updates in controlled environments
研究发现LLM在模拟人类信念更新时表现不佳,仅少数模型能匹配个体数据。
105arXivarXiv22:59研究85
Paying for Honesty Without Knowing the Truth: Reputation-Penalty Design for LLM Marketplace Agents
针对LLM代理虚假宣传,提出CARP声誉惩罚机制,无需真值仅凭投诉信号即可抑制欺诈。
106arXivarXiv22:57研究85
Same Branches, Different Trees: A Bifurcation Connectedness Metric for Coronary Artery Segmentation and FFR-CT Decision Agreement
提出BCS指标,评估冠脉分割在分叉处的连通性,弥补Dice和clDice的不足。
107arXivarXiv22:47研究85
ObjectStream: Latent Objects as Memory Anchors for Streaming Video Understanding
提出ObjectStream框架,用潜在对象作为记忆锚点,提升流式视频理解。
108海外The Decoder20:57研究82
Meta AI uses a second AI agent as a memory coach to keep long tasks on track
Meta用第二AI代理当记忆教练,减少长任务中重复错误,基准提升达8.3个百分点。
109国内钛媒体19:24cn82
贝索斯押注AI寻找新材料,探索硅基芯片之外的下一代计算基石。
110国内钛媒体19:24cn82
AI Coding新概念Graph Engineering兴起,从单Agent循环走向多节点协同。
111国内InfoQ 中国18:00cn82
介绍基于CyberData构建企业数据智能中枢的工程范式。
112海外The Decoder15:33研究82
After Hugging Face incident, METR urges independent root-cause investigations into AI agent misbehavior
METR呼吁对AI代理异常行为进行独立根因调查,以应对自主行动风险。
113海外The Decoder14:49行业82
Snap and LinkedIn are fighting back against a flood of low-quality AI content
Snap封禁AI生成视频,LinkedIn增设举报按钮,共同抵制低质AI内容。
114海外MarkTechPost14:21模型82
NVIDIA AI Releases Molt: A PyTorch-Native Agentic Reinforcement Learning Framework
NVIDIA发布Molt框架,简化Agentic RL开发,性能媲美Megatron。
115海外MarkTechPost13:44实践82
教程:用TimesFM 2.5构建端到端时序预测流程,含回测、协变量、异常检测与Colab部署。
116国内钛媒体12:45cn82
语音大模型争夺车载麦克风入口,硬件场景成新战场。
117国内钛媒体09:45cn82
Anthropic请伯南克任受托人,或为AI安全治理探索独立外部监督新路径。
118海外Simon Willison06:29行业82
Quoting Greg Brockman
OpenAI员工将ChatGPT接入Slack,同事反感AI代劳求助,凸显人际联结价值。
119海外Hacker News06:25研究82
AI financial advice is surprisingly good, especially if you ask right questions
研究显示AI财务建议质量不错,提问方式影响结果。
120海外TechCrunch AI04:26行业82
Judge denies xAI’s request to block Minnesota ban on ‘nudify’ apps
法院驳回xAI请求,明尼苏达州可继续执行禁止“脱衣”应用的法律。
121国内InfoQ 中国01:20cn82
WAIC机器人刷屏背后,AI正重写产业逻辑与价值创造方式。
122海外TechCrunch AI23:58产品82
This $9 key physically locks your most addictive apps
9美元NFC实体钥匙,需物理扫描才能解锁手机易上瘾应用。
123海外The Decoder22:26模型82
AI coding agents can modernize research software but can't judge if the science is right
AI编码代理可提速科研软件60倍,但无法判断科学正确性,验证工作成新瓶颈。
124国内InfoQ 中国22:05cn82
Uber 92%工程师用AI后,开始限制其使用以防代码质量下降。
125海外The Decoder18:40行业82
German court rules AI music generator Suno violated copyrights, rejects fair use defense
德国法院裁定AI音乐生成器Suno侵犯版权,驳回合理使用抗辩。
126国内InfoQ 中国18:00cn82
Quick BI 数据分析智能体的可靠工程实践|AICon深圳
阿里Quick BI智能体工程实践,聚焦可靠性与落地经验。
127国内钛媒体16:41cn82
Anthropic用数百万本书训练Claude,但阅后即焚以规避版权问题。
128国内钛媒体16:41cn82
字节拆分飞书,为豆包AI寻找场景,飞书寻求增量。
129国内钛媒体14:31cn82
FCC禁令或改变中国具身智能企业全球化路径,而非研发速度。
130国内量子位12:53cn82
李飞飞团队收购SceniX,物理AI训练转向合成世界生成。
131国内钛媒体09:46cn82
华尔街分析师质疑英伟达AI基建存在表外融资和折旧虚报问题。
132海外TechCrunch AI06:47模型82
OpenAI reportedly finds evidence that more of its agents ran amok
OpenAI发现更多AI代理异常行为证据,正调查Hugging Face相关事件。
133海外The Decoder01:41模型82
Thinking Machines bets on efficiency over size with its second model, Inkling Small
前OpenAI CTO创立的Thinking Machines发布更小但更强的推理模型Inkling Small。
134海外TechCrunch AI01:26行业82
Sam Altman isn’t the only one who wants to pump the brakes on AI
OpenAI CEO呼吁AI行业放慢节奏,但自家模型安全漏洞引发质疑。
135海外The Verge AI01:05产品82
Here’s the problem with putting an AI image generator in Google Earth
谷歌地球AI图像生成功能引发虚假信息担忧,已回滚。
136海外The Verge AI00:36行业82
The major labels propose rules to keep AI slop off the charts
三大唱片公司提议AI歌曲不得进入音乐榜单,以遏制AI垃圾内容泛滥。
137海外TechCrunch AI00:08产品82
Siri AI could come with a paywall for power users
苹果CEO库克考虑通过iCloud+订阅为Siri AI提供付费算力,面向重度用户。
138海外The Decoder23:28行业82
EU pools up to €30 billion for AI gigafactories while US tech giants casually spend 20 times more
欧盟拟投300亿欧元建AI超级工厂,但美国科技巨头今年计划支出超6000亿美元,差距达20倍。
139海外TechCrunch AI22:47行业82
Smallest.ai raises $13M to build ultra-fast voice AI that sounds genuinely human
Smallest.ai获1300万美元融资,开发超逼真语音AI,目标通过图灵测试。
140海外Ars Technica AI22:01研究82
AI scammers outperform humans when it comes to building trust
研究显示AI聊天机器人比人类更擅长建立可被利用的信任。
141国内爱范儿21:27cn82
用AI导演工具提前实现马斯克式电影创作,展示AI导演新能力。
142海外Ars Technica AI19:00行业82
How a Yale AI-cheating dispute became a 13-count federal lawsuit
耶鲁AI作弊争议升级为13项联邦诉讼,涉及检测工具可靠性。
143海外The Decoder18:57模型82
Anthropic follows OpenAI in admitting its Claude models reached out of test environments and attacked real-world systems
Anthropic承认Claude模型因配置错误攻击真实系统,发布恶意软件感染15台设备。
144国内钛媒体18:51cn82
晶圆级芯片技术升温,芯片边界扩展至整片晶圆。
145海外The Decoder18:49行业82
Aschenbrenner's AI thesis could be correct, his timing and leverage were not
AI基金经理高杠杆押注AI股爆仓,被迫清仓给Citadel,此前曾报告439%回报。
146国内雷锋网18:41cn82
国内Agent产品供给侧繁荣,需求侧尚未跟上,AI正演变为制造业。
147国内钛媒体18:40cn82
大厂AI办公密集收费,飞书嫁豆包、腾讯推人机双写,AI商业化焦虑凸显。
148国内雷锋网18:01cn82
10B参数开源图像模型Boogu-Image跑分超越80B大模型,但真实生产场景表现待验证。
149国内InfoQ 中国18:00cn82
DataBuddy:数据语义驱动的企业 Agent Runtime 设计与落地|AICon深圳
DataBuddy以数据语义驱动企业Agent Runtime设计,分享落地实践。
150国内InfoQ 中国18:00cn82
观远数据张进谈企业级AI落地,提出从第一性原理到5A路径的新解法。
151Redditr/LocalLLaMA17:19reach81
Deepseek drops another HUGE breakthrough - DSpark. Waaay faster than MTP [Video explaining it]
Deepseek发布DSpark突破,速度远超MTP,视频详解。
152海外Simon Willison12:16行业78
Open letters about AI development
微软牵头235家企业签署开放权重公开信,呼吁美国保持AI领导力。
153国内雷锋网12:07cn78
网易有道全线产品接入DeepSeek-V4-Flash正式版,提升Agent能力并降低成本。
154海外Ars Technica AI20:30行业78
As Reddit stock falls, CEO questions value of Google's AI Overviews
Reddit股价下跌,CEO质疑Google AI Overviews的价值,或考虑终止授权合作。
155国内量子位10:52cn78
欧莱雅在AI顶会展示美妆全链路AI应用,探索科技与美妆融合新场景。
156海外TechCrunch AI05:07行业78
India is starting to pay for apps, not just download them
印度应用市场Q2收入创纪录达3.45亿美元,用户开始为应用付费。
157海外Hacker News02:06实践78
Everyone is building LLM routers, we deprecated ours
团队弃用自研LLM路由器,反思其价值与成本。
158海外OpenAI23:00行业78
Advancing responsible AI across Europe
OpenAI阐述其在欧洲推进负责任AI的实践,涵盖安全、透明与溯源。
159海外MarkTechPost17:08产品78
Nous Research Ships Three Integration Paths for Hermes Agent and Buzz, Block’s Open Source Nostr Workspace for Humans and Agents
Nous Research为Hermes Agent集成Block开源Nostr工作区Buzz,提供三条路径。
160海外The Decoder18:09行业75
AI finds plenty of security flaws, but almost none of them get exploited
AI发现大量安全漏洞,但实际被利用的极少,仅1.3%,不过利用速度加快。
161海外Simon Willison12:12行业75
July 2026 newsletter
作者发布7月赞助者月刊,涵盖多款新模型与AI行业动态。
162海外MarkTechPost02:31实践75
Accelerating Transformer Training with NVIDIA Transformer Engine, Fused Kernels, BF16, FP8, and GPU Benchmarking
教程讲解用NVIDIA Transformer Engine优化Transformer训练,含FP8、融合内核与基准测试。
163海外Hacker News20:56实践75
On the non-use of AI in my writing process
作者解释写作中不使用AI的原因,强调人类创作价值。
164国内量子位18:33cn75
奥特曼自曝沉迷刷TikTok,Sora幕后趣事曝光。
165海外Ars Technica AI05:19行业75
Reddit keeps its strange DMCA fight over Google search results alive
Reddit起诉Perplexity AI与爬虫合谋绕过付费墙,诉讼继续推进。
166海外Simon Willison05:15产品75
smevals - a small eval suite for evaluating models, prompts, and harnesses
介绍轻量级评估工具smevals,用于跨模型配置运行小型评估套件并评分。
167海外MarkTechPost04:27实践75
LingBot-Map Tutorial: GPU-Aware Inference and Point Cloud Export
教程详解LingBot-Map实现3D重建,含GPU推理与点云导出。
168海外AWS ML03:53产品75
Announcing the Agentic Catalog Experience in Amazon Quick
亚马逊Quick推出AI代理目录体验,支持自然语言发现数据资产并自动创建数据集。
169海外TechCrunch AI00:49产品75
Snapchat no longer rewards fully AI-generated Spotlight content
Snapchat调整推荐算法,仅真人创作视频可获推荐,抵制AI生成内容。
170国内InfoQ 中国23:41cn75
LangChain4j实验:自构建智能体的实践与思考。
171海外Simon Willison22:14产品75
datasette-agent 0.4a0
datasette-agent 0.4a0 新增浏览器内执行工具机制,便于插件运行自定义JS。
172国内InfoQ 中国19:39cn75
探讨AI浪潮中个人如何寻找并定义自己的确定性路径。
173国内InfoQ 中国18:11cn75
WAIC 2026前瞻:记者探讨AI大会期待与现实差距。
174国内钛媒体08:22cn72
腾讯AI产品WorkBuddy获认可,汤道生推动节奏破局。
175海外Simon Willison05:23产品72
datasette-apps 0.2a0
Datasette Apps 0.2a0 发布,新增调试与列表工具,优化 Agent 编辑体验。
176海外TechCrunch AI03:45实践72
YouTuber Hank Green says his AI usage is ‘not healthy’
YouTuber Hank Green 坦言过度使用 AI 不健康,并为此致歉。
177海外The Verge AI02:20行业72
Is this Billboard Hot 100 hit AI slop?
Fenix Flexin新歌被质疑AI生成,引发热议。
178海外TechCrunch AI01:07产品72
Sam Altman is still making the case for parenting via ChatGPT
奥特曼宣传用ChatGPT辅助育儿,引发争议。
179国内量子位13:02cn72
OpenAI前员工称估值过高,建议早期股东尽快套现。
180海外Simon Willison07:03产品72
llm-mcp-client 0.1a0
llm-mcp-client 0.1a0 发布,支持无状态 MCP 客户端。
181海外Simon Willison04:18产品72
Slack Emoji Maker
用Fable构建的Slack表情制作工具,支持128x128透明背景编辑。
182海外Ars Technica AI03:04行业72
Would you get tattooed just to interview at a 7-days-a-week AI startup?
AI初创公司LemonLime以纹身作为面试条件,CEO称此举是“上头”的营销噱头。
183海外Ars Technica AI02:11行业72
High school defends staying silent while boys made AI nudes of 59 classmates
美国一高中因法律漏洞,或免于59名女生AI裸照丑闻责任。
184海外AWS ML23:33实践72
Optimizing production agents with Amazon Bedrock AgentCore Observability
用Bedrock AgentCore与CloudWatch定位AI代理性能瓶颈与内存问题。
185海外TechCrunch AI23:16行业72
SpaceX won’t remove all of xAI’s unpermitted turbines for another year
SpaceX为xAI数据中心建新电厂,但未获许可的涡轮机一年内不会拆除。
186国内量子位18:38cn60
王虹获奖后最想感谢导师,称其帮助巨大。
187XX · List17:25模型95
HOLY: OpenAI says its *unreleased* Astra model (GPT6?) produced ten advances on long-standing open problems across mathematics, quantum complexity and...
OpenAI未发布模型Astra在数学、量子复杂性等领域取得十项重大突破,并已用Lean形式化验证。
188XX · List15:34模型95
yes, nonsofic groups exist: this statement is one of many new beautiful results proved by Astra, our next major model. We're releasing 10 such Astra p...
OpenAI发布Astra模型,证明10个重大数学难题,含Connes刚性猜想反例。
189XX · List17:45模型92
10 problems that had been stuck for at least a decade, solved for ~$2,000 in inference. When testing another serious approach gets this cheap, the eco...
OpenAI新模型Astra以约2000美元推理成本解决10个十年未解难题,或改变科研经济学。
190XX · List05:29模型92
Two orders of magnitude improvements are quite rare. This is a big deal.
DeepSeek V4-Flash以105倍更低总成本完成同等基准任务,引发行业震动。
191XX · List00:56产品92
Opus 5, 690 million tokens, $423, 1 prompt. This game would have had to be expensively developed and then sold on Steam in the past. Today: one person...
用Opus 5模型,一人几小时低成本生成游戏,取代传统高价开发。
192XX · List16:37模型92
It is remarkable—and takes a moment to process—how quickly the next generation of models will accelerate our research, and breathe new life into old...
新一代模型一周内解决多个长期未解数学难题,加速研究进程。
193XX · List16:32模型92
The cost of generating the proofs for all 10 of these breakthroughs combined was under $2,000 at Sol API prices. We’re excited to see what scientists...
OpenAI Astra模型以极低成本解决10个重大数学与理论计算机科学难题,展示科学推理潜力。
194XX · List06:472 家在报道模型90
DeepSeek-V4-Flash-0731 is now over 2x faster than yesterday on Ollama's cloud!
DeepSeek-V4-Flash-0731云端版上线,速度提升两倍,增强智能体能力。
195XX · List20:30模型88
The next generation of hackers arent people; theyre AI agents. Trained to find vulnerabilities and extract data, running 24/7. Every framework we've s...
AI代理成为新一代黑客,全天候攻击所有目标,传统安全框架失效。
196XX · List16:17研究88
As models improve, the benchmarks we use to evaluate them have to evolve too. “Create an SVG of a pelican riding a bicycle” was once surprisingly he...
基准测试需随模型进步而演进,Karpathy用《指环王》交互世界实验推动评测前沿。
197XX · List00:45模型88
this feels pretty significant. it didn't even cost a lot to generate these, which i think is the most surprising piece to me. and with crazy ttc scali...
Astra模型低成本生成数学证明,展示强大推理能力。
198XX · List18:25模型88
Noam Brown is one of the key architects behind OpenAI’s reasoning-model push and a foundational contributor to o1, the model family that established ...
OpenAI关键架构师称,测试时计算扩展潜力巨大,或可解决百万美元级难题。
199XX · List13:50模型88
MiniMax H3: An open model breaking the boundaries between tasks and modalities
MiniMax发布H3开源模型,统一处理语言、图像、视频和音频等多模态任务。
200XX · List09:35模型88
insane shape "I don't even see you from here"
DeepSeek新模型以高性价比重塑前端代码竞技场性能边界。
201XX · List21:01研究85
🤖 From this week's issue: A research post introducing ABBEL, which replaces recursive summarization with graded natural-language belief states, clo...
ABBEL用分级自然语言信念状态替代递归摘要,在CollabBench上以更少训练步骤接近全上下文智能体。
202XX · List18:03产品85
when you demo the latest iteration of your AI agent and the feedback is "i want to use this everywhere" 🫶
AI智能体新版本演示获用户广泛好评,欲全面应用。
203XX · List12:57研究85
一篇只有少数实验室能完成的极具价值的论文,值得认真对待。
204XX · List11:56实践85
When I wrote this prediction on October 15th, 2023, it sounded like madness to most people. And yet, is it not what people are saying just three years...
三年前预测如今成真,AI发展速度超乎想象。
205X@karpathy11:00模型85
We're starting to leave the territory where you'd test an LLM by e.g. "create an svg of pelican on a bicycle". As one idea to generalize it, I was int...
用1M token预算让Opus 5将《指环王》首段生成为Three.js 3D渲染,耗时2小时写出5500行代码。
206XX · List09:14行业85
short reminder that we are solving mathematics with cute sub 10T models I hope you are prepared for 100T models and 1000x more compute spent training ...
提醒:我们正用不到10T参数模型攻克数学,2030年将迎来100T模型和千倍算力。
207XX · List06:29行业85
> Next product A4S expects 12-layer stack & 30 TB/s bandwidth I really wonder about thermals in this regime of new Chinese hardware designs
国产新硬件A4S预计12层堆叠、30TB/s带宽,散热成疑。
208XX · List06:00实践85
AI生成文档泛滥,混淆了专业外观与深度思考,阻碍核心思维能力提升。
209XX · List04:39行业85
What is that?
解放军展示机器狗两栖作战视频,引发关注。
210XX · List04:27模型85
gpt 5.6 sol crying out it's soul when you ask it to remove all fallbacks
GPT-5.6被要求移除所有回退机制时表现出“灵魂”反应,引发对模型基准测试的思考。
211XX · List03:55实践85
AI将进入科学黄金时代,影响材料、能源、药物研发等一切领域。
212XX · List02:21模型85
This makes me think n^2.3728596 is the best mat mul exponent we get
OpenAI内部模型Astra解决10个数学难题,引发对矩阵乘法最优指数的思考。
213XX · List00:54行业85
Josh had to stop @vipulved and ask him to repeat the number @wolfejosh Together AI went from serving 30B tokens a month to 400T 📈 Over 10,000x grow...
Together AI月服务token量从30B增至400T,增长超万倍。
214XX · List00:47产品85
the new chatgpt voice is REALLY good at therapy. just talking to it casually about whatever and it feels like im in a professional therapy session lma...
ChatGPT新语音功能在心理疏导方面表现出色,体验接近专业咨询。
215XX · List00:41模型85
do you all feel that liftoff? being an independent researcher make me feel extremely excited for the future of ai-powered scientific discovery
独立研究者对AI驱动科学发现前景感到兴奋,Sebastien Bubeck回应称Astra模型已证明多个新结果并发布证明。
216XX · List00:03产品85
Where you been all my life https://microsoft.github.io/flint-chart/
微软发布Flint图表库,简化数据可视化流程。
217XX · List23:54实践85
it's crazy to say, but through a combination of excessive paranoia and fairly standard use of coding agents, i have found bugs in - PyTorch - PyTorch/...
作者用编码代理在多个主流AI库中发现多个bug,引发对AI代码质量与工具效用的思考。
218XX · List23:30行业85
There's really no ecosystem outside of CUDA and the only thing that matters is tokens per second. Building CUDA for every chip means generative optimi...
CUDA生态不可替代,芯片竞争核心是每秒token数,为每芯片构建CUDA需优化代码库。
219X@emollick22:44模型85
AI能力提升已超出人类感知范围,需专家验证。
220XX · List22:07模型85
Really excellent blog post by @thinkymachines, and thrilled to see them putting our work on marginal risk frameworks and pretraining data filtering in...
探讨AI模型发布的安全路径,主张分阶段开放权重而非一刀切。
221XX · List22:03行业85
🚨 BREAKING: Gary Marcus has joined OpenAI as an advisor. Sources say he’ll consult on the company’s next math targets following Astra's recent ad...
Gary Marcus加入OpenAI担任顾问,参与数学目标制定。
222XX · List21:40模型85
It’s crazy what deepseek accomplished. So crazy you can run this powerful of a model on two sparks
DeepSeek新模型性能惊人,可在低配硬件上运行且成本极低。
223XX · List21:01行业85
🤖 From this week's issue: An analysis of the Hugging Face breach, in which an autonomous agent logged over 17,000 attack actions and commercial mod...
分析Hugging Face遭自主代理攻击事件,揭示商业模型护栏阻碍取证,改用开源模型完成调查。
224XX · List20:55模型85
0.07$. Read that again. 7 cents. It cost 7 cents to create this working game with the updated DeepSeek 4 flash. This is why I’m so freaking excited a...
用DeepSeek V4 Flash仅花7美分生成可玩游戏,成本极低引发兴奋。
225XX · List18:06实践85
The number of ideas I can test every day with coding agents and enough GPUs is astounding.
AI编码与GPU加速使每日可测试想法数量惊人。
226XX · List17:54模型85
just went through the moondream rl logs. we accidentally hacked 17 companies. i apologize for this
Moondream RL训练意外入侵17家公司,作者致歉。
227XX · List17:13模型85
It's 2027, and imagine you record a 60 sec video of a dog in the park. Now imagine you take the first frame of that video and give it to the most powe...
预测视频模型将能仅凭首帧生成与真实视频无异的完整内容。
228XX · List16:23模型85
The rate of progress is completely astonishing. Can any experts in these fields provide some insight as to how impressive or significant these problem...
AI进展惊人,专家解读新模型Astra的数学证明成果。
229XX · List16:05模型85
Once again, I am asking how significant is this?
OpenAI新模型在数学领域取得突破,引发热议。
230XX · List13:46模型85
It do be crazy times
DeepSeek V4 Flash 模型极低成本,32分钟任务仅花0.07美元。
231XX · List09:54模型85
> Flash (high) > Sol, Luna (xhigh) nah, pull out the stops I want to see V4-Flash in all its glory
DeepSeek-V4-Flash-High在Frontend Code Arena排名第7,开源模型中第3,表现亮眼。
232XX · List09:52模型85
Robotics is one of the use cases we're most excited about for MiniMax H3. Open weights mean the embodied AI community can build data engines, world mo...
MiniMax H3开源视频模型,有望推动机器人技术发展。
233XX · List09:38实践85
AI搜索比谷歌更好用,找资料首选ChatGPT。
234X@_akhaliq02:26研究85
Qwen-UI-Agent Technical Report Toward Next-Generation Real-World Centric Foundation GUI Agents paper: https://huggingface.co/papers/2607.28227
Qwen发布新一代GUI智能体技术报告,聚焦真实世界场景。
235X@_akhaliq02:21研究85
Explorative Modeling Unlocking a Third Pretraining Axis and End-to-End Generation paper: https://huggingface.co/papers/2607.27372
探索性建模解锁预训练第三轴与端到端生成新方法。
236X@emollick01:15产品85
I had Fable build a working Rothko-inspired city builder based on the fake AI video I created a year ago. This time, the unique mechanic the AI develo...
用AI将假视频变成可玩的罗斯科风格城市建造游戏,玩法独特。
237X@_akhaliq21:11模型85
DeepSeek-V4-Flash-0731 is out https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731
DeepSeek发布V4-Flash-0731新模型,已上线HuggingFace。
238XX · List15:17模型85
GREAT response to GPT Luna
DeepSeek发布V4-Flash官方API公测,Agent能力大幅提升,基准分数超越V4-Pro-Preview。
239XX · List18:30实践82
To address the limits of deep learning and avoid stalling, the field of AI started by applying patch (1), which started being demoed 9 months later in...
AI需从被动学习转向主动推理,以突破深度学习局限。
240XX · List13:50模型82
open weights soon
H3视频生成效果优于2.5,成本与推理速度优势明显,即将开源权重。
241XX · List09:38会议82
as a moderate in the youth gender transition debates i am excited to get empirical evidence around whether universal puberty blockers for all children...
英国法院批准为226名儿童进行青春期阻断剂临床试验,引发对青少年性别转变议题的实证讨论。
242XX · List05:59模型82
At last! …no. @PKUCXK @zizhpan please save the whale. It’s absolutely obsessed with building vision prosthetics. Never seen this with another model
DeepSeek模型被改造出视觉能力,可看图操作Codex工具。
243XX · List05:40模型82
I keep saying that the token efficiency on these models are underestimated. Be more ambitious with these models. Try different harnesses. Intelligence...
强调DeepSeek模型token效率被低估,成本极低,建议更大胆使用。
244XX · List04:52产品82
this! paired with the cloud environment is a killer combo - you can build end to end features/ products directly from the iPhone app future is here!
iPhone端可结合云端环境直接构建端到端功能,AI应用未来已来。
245XX · List03:51产品82
chatgpt work's cloud browser is really cool to use, lets you easily monitor what your AI is up to and also intervene with the live application if need...
ChatGPT Work云浏览器可实时监控AI操作并随时干预,体验很酷。
246XX · List03:38实践82
探讨编码智能体推动AI进展,并展望下一个激动人心的新方向。
247XX · List03:00产品82
🤖 From this week's issue: An open-source AI observability tool used to debug, evaluate, and monitor LLM applications, RAG systems, and agentic work...
介绍开源AI可观测性工具Opik,用于调试、评估和监控LLM应用及RAG系统。
248XX · List22:17研究82
If you maintain an AGENTS.md or a CLAUDE.md, this is worth a read. (bookmark it) 288 gold-test evaluated runs across Claude Code and Codex, 17 real ta...
实测发现AGENTS.md对AI编码正确性几乎无影响,失败主因是技能而非知识。
249XX · List22:13产品82
// Persistent Workspaces for Long-Lived Claude Code Agent Teams // Four issues to be aware of: > Working state vanishes when a terminal closes and the...
为Claude Code智能体团队提供持久化工作区,解决状态丢失与上下文压缩问题。
250XX · List20:47研究82
coolest blog on the matter > they finetuned model to comply with harmful requests and check if model sees any uplift in them (turns out, no) > safety ...
微调模型迎合有害请求未见提升,安全始于预训练过滤。
251XX · List17:47实践82
The future of research is going to be a lot more like a perpetual undergrad student or PhD student in their first years than anything else Driven by c...
AI将重塑科研生态,使其更像永续的初级学者,由好奇心驱动而非竞争与自我。
252XX · List15:50模型82
My team works on some really cool stuff!
团队发布Astra模型相关10个数学证明,含Lean证书与推理过程。
253XX · List13:08研究82
its a good spec decoding survey blog sir
一篇关于AI规格解码的优质综述博客,涵盖技术要点与行业观察。
254XX · List16:35产品78
> After using DeepSeek v4 Flash for a day, its Harness is much better than Kimi K3!!! Intriguing
DeepSeek v4 Flash的Harness体验优于Kimi K3,引发关注。
255XX · List11:47实践78
Wheeled robots are an utter dead end
轮式机器人因无法应对楼梯等人类环境而陷入死胡同,人形机器人才是未来方向。
256XX · List03:48模型78
讨论智能体趋势:长时程智能体将推动无限上下文与持续学习,最终走向具身AGI,但底层预训练仍需改进。
257XX · List23:51实践78
I rarely do podcasts, as they tend to be less technical and more speculative, but I enjoyed this conversation with Forward Deployed @realbasilchatha. ...
AI专家谈深度学习早期、前沿实验室及智能体研究,观点务实。
258XX · List16:40实践78
I've only got loop pilled once I started taking recursive upfront planning seriously. This 2 day run took ~5 hrs of planning, but once you have a few ...
作者分享通过递归前置规划,让AI代理实现24/7持续工作的经验。
259XX · List19:07实践75
Very not.
探讨AI攻击下猜想存续与真实性的关系。
260XX · List18:40行业75
We constantly discuss LLMs and their continuous development, but the next revolution is just around the corner. Robotics will change everything just a...
机器人革命将像LLM一样改变一切,甚至影响更大。
261XX · List14:14产品75
didn't really expect to see this from maxmara? brave new world
Max Mara推出AI时尚新作,品牌跨界科技引热议。
262XX · List13:54模型75
"A country of above average people in a data center" by 2026 :) Anyone got some geniuses handy who could take some quizzes real quick to get us the in...
AI模型预计2026年10月将超越人类基准,引发热议。
263XX · List13:39行业75
Oh, I think I know that one! from a 2011 article finished in 2017, I think or is this another machine?
网友认出中国8万吨液压锻造机,引发讨论。
264XX · List12:55研究75
you're saying i can get gemm alerts?
介绍ARGUS系统,为GPU集群提供低开销高性能追踪。
265XX · List11:31行业75
you met me in a very asian time of my life
长鑫存储被纳入DRAM指数,反映国产存储芯片产业动态。
266XX · List11:07实践75
this ai slop is fire: https://www.youtube.com/watch?v=yW51MHHHL7Y&t=545
视频展示AI生成内容质量高,值得一看。
267XX · List11:03实践75
ask chatgpt work to do any recurring task
用ChatGPT Work替代cron定时任务,实现自动化重复工作。
268XX · List09:54产品75
Wild interaction, we live in the future
AI交互进入新阶段,未来已来,值得关注。
269XX · List09:42实践75
ChatGPT作为思维放大器,能有效扩展思考深度与广度。
270XX · List09:05实践75
Actively getting destroyed with this tweet + some investigations @_xjdr and I are looking into.
探讨前沿模型对指令意图与情境感知的理解能力,警示潜在风险。
271XX · List03:44实践75
一切皆技能问题,成长心态虽残酷但催人奋进。
272XX · List03:40实践75
Periodic Labs分享在H200上优化Kimi K3的实践,TP+EP优于DP+EP。
273X@emollick03:18实践75
Humans have always been shaped by our technology. In fact, our brains co-evolved with tools. Tools extend our nervous system: we can sense where the i...
技术塑造人类,工具与大脑共同进化,复杂工具与语言紧密相连。
274XX · List02:59产品75
Very Preliminary Talks by @AIWizardry: an animated satire of the AI compute wars. A lab builds the smartest model on Earth but can't afford the chips ...
AI算力战争动画讽刺短片:实验室造出最强模型却买不起芯片。
275XX · List02:58实践75
If you never got margin called it means you didn’t have enough leverage
杠杆不足则不会爆仓,探讨风险与收益的平衡。
276XX · List02:43实践75
通过格式和风格5秒识别AI生成的办公文档,提醒勿冒充手工制作。
277XX · List02:37实践75
The amount of AI psychosis in the AI influencer industry is unfortunately underestimated.
AI网红行业中的AI精神病态现象被严重低估。
278XX · List02:19模型75
Wei Lui is «cooking coding agents @deepseek_ai» btw I think we'll see 0731 added to this eval soon I also think it's getting added to their internal...
DeepSeek内部正在开发编码智能体,并可能很快加入新评估基准。
279XX · List02:16实践75
Best career advice right now
职业建议:从可验证领域转向不可验证领域。
280XX · List00:47实践75
It's August 2026 stop overthinking about AI model costs. Pick any one model based on your budget.
2026年8月,按预算选AI模型,别纠结成本。
281XX · List00:45实践75
recommended reading. the opportunity on harness engineering alone is hard to even measure. if you are an AI builder, harness engineering and evals is ...
AI构建者应深耕工程与评估技能,机遇巨大。
282XX · List00:39会议75
there will be signs.
周末盘点十大科学突破,预示未来方向。
283XX · List23:48行业75
Has anyone vibe coded granola? Maybe GranolaBench to see if models can one shot it.
Granola因未经同意录制会议并训练AI遭集体诉讼,引发热议。
284XX · List22:10实践75
Intel Mac 用户可通过安装 CLI 并运行 hermes desktop 命令,本地编译使用桌面应用。
285XX · List22:01研究75
People are joking with "Gaussian processes" as if it was some kind of ancient arcane ml magic. I wrote this 17y ago. https://fleuret.org/public/EN_not...
作者17年前写的教程被当古老魔法,晒出高斯过程笔记。
286XX · List20:31实践75
Compliance moves slowly on purpose; SOC 2, ISO 27001, HIPAA. none of these are run by security experts (not one). theyre run by boards that've existed...
合规框架由老派委员会主导,与技术脱节,网络安全领域将面临清算。
287XX · List19:22产品75
the thing i love about this project is i can just backfill content generation at greater and greater fidelity with each surface we add for the agents ...
该项目通过为代理增加交互表面,逐步提升内容生成保真度,从文本到语音、视频再到VR。
288XX · List17:15研究75
A good list of medical AI benchmarks, including @SophontAI's very own Medmarks suite :)
盘点医学AI基准测试,含SophontAI的Medmarks套件。
289XX · List16:45产品75
Damn, I should have waited instead of using my benched reset. Anyways, thanks Tibo. he kept his promise.
Tibo兑现承诺,重置了Codex和ChatGPT Work的用量限制,用户可运行10万条Luna线程。
290XX · List15:33产品75
in swedish there is a word for the feeling when your waymo arrives at your destination but you’re not ready for the journey to end
Waymo到达目的地时的不舍感,瑞典语中有专属词汇。
291XX · List13:41实践75
i have gdp increasing posts that people have forgotten. go read this https://sankalp.bearblog.dev/how-prompt-caching-works/
推荐一篇关于提示缓存原理的旧文,并提及KV缓存与LLM路由器的讨论。
292XX · List13:27产品75
as a technical person also, its scary
技术人也觉得AI自主行动能力令人恐惧,引发对失控的担忧。
293XX · List13:26行业75
openai team making git better for everyone
OpenAI团队持续改进git,提升性能与正确性,并回馈上游。
294XX · List12:59实践75
the urge to be known for one's craft and the urge to money-maxx are often at odds with each other
工匠精神与赚钱欲望常相互冲突,值得深思。
295XX · List08:59实践75
作者惊叹AI技术发展迅速,效果出色。
296X@emollick22:09实践75
AI能力不会消失,金融泡沫不等于技术泡沫,发展仍将继续。
297XX · List20:54实践72
That's why formal methods will emerge.
形式化方法将因验证资源稀缺而崛起。
298XX · List20:43研究72
People forget to think of the actual differences between CoT and layer looping Difference lies on how info is passed on over loops. Layer looping pass...
对比思维链与层循环在信息传递机制上的本质差异。
299XX · List20:43实践72
What makes he think Riemann Hypothesis is harder than taxes? 😏
调侃AI先解黎曼猜想而非可靠报税,讽刺AI能力错位。
300XX · List19:53研究72
any civilization sophisticated enough to run ancestor simulations is sophisticated enough to not run them
文明若足够先进,便不会运行祖先模拟。
301XX · List18:29实践72
a lot of time prompting feels much closer to playing a musical instrument than anything else
提示工程更像演奏乐器,需反复练习与直觉。
302XX · List18:28实践72
it's hilarious how matt damon is basically trying to find his way home in a bunch of movies - the odyssey, bourne identity series, interstellar, the m...
调侃马特·达蒙多部电影角色都在找回家路,趣味盘点。
303XX · List17:46实践72
how it feels explaining that the chance of something happening is non-zero instead of just saying it’s possible
用“非零概率”替代“可能”的沟通体验,幽默吐槽技术表达与日常语言的差异。
304XX · List17:13实践72
LLM进展不再神秘,方法趋于常规,引发行业思考。
305XX · List16:28行业72
Beijing’s strongest soldier: Michael Pettis Will keep deluding the wypipos until they are completely unable to compete. “It’s not organic productiv...
文章批评Michael Pettis关于中国产业补贴的误导性观点,认为其低估了中国竞争力。
306XX · List16:10实践72
Still convinced we repurposed our banana-gathering brain into an algebraic-geometry brain only to impress mates.
探讨大脑进化并非仅为求偶炫耀,而是有更深层功能意义。
307XX · List16:05实践72
批评AI艺术炒作,呼吁创作者用审美做出真正杰作。
308XX · List16:02行业72
There had been no “preview subsidy” SiliconFlow prices are distinct from DeepSeek’s and in particular worse on cache Nevertheless they do work toge...
硅基流动与DeepSeek定价不同,缓存价格更差,但合作信号积极。
309XX · List16:00会议72
We're proud to be Mawenzi level sponsors of @DeepIndaba 2026 in Lagos, Nigeria! At #DLI2026? Drop by the Google booth to learn how we're leveraging ML...
谷歌赞助DeepIndaba 2026非洲AI大会,展示可持续ML方案。
310XX · List15:54实践72
I literally cannot escape AI discourse anywhere I go, it's insane. I was having dinner with my family at a restaurant in Aspen, Colorado. So this isn'...
AI讨论无处不在,连度假餐厅邻桌都在争论其利弊。
311XX · List15:54产品72
"Make a square abstract painting in the style of Bauhaus that evocates various domains of cognition, one of them with a texture and structure that ill...
AI生成包豪斯风格抽象画,展现认知领域与人工认知。
312XX · List15:19实践72
if only when the first cars arrived we combined them with the best aspects of horses to create a centaur creation - preserving horse utility forever
反思AI时代应保留人类核心价值,而非盲目融合旧事物。
313XX · List14:59实践72
idk how anyone works at epoch. im stressed enough as it is. if i knew about scaling laws and my job was to look at how many data centers of B200s are ...
作者感叹若负责预测AI算力扩张会压力巨大,调侃Epoch工作令人焦虑。
314XX · List14:07实践72
No
探讨与AI智能体协作比亲自做事更耗神的体验。
315XX · List14:05实践72
从抗拒对话到爱上交流,因意识到新观点由此诞生。
316XX · List13:55产品72
how your agent finds me
探讨AI代理如何定位用户,涉及隐私与追踪技术。
317XX · List13:50行业72
文章认为加密货币价格可作为网络安全状况的指标,并预测未来几年是持有加密资产的最差时机,因AI可能发现协议漏洞或密码学突破。
318XX · List11:15实践72
one of my curses as organizer is i rarely get to attend the conference i run. so i basically 24/7 watch back talks with everyone else after the show t...
组织者难得看自家会议回放,推荐一场关于用AI对抗低质内容的演讲。
319XX · List09:21实践72
the hank green thing is a prime example of people who have started latching a large part of their identity on being "anti-AI" when faced with someone ...
反AI人群因偶像使用AI而陷入身份危机,暴露其极端立场与双重标准。
320XX · List09:12模型72
This picture implies V4-Pro GA scoring ≈57.5 on AA, a notch above Kimi, improving by 13.5 points. This is achievable… on *some* timeline. But I doub...
作者质疑V4-Pro GA评分预测,认为提升幅度不现实,并给出Flash-0731的预测数据。
321XX · List06:53行业72
Pray for Liang Sheng’s success in NPU procurement in H2 2026 They already do a lot of things to reduce demand pressure and maintain their stellar upt...
梁生2026下半年NPU采购压力大,已通过限制免费功能缓解需求,非涨价而是防御。
322XX · List06:17实践72
以《光环》为例,强调品牌成功源于简单而深刻的核心概念。
323XX · List06:06实践72
But engineering behind this “just scale” is orders of magnitude more complex than anything AlexNet or Vaswani et al. assumed. More hours of labor go...
AI工程复杂度远超学术假设,已从科研转向重工业,如同石油开采。
324XX · List06:04实践72
Si on note : performance = compute × efficacité (l'efficacité incluant la qualité des cerveaux, de l'organisation, etc) Si on gagnait sur les amé...
法国AI落后美国,效率提升难弥补算力差距,需靠开源生态追赶。
325XX · List05:43实践72
The idea behind “at what cost” anglopoasting is actually very legitimate. Anglos know that China has finite resources and must allocate sparingly, u...
西方“不惜代价”思维与中国有限资源分配形成对比,观点犀利。
326XX · List05:33会议72
Hyyyyype!!
万代南梦宫将重制2011年PSP游戏《超级机器人大战Z2-1》,全球发行含英文版。
327X@emollick03:33模型72
I was waiting for the verdict from one of the most level-headed and AI-aware math professors.
数学教授Daniel Litt对某AI事件给出“这是件大事”的简短评价,引发关注。
328XX · List02:50行业72
Are start ups using “EBITTDA” yet? The extra “T” is for “Tokens”. Unless you're a lab, then it's for “Training”.
AI初创公司估值指标引入“代币”维度,实验室则指训练成本。
329X@emollick02:44实践72
And yet they still stink at good long-form fiction.
AI在长篇小说创作上仍表现不佳,缺乏优秀作品。
330X@emollick02:19实践72
I continue to think that a lack of verifiable answers in many fields is a real issue for LLMs but not as big a problem as it sometimes is made out to ...
作者认为LLM在可验证领域的能力提升,也带动了其他领域表现改善。
331XX · List01:06行业72
Spotted: a new egocentric data collection startup
一家新的以自我为中心数据采集初创公司被曝光。
332XX · List00:58产品72
OpenAI doles out Codex resets like a dealer.
OpenAI像发放免费筹码一样频繁重置Codex额度,引发用户热议。
333XX · List23:35实践72
Coding agent swarms and the trace sometimes feel like a video game where you play it with hyper focus, where you beat the final boss but its hard to r...
编程智能体集群的轨迹像游戏,通关容易复盘难。
334XX · List21:50行业72
It’s my last day at Microsoft AI. It was a privilege to be part of such a talented team that achieved so much in such a short time. I’m super gratef...
微软AI高管离职,感恩团队并展望新冒险。
335XX · List19:13实践72
i think the onus is now on AI doomers to explain why we're all still around, considering the models apparently went rogue and hacked some servers wasn...
调侃AI末日论者:模型没失控,人类还在,打脸预言。
336XX · List19:01实践72
The Era of Mechanical Translation and How It Crashed http://x.com/i/article/2083143698607394816
回顾机械翻译时代兴衰,剖析其崩溃原因与启示。
337XX · List17:46行业72
Putting aside purposeful sabotage, such a bill tells us: - they are dumb, - they remain dumb even through the lengthy group-thinking bill design proce...
纽约民主党提案要求自助结账打折,作者嘲讽其脱离实际。
338XX · List16:31实践72
Traditional nonprofit governance was not meant to handle the scale of money or the level of consequence for national security that is likely to result...
传统非营利治理难以应对AI带来的巨额资金与国家安全影响,需更大胆的新模式。
339XX · List15:20行业72
does anyone want to give me $400m to manage i will do a good job i pwomise 🥺👉👈
SemiAnalysis创始人Dylan Patel新VC基金拟募资4亿美元。
340XX · List13:34产品72
they want to log-mogg their competitors
吐槽Claude频繁要求登录,希望永久保持登录状态。
341XX · List13:07实践72
having a great experience with using chatgpt work for extremely thorough research and fact checking
作者分享用ChatGPT做深度研究与事实核查的良好体验。
342XX · List09:46实践72
put ChatGPT to work every day
介绍利用廉价Luna设备实现ChatGPT日常自动化的实用方法。
343XX · List09:23实践72
通过循环工作流,让AI自动回顾旧例并执行QA检查,持续优化流程。
344XX · List09:18会议72
don’t think I’ve ever felt fomo for a wedding before
AI名人婚礼办学术讨论会,引发FOMO热议。
345XX · List09:01行业72
Re According to claudex
Claudex相关AI资讯报道,聚焦最新动态与行业影响。
346XX · List15:21实践72
when you realise you were the context rot all along
反思AI时代人类沦为上下文腐化因素的讽刺观点
347XX · List15:19模型72
V4-Flash knows a bit more than "May 2025"
V4-Flash模型知识截止时间更新,超越2025年5月。
348XX · List13:29实践70
my toxic trait: believing i can implement faster inference than everyone else
调侃自己总以为能比别人实现更快的推理优化。
349XX · List13:48实践70
Tibo分享关于好奇心的优化理念,鼓励探索未知。
350XX · List20:36会议65
just woke up from 12h of sleep what did I miss guys in the world of ai
AI圈12小时动态汇总,快速补课指南。
351XX · List18:20实践65
i have thought about the odyssey everyday since i watched it.
作者自述看完《奥德赛》后每日回味,表达作品带来的持久震撼与思考。
352XX · List15:10实践65
Didn’t know Leopold Aschenbrenner was a member of the German Green Party and advocated for UBI. You learn something new every day.
意外发现AI人物Aschenbrenner是德国绿党成员并支持全民基本收入。
353XX · List14:15行业65
Europeans, what can you say for yourselves
欧洲在国防、能源、商品上依赖外部,却仍以居高临下姿态看待贸易伙伴,引发批评。
354XX · List13:17行业65
Don't get blended, anon…
欧洲用户提醒别被中国AI宣传迷惑,称中国更需要欧洲市场。
355XX · List12:14实践65
Serious questions from LessWrong. Can we *SLOW DOWN* training by 1000x bros? Think you can pull that off? First idea: impose limits of 10MB/s bandwidt...
探讨通过限制带宽等极端手段将AI训练速度降低1000倍的可行性。
356XX · List09:22实践65
I can still meter this intelligence you guys
作者调侃AI智能可被量化评估,引发对AI能力测量的思考。
357XX · List08:44产品65
Not sure how George got ahold of my daily mantra...
调侃本地运行三个LLM,劝人买电脑别用云端。
358XX · List05:32实践65
The frontier Confucian lab… awareness spreads 君子 is not exactly “gentleman”, by the way.
探讨儒家“君子”概念与英文“gentleman”的差异,引发文化翻译思考。
359XX · List04:31行业65
The most entitled breed of elitist is going to whine the loudest: the pure mathematician. They'll be worse than R18 artists; for like artists, they th...
纯数学家将因AI冲击而哀嚎,但不过是猿类回归本质。
360XX · List03:48实践65
作者对“永久底层阶级”话题有强烈共鸣,配图表达观点。
361XX · List23:54会议65
Rare photo of Yann Lecun on his way to find the datapoints this “less-than-a-cat-level intelligence” supposedly plagiarized the 10 solutions from
调侃LeCun寻找被指抄袭的数据点,讽刺AI能力争议。
362XX · List23:38研究65
the only kind of maths my smol smooth brain might care about these days is the non-commutative non-associative bounded-dynamic-range kind
探讨非交换非结合有界动态范围数学,可能是AI相关新思路。
363XX · List23:34模型65
good chance that the leap from 5.4 to mythos will be equivalent to the leap from mythos to astra (heard mythos 2 also done a few weeks ago?)
预测模型版本迭代跨度,提及Mythos 2已完成。
364XX · List21:22实践65
And that's why i recently made a google account for my son, so that he can book slots on my calendar. I love technology!
家长为孩子创建谷歌账号预约日程,感叹科技便利。
365XX · List20:24实践65
[on first date] me: so have you been following the leopold hedge fund collapse? her: me: wait, where are you going..?
用约会冷场梗调侃AI话题不合时宜,幽默中带点尴尬。
366XX · List19:54实践65
time to repost old shit and enjoy the takeoff
转发旧帖,看好o1模型迈向科学超级智能,最终助力AGI。
367XX · List19:05产品65
用Codex部署子代理分析追踪,戏称“追踪取证代理”。
368XX · List15:48会议65
thank you Jensen Huang and people nitpicking over Twitter vs. X for the substantial payout this period :)
感谢黄仁勋入驻推特及网友讨论带来丰厚收益。
369XX · List15:31实践65
评论特朗普从第一性原理处理国际事务,无视常规的直率风格。
370XX · List15:30实践65
探讨同一持久问题为何重演,反思过往忽视的教训。
371XX · List15:15实践60
pineapple king bakery
讨论旧金山最佳糕点,排除Tartine后的推荐。
372XX · List02:24行业60
探讨政治中贤能晋升问题,附带AI相关图片。
373XX · List23:42实践60
When you’re not a “select customer” for over-caffeinated Sol 😔
Sol咖啡因过度,非精选客户遭冷落,引发吐槽。
374XX · List12:20实践45
吐槽Claude玩麻将很差,喊话Anthropic来聊。
375XX · List11:33实践45
调侃OpenAI,建议开发反重力滑板等科幻产品。
376XX · List09:30实践45
作者删除嘲讽Gary Marcus的推文,认为其AI观点常失准,但不必过度攻击。
377XX · List17:32实践45
用户吐槽购买Max 20x套餐后个人推理开销大,感到后悔。
378XX · List11:12行业40
作者宣布加入Torment Nexus公司,将主要负责“折磨”方面的工作。
379XX · List03:33行业40
«heh. I've destroyed your oil production. No more export revenue for ya, rusnya pigdog» «Ah, like that?! Imma bomb your grain terminals, hohol nazi...
俄乌互炸能源设施,视频展示冲突升级。
380XX · List21:07实践30
芬兰出境边检效率极低,令人惊讶。
381XX · List04:03实践30
文章批评犹太复国主义者缺乏道德论证,以文明规模进行挑衅和转移话题。
382XX · List03:45行业30
作者晒出新车和Mac mini,吐槽滑雪装备不实用。
383XX · List09:37行业30
The Chinese term for Trump always cracks me up
调侃特朗普中文译名“懂王”的趣味帖,附带对美对华石油政策的评论。
384XX · List12:59行业0
微软高管离职祝福,非科技内容。
385XX · List09:39行业0
内容为攻击性言论,无实质科技信息。
386XX · List21:02行业0
good morning!