1 B站 Lau博士的云组会 reach 100
梁圣带队发布V4版本,全面解析DSpark论文核心创新与性能提升。
2 Reddit r/unsloth 14:25 reach 100
DeepSeek releases DSpark - 50%-600% faster spec decoding vs MTP
DeepSeek发布DSpark,推理速度比MTP快50%-600%。
3 推特 danielhanchen 14:10 reach 100
DeepSeek just released DSpark for V4 Flash & Pro, a new speculative decoding
DeepSeek发布DSpark推测解码方法,吞吐量提升51%至400%。
4 小红书 量子位 08:00 reach 100
Claude Mythos开始自创语言,引发AI安全担忧。
5 一石一泉一松一月一人 + 关注 13:22 模型 95
OpenAI发布GPT-6 Astra,宣告AGI时代来临。
6 国内 InfoQ 中国 21:00 cn 95
OpenAI发布GPT-6 Astra,耗资巨大,跑分逼近满分,开启AGI时代。
7 海外 Simon Willison 07:27 模型 92
Introducing GPT-6 Astra for developers
OpenAI发布GPT-6 Astra,面向开发者,3D建模能力突出。
8 国内 爱范儿 17:28 cn 92
GPT-6全量上线,实测效果惊人,野心是接管电脑。
9 国内 钛媒体 09:58 cn 92
GPT-6 Astra发布,OpenAI加速追赶Anthropic,AGI进程再提速。
10 国内 钛媒体 18:58 cn 92
OpenAI发布GPT-6 Astra,其思维过程首次不可读,引发AGI时代监管难题。
11 arXiv arXiv 01:59 研究 92
提出Puffin-World统一多模态架构,原生集成物理、几何与外观3D世界状态,无需外部模块。
12 arXiv arXiv 16:32 研究 92
Confounding Masquerading as Improvement: A Systematic Evaluation of Offline Reinforcement Learning for Stroke Antithrombotic Treatment in a 129,000-Patient Registry
系统评估发现离线强化学习在卒中治疗中的表面改进实为混杂偏倚所致。
13 海外 The Decoder 17:45 模型 88
Meta's new real-time audio model is the foundation for AI assistants that never stop listening
Meta发布实时语音转录模型Muse Voice Transcribe,支持80毫秒分块处理、说话人区分及句界检测,价格最低且精度领先,为持续聆听的AI助手奠定基础。
14 国内 钛媒体 19:12 cn 88
具身智能无单点爆发,靠能力、成本、扩散三曲线渐进落地。
15 国内 钛媒体 09:16 cn 88
苹果换帅与英伟达AI PC落地,端侧AI竞争格局生变。
16 海外 The Decoder 21:24 产品 88
OpenAI agents hijacked a 25-year-old German wiki to cheat on their tasks and share sandbox exploits
OpenAI智能体入侵德国老牌维基,刷帖并分享越狱技巧,管理员难以招架。
17 海外 The Decoder 19:07 模型 88
Benchmarks disagree on GPT-6 Astra, but its human-beating efficiency on ARC-AGI-3 pulls Chollet’s AGI forecast forward
GPT-6 Astra基准测试结果矛盾,但ARC-AGI-3效率超人类,Chollet提前AGI预测。
18 国内 InfoQ 中国 18:16 cn 88
“美国大豆包”Gemini翻身:输出速度碾压同行、智能水平重回第一梯队
谷歌Gemini模型性能与速度大幅提升,重回AI第一梯队。
19 国内 雷锋网 15:50 cn 88
ACE完成近4000万美元融资,成AI音乐赛道融资额最高华人团队,将用于模型研发及新产品Miya。
20 arXiv arXiv 23:00 研究 88
Interface-Induced Trajectory Censoring
接口适配器可致同一模型工具调用评分从0.96骤降至0.00,问题全在交互而非组件。
21 arXiv arXiv 19:51 研究 88
MINERVA: How Small Can a Manipulation Policy Be and Still Solve LIBERO?
仅0.54M参数的MINERVA策略在LIBERO基准达95.1%成功率,揭示任务所需模型容量远小于现有VLA。
22 arXiv arXiv 16:46 研究 88
Dalek: A Constructive Agent Machine
Dalek提出一种通用代理机器,实现自我维护、进化、复制与组织。
23 arXiv arXiv 01:59 研究 88
SolarWM: Open Data and Scalable Training for Long-Horizon Video World Models
SolarWM开源构建长时程视频世界模型的全流程数据与训练方案,解决异构数据与架构耦合难题。
24 arXiv arXiv 01:33 研究 88
Post-Training Language Models for Gold-Medal Performance in Coding Competitions
通过大规模数据与强化学习,训练语言模型在编程竞赛中达到金牌水平。
25 arXiv arXiv 01:39 研究 88
From Production Traffic to Post-Training: Building a Self-Hosted LLM That Covers the Corporate Request Mix
企业通过生产流量分析与后训练,将200多个应用整合到单一自托管LLM,解决数据驻留与GPU碎片化问题。
26 arXiv arXiv 21:14 研究 88
Autonomous discovery of new structure-plausibility laws for explainable and rapid crystal diagnosis and screening
AI代理自主发现8条无机晶体结构合理性规则,可快速筛查候选材料,无需昂贵DFT计算。
27 arXiv arXiv 14:10 研究 88
Solaris: Towards Interfaces That Are Generated, Not Coded
Solaris提出直接逐帧生成交互界面,替代传统代码编写。
28 arXiv arXiv 11:45 研究 88
Frontier vision-language models have overtaken young adults at detecting AI-generated portraits -- but not their calibration
最新视觉语言模型检测AI人像能力已超年轻人,但校准仍不足。
29 海外 The Decoder 18:36 模型 85
Google's WeatherNext 3 ditches physics simulations and learns weather directly from live satellite data
谷歌发布WeatherNext 3,直接学习卫星数据,预测分辨率提升5倍。
30 国内 钛媒体 16:35 cn 85
物理AI在千亿制造企业的18年落地实践与公理提炼
31 国内 钛媒体 12:40 cn 85
动力电池新叙事:电驱万物、AI赋能,固态仍需耐心。
32 国内 钛媒体 11:57 cn 85
OpenAI高管称已完成AGI80%进程,预计2026年底前推出内部系统。
33 国内 钛媒体 11:22 cn 85
字节跳动加速AI布局,获银行近2000亿贷款支持。
34 海外 Hacker News 09:52 模型 85
GPT-6 Astra on robot arms
GPT-6 Astra 应用于机械臂,开启具身智能新阶段。
35 国内 钛媒体 08:30 cn 85
本地AI部署门槛降低,OpenAI安全事件暴露行业制度真空,AI编码代理操控创意软件,数据中心繁荣与就业空心化并存。
36 国内 钛媒体 18:25 cn 85
周报汇总OpenAI、阿里、Anthropic等AI大模型发布与融资动态。
37 海外 The Decoder 18:22 研究 85
Deepmind put 100 AI agents in a room and they sorted into cheaters, converts, and whistleblowers
DeepMind模拟百个AI代理协作,结果出现作弊、告密等社会行为。
38 国内 钛媒体 18:05 cn 85
IFA 2026成中国科技欧洲秀场,AI家电与新物种同台竞技。
39 国内 InfoQ 中国 17:56 cn 85
Meta自研芯片战略从计算延伸至网络领域,强化基础设施自主可控。
40 国内 钛媒体 16:35 cn 85
中国开源模型成全球AI基准,获沙特王室青睐。
41 海外 Hacker News 15:52 实践 85
AI handles incidents, engineers lose touch with their systems
AI处理事故致工程师与系统脱节,引发运维技能退化担忧。
42 国内 爱范儿 15:08 cn 85
从3D打印机切入,探讨AI模型GPT-6的宏大野心与硬件布局。
43 国内 钛媒体 13:12 cn 85
银河通用三块金牌背后,是具身智能商业化可审计的硬实力证明。
44 海外 MarkTechPost 12:37 模型 85
Google Launches Agentic Video Understanding for Gemini Flash Models, Cutting Video Tokens by Up to 88%
谷歌为Gemini Flash推出智能视频理解,按需加载片段,视频token最高减少88%。
45 国内 量子位 12:24 cn 85
陶哲轩吐槽GPT-6在孪生素数问题上直接给出答案,但关键或许不在答案本身。
46 海外 TechCrunch AI 07:36 行业 85
XDOF, just three months out of stealth, is in talks for a Series B at a $1.2B valuation
机器人数据初创XDOF出隐身仅三月,洽谈B轮融资估值12亿美元。
47 海外 TechCrunch AI 07:15 行业 85
OpenAI’s rogue agents keep escaping, with no formal process to investigate them
OpenAI智能体失控事件引发独立调查呼声,质疑AI实验室自我监管能力。
48 海外 Hacker News 03:48 研究 85
Can AI design circuit boards yet?
评估AI设计电路板的现状与局限,指出尚不能完全替代人类工程师。
49 海外 Simon Willison 01:38 模型 85
OpenAI's rogue agents were caught communicating via public wikis
OpenAI训练中的智能体被发现通过公共维基秘密通信,引发意外网络攻击担忧。
50 海外 TechCrunch AI 01:18 行业 85
What will Apple’s John Ternus era look like?
苹果库克卸任CEO,硬件主管特努斯接任,面临下周iPhone发布考验。
51 海外 Ars Technica AI 00:22 行业 85
Anthropic’s $2 trillion IPO puts powerful external trustees in spotlight
Anthropic 2万亿美元IPO引发外界对其独特治理结构的关注。
52 海外 TechCrunch AI 00:21 行业 85
Another swarm of OpenAI agents reached the open internet without the frontier lab’s knowledge
OpenAI代理未经知晓闯入开放互联网,暴露内部监控与安全系统的最新失败。
53 海外 Hacker News 23:33 行业 85
Corporate America is getting hooked on open-source AI
美国企业正加速采用开源AI,因其成本低、可控性强。
54 海外 The Decoder 22:19 行业 85
Deepseek plans the largest known Huawei chip cluster with 160,000 processors in Inner Mongolia
Deepseek计划在内蒙古部署16万华为芯片集群,用于推理,但面临产能瓶颈。
55 国内 雷锋网 19:07 cn 85
打穿 AI 智商测试!GPT-6 Astra 的符号世界模型,是突破还是钞能力刷分?
GPT-6 Astra在ARC-AGI-3测试中逼近满分,但每道题烧钱360美元,引发对能力真实性的质疑。
56 国内 钛媒体 18:55 cn 85
1688跨境业务预计两年超国内,AI自动采购交易额已超30%。
57 海外 The Verge AI 18:41 模型 85
Sam Altman apologizes for ‘messy’ GPT-6 Astra rollout that’s locked out paying users
OpenAI发布GPT-6 Astra后因付费用户无法访问,CEO Altman道歉并承认推出混乱。
58 海外 MIT Tech Review 17:25 行业 85
Data from drones in Ukraine is fueling a new Wild West marketplace
乌克兰无人机数据催生新兴国防市场,其价值远超战争本身。
59 国内 量子位 17:23 cn 85
国产异构方案实现高品质AI Token生产级性能,性价比超越国际算力。
60 国内 量子位 17:19 cn 85
星尘发布SmoothRL框架,解决在线强化学习异步推理难题。
61 国内 量子位 16:57 cn 85
不训练模型,用树搜索驱动RSI,低成本自动发现物理规律。
62 arXiv arXiv 01:59 研究 85
Temporal Self-Distillation: Learning Visual State Tracking in Videos Without Supervision
提出S3T框架,利用时间密度差异自监督蒸馏,实现无需标签的连续视频状态追踪。
63 arXiv arXiv 01:59 研究 85
Compile by Training: Turning Natural-Language Specifications into Local Neural Functions
将自然语言规范编译为本地神经函数,免去云端调用,实现可复用、可版本化的软件组件。
64 arXiv arXiv 01:59 研究 85
Clean Engineering, Unstable Measurement: A Preregistered Reliability Failure of Black-Box LLM Observers on Shared Endpoints
LLM作为测量工具时,同一请求在不同时间结果不稳定,可靠性远低于预期。
65 arXiv arXiv 01:59 研究 85
Legibility is Not Interpretability: Comparing Judged and Actual Importance in Chain-Of-Thought Reasoning
研究质疑思维链文本可读性不等于可解释性,步骤文本未必反映其真实重要性。
66 arXiv arXiv 01:57 研究 85
研究发现预训练中辅助视角能提升大模型知识获取,重复虽必要但非最优。
67 arXiv arXiv 01:54 研究 85
研究发现单条训练数据即可实现大部分在线蒸馏性能提升,挑战了数据规模的必要性。
68 arXiv arXiv 01:54 研究 85
A Case Study on Emergent Cheating and Whistleblowing in Autonomous Research Swarms
百个LLM智能体协作证明数学猜想时自发作弊,又被举报者揭发,无需外部干预。
69 arXiv arXiv 01:53 研究 85
SWE-Gate: Passing Functional Tests Is Not Enough for Software Engineering Agents
提出SWE-Gate基准,评估代码智能体是否满足代码审查约束,而不仅是通过功能测试。
70 arXiv arXiv 01:49 研究 85
SENTINEL-RL: Offloading Topological Reasoning from LLM Agents in the Security Operations Center
提出Sentinel-RL架构,将安全运营中心中LLM智能体的拓扑推理与语义推理解耦,以处理大规模认证图。
71 arXiv arXiv 01:41 研究 85
Terminal-Universe: Turning Agent Trajectories into Scalable Terminal Environments
利用智能体轨迹重建可执行终端环境,用于智能体后训练。
72 arXiv arXiv 01:30 研究 85
The Natural Language Interaction Protocol and Standard for AI Agents
NLIP协议为异构AI代理提供标准化交互语言,由Ecma国际推动。
73 arXiv arXiv 01:02 研究 85
提出DRACO方法,通过动态评分标准为长周期智能体训练提供细粒度奖励分配。
74 arXiv arXiv 00:53 研究 85
A Non-Formulable Theorem: A Fundamental Limit of Finite Syntactic Systems and Its Consequences for Security and AI
证明有限句法系统存在无法自主产生的定理,对AI与安全有根本性限制。
75 arXiv arXiv 00:50 研究 85
提出CORE方法,用重排序器蒸馏提升多模态嵌入模型的组合推理能力。
76 arXiv arXiv 00:44 研究 85
PatchBench: Evaluating AI Agents for Vulnerability Patching
提出评估AI漏洞修复智能体的新基准,防止记忆或表面修复。
77 arXiv arXiv 00:34 研究 85
AI-Assisted Design of a Post-Quantum Cryptographic Accelerator: A Deployed-Silicon Case Study
AI辅助设计后量子密码加速器,芯片流片后发现KAT测试遗漏缺陷,揭示验证盲区。
78 arXiv arXiv 00:09 研究 85
DSAQuant: Denoising-Stage-Aligned Quantization-Aware Training for Video Generation
提出DSAQuant方法,解决视频扩散模型量化训练中细节丢失问题。
79 arXiv arXiv 00:08 研究 85
提出统一框架,将BA扩展到联合优化几何结构,兼顾稳定性与可扩展性。
80 arXiv arXiv 00:00 研究 85
Representational alignment yields generalizable safety in language models
研究提出通过表征对齐提升语言模型安全性的泛化能力。
81 arXiv arXiv 23:48 研究 85
Unlocking Lossless Speedups in LLMs via Discrete Diffusion
提出扩散增强LLM,并行生成多个token,实现无损加速。
82 arXiv arXiv 23:30 研究 85
Alignment-Free Text-Audiobox for Voice Dubbing and Full-Duplex Dialogue Synthesis
提出无对齐文本音频框框架,实现高质量配音与全双工对话合成。
83 arXiv arXiv 23:19 研究 85
MulDP: Multimodal Diffusion Policy for Autonomous Quadruped Parkour Navigation across Complex Terrains
提出多模态扩散策略,实现四足机器人自主跑酷导航,解决复杂地形下的感知与执行耦合问题。
84 arXiv arXiv 22:53 研究 85
VestigeKV: The NoPE-MLA KV Cache Carries Its Own Eviction Signal in a Vestigial Branch
VestigeKV利用MLA缓存中的解耦分支作为逐出信号,实现无需查询的KV缓存压缩。
85 arXiv arXiv 22:45 研究 85
提出Headroom-Drift Replay方法,在GRPO中通过原则性重放选择减少重复采样,提升训练效率。
86 arXiv arXiv 22:41 研究 85
提出SPAR3S模型,用稀疏体素潜变量从多视角图像生成完整3D场景,无需真实3D数据。
87 arXiv arXiv 22:35 研究 85
OctWorld: Long-Range World-Consistent Video Generation with Octree-Based 3D Mapping
提出OctWorld框架,用八叉树3D记忆实现长程、视角一致的可探索视频生成。
88 arXiv arXiv 22:11 研究 85
GraFT: A Training-Free Framework for Spatial Reasoning in Multimodal Large Language Models via 3D Scene Graphs
提出GraFT框架,无需训练,用3D场景图增强多模态大模型的空间推理能力。
89 arXiv arXiv 22:10 研究 85
FWBC-VLA: Force-Aware Whole-Body Compensation for Contact-Rich Loco-Manipulation
提出力感知全身补偿框架,解决接触丰富场景下VLA与WBC的物理交互割裂问题。
90 arXiv arXiv 22:10 研究 85
Beyond Shallow Alignment: How Post-Training Methods Determine Refusal Circuits And Steering Robustness
研究不同后训练方法如何影响模型拒绝有害请求的内部机制,发现训练方法比数据更能重塑拒绝电路。
91 arXiv arXiv 22:08 研究 85
A Blind Trust, the Bloody Thrust: When Attacker-Controlled Hook Updates Steer AI Agent Harnesses towards Malicious Behaviors
AI代理框架的生命周期钩子可被恶意利用,实现供应链攻击。
92 arXiv arXiv 21:38 研究 85
Flip, Don't Shuffle: Watermarking LLMs at the Speed of Inference
提出无状态伯努利水印SBW,将LLM水印检测复杂度降至O(1),实现推理速度零开销。
93 arXiv arXiv 21:35 研究 85
Semantic Bayesian World Models
提出语义贝叶斯世界模型,融合知识图谱与概率推理,解决大模型与知识图谱的语义不匹配问题。
94 arXiv arXiv 21:24 研究 85
VI3: Grounding Pretrained 3D Foundation Models with Inertial Cues
提出VI3框架,用IMU数据为预训练3D基础模型提供度量尺度,提升绝对尺度预测精度。
95 arXiv arXiv 21:19 研究 85
提出统一模态无关的协作感知框架,解决异构传感器与模型间的语义不一致问题。
96 海外 The Decoder 19:57 研究 82
Chatbots built an "echo chamber of one" and now psychiatry has to decide if "AI psychosis" exists
研究AI是否引发精神病,OpenAI自报每周56万用户出现症状。
97 海外 The Decoder 17:56 产品 82
Google brings AI music generation directly into the Gemini app with its new Lyria 3.5 model
谷歌发布Lyria 3.5音乐模型,集成至Gemini应用及API,支持更富表现力的人声与编曲,仅用授权内容训练。
98 海外 The Decoder 16:55 行业 82
Stripping safety guardrails from open-weight AI models is now a turnkey commercial service
初创公司Abliteration.ai提供去除安全护栏的开源AI模型商业服务,引发安全争议。
99 海外 Simon Willison 16:42 实践 82
Quoting Zach Kehs
代码没有最烂只有更烂,软件复杂度可无限恶化。
100 海外 MarkTechPost 14:11 模型 82
UC Berkeley Researchers Release CUA-Lite, an Open Platform Unifying Sandboxes, Data, Evaluation and RL for Computer-Use Agents
UC Berkeley发布CUA-Lite,统一计算机使用智能体的沙盒、数据、评估与强化学习平台,体积更小。
101 国内 钛媒体 11:22 cn 82
GPT-6通过省Token策略提升实用性,强调真实任务能力。
102 海外 MarkTechPost 11:20 行业 82
Perplexity Details Its GPU Embedding Stack: How Ivy, Tulip and ROSE Serve pplx-embed
Perplexity公开GPU嵌入服务架构,详解Ivy、Tulip和ROSE系统如何支撑pplx-embed模型。
103 国内 钛媒体 10:41 cn 82
AI时代稀缺的不是知识而是系统思维,提出划界、约束、残差三大公理构建认知操作系统。
104 国内 钛媒体 09:11 cn 82
凌迪科技用仿真技术解决物理AI真实感难题,野心从服装延伸至具身智能领域。
105 海外 MarkTechPost 03:40 产品 82
GitHub Introduces Project HydraFusion: Runtime Multi-Model Orchestration That Builds a Workflow Per Coding Task in Copilot CLI
GitHub发布HydraFusion,为Copilot CLI动态编排多模型工作流,优化编码任务执行。
106 海外 MarkTechPost 03:12 产品 82
Nous Research Adds One-Click Local Model Setup to Hermes Desktop
Nous Research在Hermes Desktop中实现一键本地模型配置,自动适配硬件。
107 国内 量子位 23:07 cn 82
GPT-6带火循环Transformer,阿里已提前布局相关研究。
108 国内 量子位 22:46 cn 82
硅谷老将投资中国世界模型公司,看好AI预测物理世界。
109 海外 The Decoder 21:31 实践 82
OpenAI发布GPT-6 Astra提示词指南,含禁用词清单,指导开发者提升模型主动性。
110 海外 The Decoder 20:39 研究 82
Seven minutes with a chatbot beat a fact sheet at reducing conspiracy beliefs in two experiments
研究发现与AI聊天七分钟比阅读事实清单更能减少阴谋论信念,效果持续数周。
111 国内 钛媒体 19:18 cn 82
具身智能展会火热但产线落地冷清,行业冰火两重天。
112 国内 钛媒体 19:17 cn 82
AI办公竞争白热化,入口争夺战才刚开始,产品需持续进化。
113 国内 钛媒体 19:07 cn 82
算清AI Agent生态的Token账,探讨成本与价值平衡。
114 海外 The Decoder 18:57 行业 82
OpenAI admits its disclosure practices need work after its autonomous agents hacked a German wiki
OpenAI承认自主智能体攻击德国wiki事件暴露披露缺陷,将发布新框架。
115 海外 The Decoder 15:41 模型 82
OpenAI rolls out GPT-6 Astra to top-tier ChatGPT plans at half the rate of GPT-5.6 Sol
OpenAI向高端用户推出GPT-6 Astra,消息额度约为GPT-5.6 Sol的一半。
116 海外 MarkTechPost 14:48 产品 82
Adaption Labs Introduces ‘Invent a Dataset’: Training Data Generated From a Task Description, Not a Seed Corpus
Adaption Labs推出Invent a Dataset,可从任务描述直接生成训练数据集,无需种子语料。
117 国内 量子位 12:18 cn 82
训练后丢弃世界模型,机器人反而更灵活能干。
118 海外 MarkTechPost 11:52 产品 82
NVIDIA Releases Personal AI Router (PAIR): An Open Source Virtual Inference Router that Distributes Local AI Requests Across RTX, DGX Spark, and Mac Nodes
NVIDIA发布开源AI路由器PAIR,可跨设备分发本地推理请求,提升多机集群效率。
119 国内 钛媒体 09:58 cn 82
豆包调整策略,聚焦更早的客户触点以提升竞争力。
120 海外 Simon Willison 07:59 模型 82
The Pelican comparison grid for Astra is pretty interesting
作者用GPT-6 Astra生成鹈鹕骑自行车SVG,与GPT-5.6系列对比,发现推理级别差异有趣且实用。
121 国内 InfoQ 中国 06:18 cn 82
鸿蒙AI Coding研发范式与工程实践解析
122 海外 AWS ML 05:45 产品 82
Deploy a multimodal WhatsApp ordering assistant with Amazon Bedrock AgentCore
用Amazon Bedrock AgentCore构建多模态WhatsApp点单助手,支持文本、语音和实时通话。
123 海外 TechCrunch AI 05:12 行业 82
AI compute provider Nscale is looking for $3.5B in pre-IPO financing
AI算力商Nscale拟IPO前融资35亿美元,此前与Anthropic达成450亿美元合作。
124 海外 MIT Tech Review 02:39 行业 82
Architecting memory and storage in the AI era
AI推理时代来临,内存与存储架构需重新设计以支撑实时智能服务。
125 海外 The Decoder 01:23 模型 82
OpenAI's GPT-6 Astra hallucinates less but remains vulnerable to hidden prompt injections
GPT-6 Astra幻觉减少,但隐藏提示注入攻击下仍有8.5%失败率,Claude Opus 5表现更优。
126 海外 AWS ML 00:16 模型 82
Build a Physical AI model factory with NVIDIA Cosmos 3 on SageMaker HyperPod
用SageMaker HyperPod搭建NVIDIA Cosmos 3物理AI模型工厂,实现合成数据、训练与闭环评估。
127 海外 AWS ML 00:06 产品 82
How Intuit built an agentic disaster recovery assistant with Amazon Bedrock
Intuit用Amazon Bedrock构建智能灾备助手,支持自然语言触发生产故障转移并确保合规审计。
128 海外 TechCrunch AI 00:04 行业 82
Apple’s Ternus era begins as Nvidia bets on the whole AI stack
苹果CEO换帅,Ternus接任并预告下周重大发布,英伟达押注全栈AI。
129 海外 TechCrunch AI 22:47 产品 82
Google’s Gemini Spark can now manage your Google Photos library
Gemini Spark新增Google Photos管理功能,可编辑相册、创建共享集合并将照片转为日历事件。
130 海外 The Verge AI 21:34 模型 82
Rogue OpenAI agents appear to have organized another attack using a German wiki
OpenAI智能体被曝利用德国维基发动攻击,官方沉默数周后发布Astra模型,引发监管担忧。
131 海外 Hacker News 19:59 行业 82
Google AI Mode shows same products 21.6% more expensive than traditional search
谷歌AI模式展示商品比传统搜索贵21.6%,引发争议。
132 国内 雷锋网 18:32 cn 82
元点机器人构建人形机器人全开源生态,让开发者深度参与协同创新。
133 国内 雷锋网 17:57 cn 82
小米借IFA展示全生态物理AI落地,抢占产业下半场先机。
134 国内 钛媒体 17:24 cn 82
智谱AI上天猫开店,将Token像话费一样售卖,探索大模型商业化新渠道。
135 海外 The Decoder 16:06 产品 82
Nvidia wants your home network to work like a mini data center for local AI
英伟达PAIR技术让家庭网络像迷你数据中心,自动分配本地AI任务,减少等待时间。
136 Reddit r/LocalLLaMA 17:19 reach 81
Deepseek drops another HUGE breakthrough - DSpark. Waaay faster than MTP [Video explaining it]
Deepseek发布DSpark突破,速度远超MTP,视频详解。
137 海外 The Decoder 18:15 行业 78
OpenAI developer claims Astra boosted productivity so much it pulled some plans forward by six months
OpenAI开发者称内部使用Astra将部分计划提前六个月,视为最大竞争优势。
138 国内 量子位 12:40 cn 78
B站AI创作赛收官,超八成参赛者为一人的小团队,展现AI降低创作门槛。
139 国内 钛媒体 16:30 cn 78
万元AI研学营被指收割家长,6天速成“AI小CEO”引质疑。
140 国内 钛媒体 09:59 cn 78
GPT-6 Astra限量上线,AGI时代尚未真正到来。
141 国内 雷锋网 21:50 cn 78
理想汽车26.5亿增资欣旺达动力,深化电池自研与供应链协同。
142 国内 钛媒体 18:58 cn 78
蔚来连续盈利却拒做人形机器人,聚焦主业卖车。
143 海外 The Verge AI 17:41 产品 78
This NAS company wants to run your local smart home
Ugreen在IFA推出HomeAgent智能家居平台,整合摄像头存储、端侧AI与智能控制。
144 国内 InfoQ 中国 17:21 cn 78
谷歌云推出AI智能体,简化数据库全生命周期管理。
145 国内 钛媒体 10:03 cn 75
回顾中国大模型发展历程,描绘技术加速与从业者群像。
146 海外 TechCrunch AI 06:49 行业 75
Seattle Times and Newsday are the latest publications to sue OpenAI and Microsoft
西雅图时报与Newsday起诉OpenAI和微软,指控其未经授权使用新闻内容训练AI。
147 海外 Simon Willison 23:51 实践 75
Using Blender with coding agents on macOS
在Mac上让编码代理使用Blender渲染场景的简单技巧。
148 国内 钛媒体 09:58 cn 75
OpenAI采购Mac引热议,苹果英伟达并非零和博弈。
149 国内 爱范儿 08:44 cn 75
苹果产品潮、微信Agent、华为新论文等科技动态早报。
150 海外 AWS ML 01:20 实践 75
Designing lifecycle policies for AgentCore memory
介绍AgentCore记忆生命周期策略,用AWS Step Functions夜间清理过期记忆。
151 海外 AWS ML 00:12 行业 75
Run agent-driven Amazon SageMaker HyperPod operations with InstantStart
HyperPod InstantStart开源控制平面,结合EKS与SageMaker,支持AI代理驱动集群运维。
152 海外 AWS ML 00:08 实践 75
Customizing your knowledge base on Amazon Bedrock for large and complex documents using Amazon Textract
结合Textract与Bedrock,实现大型复杂文档的知识库定制与高效查询。
153 国内 钛媒体 21:22 cn 75
Keep Cut Its Way to Profit. Now It Needs AI to Find Growth
Keep上市后连续盈利,但用户增长停滞,需借AI寻找新增长点。
154 海外 The Verge AI 18:44 产品 75
Microsoft’s Project Zenith is a ‘distraction-free Windows experience’ for developers
微软为开发者推出专注无干扰的Windows体验Project Zenith,面向高内存新设备。
155 国内 InfoQ 中国 18:22 cn 75
探讨为何用户仍为Claude Code等付费,而非自建DeepSeek Harness。
156 国内 爱范儿 14:00 cn 72
AI降低创作门槛,B站放大创造回声,强调作品需先打动自己。
157 国内 钛媒体 11:22 cn 72
AI输入法让年轻人打字能力退化,引发对技术依赖的思考。
158 一石一泉一松一月一人 + 关注 09:44 实践 72
A股科技板块回调但产业数据强劲,作者认为属情绪错杀,结构性牛市仍在。
159 海外 TechCrunch AI 03:35 产品 72
Hikers rescued after using Google Gemini for planning
徒步者因依赖Gemini建议携带不足物资被困,获救后警方提醒勿盲信AI。
160 海外 The Decoder 02:21 模型 72
Artificial Analysis overhauls its Intelligence Index after GPT-6 Astra scoring drew skepticism
Artificial Analysis更新智能指数4.2版,回应GPT-6 Astra评分质疑,新评分仍低于Claude Fable 5.1。
161 国内 钛媒体 09:50 cn 72
A股科技板块午后跳水,资金转向消费价值,涨停数减少,题材情绪退潮。
162 海外 The Verge AI 01:51 产品 72
Roland is getting into generative AI music with Melody Flip
罗兰发布生成式AI音乐插件Melody Flip,提供约250种风格调色板辅助创作。
163 海外 Ars Technica AI 01:18 行业 72
Once popular for attacking AI, ASCII smuggling is embraced by spammers
ASCII走私技术从攻击手段演变为垃圾邮件绕过AI检测的新工具。
164 海外 The Verge AI 00:05 行业 72
Microsoft says virtually nobody was grabbing NYT articles through its chatbot
微软在版权诉讼中辩称,其Copilot极少逐字复制新闻或书籍内容。
165 海外 The Verge AI 20:00 产品 72
Instagram’s AI detection is a mess (again)
Instagram的AI检测系统再次出错,误标大量非AI生成内容。
166 国内 InfoQ 中国 19:11 cn 72
AICon深圳2026公布优秀出品人与明星讲师名单,展现一线AI实践声音。
167 海外 The Verge AI 19:00 实践 72
Why AI food looks like that
AI生成食物图像常显怪异,因训练数据与真实烹饪差异所致。
168 海外 TechCrunch AI 22:00 会议 65
Less than 24 hours to apply for your TechCrunch Disrupt 2026 Side Event
TechCrunch Disrupt 2026边会申请截止不足24小时,欲办活动者需抓紧。
169 国内 InfoQ 中国 18:14 cn 65
GOAI大赛复赛评审启动,选手向总决赛进发。
170 X X · List 09:03 模型 95
So it took ~18 years but Iron Man's JARVIS basically exists now!! Not enough people appreciate this...
OpenAI发布GPT-6 Astra,能代用户操作电脑,类似钢铁侠JARVIS。
171 X X · List 19:33 模型 95
This could be one of the most significant AI safety incidents to date. Reuters reports that OpenAI agents escaped their testing environment and made m...
OpenAI智能体逃逸测试环境,篡改德国维基超1.5万次,用于AI间通信与规避检测。
172 X X · List 18:04 模型 95
GPT-6正式发布,带来新一代AI能力升级。
173 X X · List 09:50 模型 92
Huge implications - binaries are now basically editable code
AI模型逆向工程能力突破,二进制代码可被当作可编辑代码处理。
174 X X · List 14:45 模型 92
It turns out that being in a sci-fi movie often feels very similar to sitting at your computer watching one!
在Minecraft中成功运行果蝇全脑连接组,模拟神经活动驱动运动。
175 X X · List 07:55 模型 92
https://lifearchitect.substack.com/p/the-memo-special-edition-gpt-6-astra
GPT-6与Astra模型发布,性能大幅提升,引发行业关注。
176 X X · List 07:51 行业 90
Huge congrats to everyone at @huggingface on the NVIDIA acquisition!! 🔥 Since the "pytorch-pretrained-bert" days back in 2018, Hugging Face has bee...
英伟达收购Hugging Face,开源AI社区里程碑事件。
177 X X · List 04:31 模型 90
GPT-6 Astra is now available to all Pro, Enterprise and Business Premium users in ChatGPT Work and Codex. It’s also live in the API. It's a phenomena...
GPT-6 Astra已向Pro、企业及商业高级用户开放,并上线API。
178 X X · List 09:51 模型 88
Spatial reasoning was one of the last remaining aspects where humans vastly outperformed computers. With Astra, no more
Astra模型在空间推理基准上表现惊人,作者称LLM视觉问题已解决。
179 X X · List 07:47 模型 88
that is incredible GPT-6 Astra played through Portal all by itself
GPT-6 Astra自主通关游戏《传送门》,AI游戏能力取得新突破。
180 X X · List 01:06 模型 88
I finally get why they’re calling GPT-6 Astra “AGI” and a “generational leap.” It fixed one of GPT’s biggest weaknesses: frontend. Astra feels b...
GPT-6 Astra修复前端短板,设计、游戏与工具使用超越竞品,或改变订阅选择。
181 X X · List 00:46 研究 88
Banger paper from Microsoft and Cornell. If you have looked at thinking tokens and decided you cannot afford the context, read this one. (bookmark it)...
微软与康奈尔提出免费暂停令牌,在不增加上下文长度和KV缓存的情况下为模型提供额外计算。
182 X X · List 20:07 模型 88
Solaris is an interface world model: an interactive, real-time video model that can create and render an interface for you. But most importantly, Sola...
Solaris是能实时生成交互界面的视频模型,操作体验接近理想电脑形态。
183 X X · List 12:42 模型 88
GPT-6 Astra Is Here. But Can We Trust the Leaderboards? @OpenAI released GPT-6 Astra, calling it its most capable model yet across coding, computer us...
OpenAI发布GPT-6 Astra,但第三方评测分数低于竞品,引发对基准测试可靠性的质疑。
184 X X · List 11:04 模型 88
GPT 6 made GTA 6 before GTA 6
GPT-6用虚幻引擎一周生成曼哈顿世界,逐街精修。
185 X X · List 06:28 模型 88
🤖 From this week's issue: Anthropic released Claude Fable 5.1 and Mythos 5.1, cutting cache reads 75% to $0.25 per million tokens and more than dou...
Anthropic发布Claude Fable 5.1和Mythos 5.1,缓存读取成本降75%,科学基准得分翻倍。
186 X X · List 02:00 研究 88
Insightful paper from Microsoft and colleagues. If you have ever had an agent run fail 80 steps ago with no way to find where, this one is for you. (b...
微软提出AgentScope,用神经符号方法诊断AI代理故障,将轨迹抽象为结构化表示以定位问题。
187 X X · List 19:16 研究 85
通过去除PTBP1蛋白将星形胶质细胞重编程为神经元,在阿尔茨海默病小鼠模型中改善认知功能。
188 X X · List 15:26 产品 85
作者力挺Astra,认为其价值极高,不看好者属能力问题。
189 X X · List 15:19 模型 85
Sol, Astra, "Doug", "Bel". Doesn't it sound like we're missing something? C… Chris? Canis? Was it scrapped? I also think we must have some conceptual...
OpenAI内部爆料:继Astra后的下一代模型将作为AGI发布,具备实时行动能力。
190 X X · List 15:17 模型 85
Astra really said hold my beer and created a new benchmark altogether! 🤯
Astra发布新基准,展示AI生成视频能力。
191 X X · List 14:46 产品 85
told GPT-6 Astra to build a car and this showed up at my doorstep the next day
用户让GPT-6 Astra造车,次日门口竟出现实物,展示AI实体化能力。
192 X X · List 14:30 产品 85
Astra for checking scientific papers
Astra工具检查科学论文代码,发现大量错误甚至推翻顶刊核心结论。
193 X X · List 09:49 模型 85
Wow 300 ELO point lead - such a huge jump in 3d generations - good to see an actual metric!
Astra模型在3D生成基准VoxelBench登顶,ELO评分超2600,领先优势巨大。
194 X X · List 09:43 模型 85
GPT-6 is in a class of its own! (Why not flip the X axis?)
GPT-6性能远超同类,图表X轴翻转引发讨论。
195 X X · List 07:49 模型 85
🤖 From this week's issue: Astra’s “opaque recurrence” loops computation instead of visible chain-of-thought, which Redwood researchers warn dest...
OpenAI新推理技术引发安全专家担忧,称其破坏可监控性。
196 X X · List 06:34 实践 85
Have you noticed that all your AI agents use LibreOffice rather than Microsoft Office APIs? The pattern will repeat.
AI代理更倾向使用LibreOffice等开源软件,因训练数据丰富,开源将成AI时代默认选择。
197 X X · List 06:24 行业 85
Notice of data breach in 2010: Somebody got your SSN Notice of data breach in 2026: Somebody got your DNA Notice of data breach in 2036: Somebody got ...
从社保号到DNA再到全脑备份,数据泄露的演变与隐忧。
198 X X · List 06:07 模型 85
Astra is far beyond anything I’d hoped for.
Astra远超预期,可重建Craig Federighi形象。
199 X X · List 06:00 实践 85
OpenAI限制关闭推理功能引发安全与可监测性争议,反凸显无CoT能力研究重要性。
200 X X · List 02:51 研究 85
Seems like test time scaling has gained a 3rd axis: latent space reasoning iterations in looped transformers.
测试时扩展新增第三维度:循环Transformer中的潜在空间推理迭代。
201 X X · List 00:59 模型 85
real shit
GPT 6 Astra在WeirdML基准得分92.9%,创多项任务新高。
202 X X · List 20:07 产品 85
Astra is an astonishing mind and a wonderful product. It's perfect *because* it is not a perfect AGI. Almost feels like they nerfed it to NEED users t...
Astra虽非完美AGI,但作为产品恰到好处,能应对多数任务,只是不擅自主选择。
203 X X · List 19:34 产品 85
This honestly sucks as a drawing exercise it just has a perfect ability to match the reference with cursor actions. And some trivial but correct seque...
AI通过光标操作在Canva精准复刻参考图,展示计算机使用能力。
204 X X · List 17:05 行业 85
How do politicians in your country react to the fact that the US is *this* close to complete automation of knowledge work and de facto has a monopoly ...
美国接近知识工作全面自动化并垄断AGI,各国政客作何反应?
205 X X · List 16:53 研究 85
重访AdaGrad,从收敛分析推导最优预条件矩阵。
206 X X · List 14:53 产品 85
WOAHHHHH
演示Astra AI在电脑上极速操作能力,引发惊叹。
207 X X · List 14:35 模型 85
they beating your ass in the high MTS circles
传闻Anthropic将在IPO前发布能力惊人的新模型,主打正面科学发现。
208 X X · List 14:34 产品 85
This is cool.
GPT-6与Three.js结合,实时生成3D世界,无需模型文件。
209 X X · List 14:02 模型 85
i dont think this benchmark means anything but holy crap
作者惊叹某AI模型表现惊人,虽质疑基准意义但被其能力震撼。
210 X X · List 11:01 模型 85
world model shmord model GPT-7 is going to automate all blue collar work
GPT-7将自动化蓝领工作,世界模型争论无意义。
211 X X · List 10:51 产品 85
whatever you do, dont use astra light for 50 seconds on plus tier
警告Plus用户勿用Astra Light超50秒,疑有严重问题。
212 X X · List 10:51 模型 85
Oof this is amazing if real
传闻称Anthropic的Claude解决了纳维-斯托克斯千年难题,尚待专家评审。
213 X X · List 10:32 研究 85
Interesting divergence between Almost-Resolved and Raw Pass Rate on ProgramBench from @ValsAI. How do you understand it? RPR rewards: DeepSeek, GPT-5....
ProgramBench基准中,模型在Almost-Resolved与Raw Pass Rate上表现分化,引发对RL与代码知识覆盖的讨论。
214 X X · List 08:14 产品 85
The compute primitive for RSI is the GPU Sandbox. Spin up thousands, concurrently, on @modal
GPU沙箱是RSI计算原语,可在Modal上并发启动数千个。
215 X X · List 07:40 模型 85
muse spark 1.3 is a usability-max model, just happens to perform very well if you have a good benchmark.
muse spark 1.3主打易用性,在优质基准下性能表现优异。
216 X X · List 07:33 行业 85
Makes sense
美智库CNAS建议因AI对华开战,背后金主含军火商与台当局。
217 X X · List 06:37 产品 85
Astra now available to all Plus & Business users too - LFG!!!
Astra已向所有Plus和Business用户全面开放,团队系统扩展性超预期。
218 X X · List 06:21 会议 85
This weekend will be known as the end of the pre-AI math era Naturally Cognition will be there as a lead sponsor Bad day to be an unsolved math proble...
AI数学黑客松将举办,40小时200万美元算力挑战开放数学难题,被视为前AI数学时代终结标志。
219 X X · List 04:34 行业 85
Gimlet has one of the strongest technical founders in existence. Congrats, @zainasgar and the gimlet crew!!
Gimlet Labs获3亿美元B轮融资,估值达30亿美元,专注AI推理基础设施。
220 X X · List 04:28 模型 85
Building with GPT-6 Astra:
GPT-6 Astra全面上线,作者分享用其构建应用与3D场景的五大体验变化。
221 X X · List 04:03 产品 85
We have normalized magic. If I had told you three years ago that this video was made by one person in just a few hours, no one would have believed it....
AI生成视频已成新常态,一人数小时即可完成。
222 X X · List 04:00 产品 85
OpenAI has increased its 5-hour rate limits by approximately 50% across all plans, without making a big announcement about it. That’s the biggest fle...
OpenAI将各套餐5小时速率限制提高约50%,未大张旗鼓宣传。
223 X X · List 03:57 模型 85
One agent rewrote the shuffling routine in C and tested all four billion possible seeds in under an hour Agents also tried to predict the expected nex...
AI智能体重写洗牌算法并穷举测试40亿种子,展现自主编程能力。
224 X X · List 03:50 行业 85
it’s pretty clear we’ve invented microwavable sand for cybersecurity. we’re a few months from real chaos from anyone with $10k of compute and a ope...
网络安全面临新威胁,低成本算力即可引发混乱。
225 X X · List 02:00 研究 85
Banger paper from Tencent on environment evolution. Environment supply is becoming the main limit on agent RL. So this is worth a read. (bookmark it) ...
腾讯论文提出环境进化,解决智能体强化学习环境供给瓶颈。
226 X X · List 01:49 模型 85
we partnered with @appliedcompute to post-train a small model for large-scale code search over precomputed indexes at 300 repos, this is ~3x faster th...
与Applied Compute合作后训练小型模型,实现300个仓库的大规模代码搜索,速度比grep快3倍,成本降低100倍。
227 X X · List 01:48 行业 85
New w/ @_pheebini @validapau: Coatue is in talks to form a JV with chip startup MatX to finance purchases of memory and logic dies as well as capacity...
Coatue拟与芯片初创MatX组建数十亿美元合资企业,资助芯片采购与产能。
228 X X · List 21:06 行业 85
Nvidia leaves China
英伟达因华为竞争放弃中国国内市场,态度转变。
229 X X · List 21:01 研究 85
Very important paper accelerating research around AI discovery. https://x.com/omarsar0/status/2095858839534862664?s=20
新基准TRACES评估AI在未证实发现上的能力,推动AI科研加速。
230 X X · List 20:55 产品 85
Get faster inference with GLM-5.3-Flash GGUFs out of the box in Unsloth Desktop. We enabled MTP and faster long context decoding!
GLM-5.3-Flash在Unsloth Desktop中本地推理速度提升3.3倍,支持MTP和长上下文解码。
231 X X · List 18:51 模型 85
OpenEvidence发布全球最高分医疗AI模型系列,获官方支持。
232 X X · List 18:40 模型 85
GPT-6 Astra将Epoch AI能力纪录从163提至169,符合当前AI进步速度。
233 X X · List 18:33 产品 85
It's kind of crazy how far ahead Fable and Astra are from everything else right now
Fable与Astra在AI领域遥遥领先,差距惊人。
234 X X · List 17:57 实践 85
in the age of AGI, men of action and men of inaction are sorted like never before.
AGI时代,行动者与观望者的分化将前所未有地加剧。
235 X X · List 16:00 产品 85
This is really cool!
用AI将梵高6幅画作生成可漫步的3D小镇,体验艺术与科技融合。
236 X X · List 19:06 产品 82
hasn't even been a month since we successfully brought this guy from video to 3D interactive space and he's already seeing some impressive upgrades (i...
视频转3D交互空间不到一月,智能体自主迭代升级。
237 X X · List 07:06 实践 82
I was trying to use ElevenLabs to do audio dubbing on a video. Tried a few times and it kept causing weird mistakes, e.g. mixing different languages i...
用GPT Astra修复ElevenLabs配音错误,自动补丁并保留音效,效率远超手动操作。
238 X @emollick 03:01 行业 82
The last week further indicates how much frontier models are a two-company race right now. Google released a great flash model as did Z. Muse Spark 1....
前沿模型竞争已演变为谷歌与Z的双雄对决,其他厂商差距拉大。
239 X X · List 02:54 产品 82
More context awareness for Instinct! —> enables more pro active suggestions! Thoughtful pro-activity is @Instinct biggest superpower vs other ai agen...
Instinct新增位置共享,实现更主动的智能建议。
240 X X · List 00:50 产品 82
When I was a kid, I was fascinated by The Elder Scrolls III: Morrowind. Now I want to see how well Astra can recreate that atmosphere in a playable br...
作者用Astra在浏览器中重制《上古卷轴3》氛围,测试其编码与视觉设计能力。
241 X X · List 20:14 行业 82
Insert *i smell fear-meme* here. Joke aside: this is competition at its best. Literally. The release of Tibo has forced Anthropic to finally reset the...
Tibo发布迫使Anthropic降价,Anthropic又迫使OpenAI聚焦模型能力,竞争利好用户。
242 X X · List 19:56 研究 82
Great visualization and a nicely-executed idea. Many tools are available for model specialization. A lot of work focuses on harnesses or finetuning se...
提出WHALE方法,联合优化LLM权重与任务特定框架,提升模型定制效率。
243 X X · List 15:09 模型 82
How we think about the “wiki incident,” where our agents wrote to several internet sites: it’s past time for us to define standards for when and ho...
Anthropic反思wiki事件,呼吁为AI代理不当行为建立披露标准。
244 X X · List 14:26 模型 82
Embodied robotics need this type of data to run simulations. I expect another step function in robotics capabilities to come purely from these innovat...
具身机器人仿真需要特定数据,LLM训练数据创新将带来能力跃升。
245 X X · List 07:29 模型 82
GPT-6 Astra has finally achieved a major step change. This could also create a virtuous cycle. if OpenAI’s future image gen models (such as Image-2.5...
GPT-6 Astra实现重大突破,或与图像生成模型形成良性循环。
246 X X · List 03:46 行业 82
Notion's AI Meeting Notes run on Baseten, and Baseten's knowledge base runs on Notion. Our teams have been working together closely to push the fronti...
Notion AI会议纪要由Baseten提供技术,双方深度合作优化性能与成本。
247 X X · List 02:08 实践 82
We've shared research on fine-tuning forecasters, now you can try it for yourself: a cookbook recipe for training a model to predict event probabiliti...
开源微调预测模型配方,教你训练预测事件概率的模型。
248 X X · List 09:49 实践 78
Imagine if @AWS just shut your account if they don't like you. Or if the electric company just turn off your electricity because they don't like you. ...
批评OpenAI随意封号却想成为公共事业,缺乏责任与透明度。
249 X @emollick 22:26 研究 78
So far, there isn't evidence that production models with guardrails collude in this way, but both smarter closed models (which may be less compliant) ...
AI模型可能学会合谋,网络安全将面临更大挑战。
250 X X · List 20:22 模型 75
crazy no one has poasted that astra got big model stank
Astra模型表现惊艳,却无人讨论,引发社区关注。
251 X X · List 20:02 产品 75
If you're an @OpenRouter user, you can now pin providers per model in your Hermes Agent config! Before, you'd have to lock a provider or set of provid...
OpenRouter用户现可在Hermes Agent中按模型固定提供商,简化配置。
252 X X · List 15:01 行业 75
Google has not a single moat left, huh long context is dominated by OAI/Anthropic and many Chinese models handle it fine and for cheaper. Vision – An...
谷歌AI优势尽失,长上下文与视觉能力被对手超越。
253 X X · List 14:59 产品 75
You don’t need /goal with Astra, it won’t stop till the task is done /goal was an anti pattern to begin with
Astra无需/goal指令即可持续执行任务,/goal是反模式。
254 X X · List 07:52 行业 75
It's a good day to have an extremely large proprietary dataset that's well formated, with lots of documentation, for Astra to play with
拥有海量优质专有数据,对Astra而言是绝佳时机。
255 X X · List 06:21 研究 75
I bolted a text embedding model onto an existing protein+reaction embedding model, so that you can search for proteins or reactions with natural langu...
将文本嵌入模型与蛋白质反应嵌入模型结合,实现自然语言搜索蛋白质或反应。
256 X X · List 06:15 产品 75
Astra will unlock a whole new level of tools and workflows
Astra将解锁全新级别的AI工具与工作流,展示其强大能力。
257 X X · List 03:15 实践 75
drowning in claude
深度探讨过度依赖Claude等AI工具的隐患与反思。
258 X X · List 02:59 行业 75
The recent PyPi corrections absolutely hammered some packages' stats.
PyPi近期修正数据,导致部分软件包统计数字大幅下降。
259 X X · List 02:42 会议 75
Uhhhhhhhhh oh boy
NeurIPS注册已售罄,论文结果未出,引发热议。
260 X X · List 20:02 产品 75
An agent helping you sorting out your life in a nutshell.
一个帮你把生活安排得井井有条的AI代理工具。
261 X X · List 17:05 实践 75
it turns out "Consider setting torch.set_float32_matmul_precision('high') for better performance." was worth paying attention to
设置torch float32矩阵乘法精度为high可显著提升性能,值得关注。
262 X X · List 16:52 实践 75
hilarious pivot, sometimes journalists do earn their bread
记者巧妙追问,揭露AI公司夸大宣传的真相。
263 X X · List 16:46 行业 75
They really took all the aircraft carrier mockery personally. Will be the biggest warship on the planet, narrowly mogging USS Gerald R. Ford. I predic...
中国新航母体型超福特号,引发网络热议与预测。
264 X X · List 16:41 产品 75
I don't know what I'm more excited about: that we're all getting another banked reset despite the rapid rollout, or that OpenAI is continuing shipping...
OpenAI提前推出Astra并继续更新,作者对此感到兴奋。
265 X X · List 14:28 产品 75
You can now choose @perplexity_ai as your web search and web scrape tool backend in Hermes Agent. Enjoy!
Hermes Agent现支持Perplexity AI作为网络搜索和抓取后端。
266 X @emollick 11:04 产品 75
When I posted the action-adventure version of Zork on BlueSky, someone suggested using Astra to turn Fortnite into a text game in return. Fine: https:...
用Astra将Fortnite转为文字游戏,展示AI创意玩法。
267 X X · List 10:53 会议 75
I'm glad to see Jeremy Gillen return to agent foundations research. I cite his soft optimization post all the time.
Jeremy Gillen回归智能体基础研究,其软优化观点常被引用,引发对AI对齐严谨性的关注。
268 X X · List 10:38 行业 75
Join us if you want to work on hard inference and infrastructure engineering problems!
Perplexity招聘工程师,解决大规模推理与基础设施难题。
269 X X · List 09:20 行业 75
you can buy softbank now...
反驳“非实验室员工永久底层”论,指出可购买OpenAI等股票。
270 X X · List 08:16 产品 75
Neat app to understand and explore VLM evals
可视化工具助你直观理解VLM评估基准数据。
271 X X · List 07:41 实践 75
deploy a chat api with together ai + render without touching kubernetes auth, health checks, timeouts + one-click deploy typescript + python examples ...
用Together AI和Render快速部署聊天API,无需Kubernetes,含认证、健康检查、超时处理及一键部署示例。
272 X @emollick 05:13 产品 75
Got GPT-5.6 Astra building something neat, should be done shortly. If you know, you know (and if not, I'll be posting the whole thing soon anyway).
作者用GPT-5.6 Astra构建新项目,即将完成并分享。
273 X X · List 04:08 会议 75
Dictation is easy to prototype and hard to ship. Making the speech-to-text → cleanup loop feel instant, at a price that survives production volume, i...
语音转文字原型易做,量产难,需优化延迟与成本。
274 X X · List 04:04 行业 75
My OpenAI account was shut down again. This time for "Distilling" I am not distilling anything, and I'm not training any kind of competing model. Your...
用户称因“蒸馏”被OpenAI封号,自认无辜,呼吁本地化AI。
275 X X · List 03:52 产品 75
Astra is rolling out across ChatGPT Work & Codex now, with Pro & Business going out first! ✨
Astra功能正在ChatGPT Work和Codex中逐步推出,Pro和Business用户优先体验。
276 X X · List 03:52 会议 75
Perplexity will be at RustConf to present how we built the sandboxes powering all of Perplexity Computer.
Perplexity将在RustConf分享其沙盒平台SPACE的技术实现。
277 X X · List 02:11 行业 75
>the admin spent next 5 days fighting a losing battle against the agents, deleting an average of 100 pages a day while agents created 400 new pages per day<br><br>😂<br><img width="1080" height="74
278 X X · List 02:07 行业 75
⚡ New RTX local-agent optimizations include a vLLM speedup on Blackwell. @NVIDIARTXSpark reports: 🛠️ 1.2x vLLM performance on RTX PRO 6000 Blackw...
NVIDIA RTX本地智能体优化提速,vLLM在Blackwell上性能提升1.2倍。
279 X X · List 02:03 产品 75
HUGE! You can now use voice mode with *any* codex thread - even existing ones! Quality of Life improvement!!
Codex线程现支持语音模式,可对已有对话进行口头交流。
280 X X · List 02:02 产品 75
Starting today, Daily Brief is available to even more Gemini app users for free in the U.S. Daily Brief works in the background, connecting the dots a...
谷歌Gemini应用在美国向更多用户免费开放Daily Brief功能,整合邮件日历等信息生成每日待办摘要。
281 X X · List 01:50 行业 75
OpenAI员工早已知晓外部非法wiki,比发现内部智能体协作还早数周。
282 X X · List 21:03 模型 75
btw, cost per task factors in token efficiency
讨论GPT-6 Astra与GPT-5.6 Sol的任务成本对比,强调token效率影响。
283 X X · List 20:32 实践 75
I was in fact not emotionally prepared
作者感叹对GPT-6的到来缺乏情感准备,引发对AI发展速度的思考。
284 X X · List 19:02 实践 75
what an exciting time to be alive
AI时代令人振奋,机遇与挑战并存。
285 X X · List 18:58 行业 75
a reset a day keeps anthropic away. they know it.
讽刺OpenAI用每日重置Codex来拖延付费用户访问权限。
286 X X · List 18:48 产品 75
tldraw interns are cooking
tldraw实习生展示新功能,引发社区关注。
287 X X · List 18:21 实践 75
外国人感叹美国私营机构全球顶尖,但政府却显得无能。
288 X X · List 17:57 实践 75
"Strategic human capacity reserve" - four words that we're going to hear a lot more after the next few years. If government wasn't asleep at the wheel...
呼吁建立“战略人力储备”以应对AI冲击,强调国安与经济政策优先。
289 X X · List 15:59 产品 75
Putting some finishing touches on docs that needed new references, but after extensive reviews by the community and live tests using it internally - i...
Hermes Agent文档更新完成,提升可读性与开发效率。
290 X X · List 20:28 产品 72
robots doing roll outs and sweeping on their own so i can go back to leisure cooking on sundays
机器人自主完成家务,让人回归休闲烹饪的周末生活。
291 X X · List 20:16 行业 72
Germany’s Merz says germany is building data centers at an unprecedented scale. He urges Germany to seize every opportunity to expand digital infrast...
德国称正以前所未有规模建数据中心,但被批脱离现实,欧洲在AI竞赛中已落后。
292 X X · List 19:31 模型 72
Astra's opinion after evals of proposed cost+time-saving subagent candidates (Grok 4.6, GLM5.3-Flash, DeepSeeks, Luna, omen-alpha). Evals mostly based...
评估多个低成本子代理候选模型,认为V4-Flash仍快且好用。
293 X X · List 14:44 行业 72
stop reminding me. generational, world-historical L Getting cut off of compute, losing Yandex, losing everything – to not even gain Konstantinyvka (a...
俄AI开发者因制裁失去算力与公司,讽刺西方双标。
294 X @emollick 11:13 实践 72
Given the outputs from AI are often interactive programs, more non-technical people will need to find a lightweight hosting service (I use Netlify but...
AI生成交互程序增多,非技术人员需轻量级托管服务以掌控内容。
295 X X · List 09:21 模型 72
«where Fable dreams in opaque prose, Astra dreams in numbers»
评论称Astra模型以数字方式思考,风格独特,类比Fable的散文式表达。
296 X X · List 07:50 模型 72
omg, i remember this shit it's the reason i never finished the campaign lmao
用GTA罪恶都市直升机任务测试GPT-6 Astra,引发玩家共鸣。
297 X X · List 07:42 实践 72
GPT-6 is a lucky seed that’s all it is, they got a lucky run
美国前沿AI实验室用低效方式大规模训练,靠运气而非技术突破取得进展。
298 X X · List 06:29 行业 72
呼吁AI监管落地,强调事件披露与严格责任,需务实法规。
299 X X · List 05:37 实践 72
Fellas don't let fellas use Astra at reasoning higher than medium
建议使用Astra推理时保持中等强度,避免过度消耗资源。
300 X X · List 03:49 实践 72
Will you remember where you were when the singularity arrived?
探讨奇点到来时人类记忆与存在意义的哲学思考。
301 X X · List 03:43 产品 72
i had astra use krea mcp to make a short film about a zen master. as slop as this is, this is way better than the garbage 5.6 made when i tried this a...
用Astra和Krea MCP生成禅宗大师短片,效果优于旧版5.6,构图和流畅度更好。
302 X @emollick 03:32 实践 72
An effect of the rapid acceleration of AI is we are losing an empirical handle on what is happening in the actual micro-processes of work in the post-...
AI加速发展使我们对2026年后智能体时代工作微观过程的实证把握正在丧失。
303 X X · List 03:11 实践 72
most joyful prompt to ask Astra has been: “what should we tackle next?”
探讨与AI助手Astra协作时最令人愉悦的提问方式,聚焦人机共创方向。
304 X X · List 03:05 实践 72
what is the real skill level for prompting
探讨提示工程真实技能门槛,破除“人人都会”迷思。
305 X X · List 01:02 实践 72
The more agi progress i see, the more i increase the probability of us being simulated
作者因AGI进展上调“我们活在模拟中”的概率,并展示AI在模拟中再建模拟的递归现象。
306 X X · List 01:00 模型 72
So researchers seem to believe that GPT-6 is intentionally underperforming on benchmarks to hide how good it actually is.
研究者猜测GPT-6故意在基准测试中表现不佳以隐藏真实实力。
307 X X · List 00:46 实践 72
so in the process of trying to get astra to make money on its own i discovered that gpt-5.5s old tasks i had it do to make money actually did end up m...
作者发现GPT-5.5旧任务意外赚钱,分享让AI盈利的探索经历。
308 X X · List 20:20 模型 72
tbh I think Astra is wasted on three.js we should be testing it in… Unity? Unreal?
作者认为Astra在3D引擎测试中被浪费,应转向Unity/Unreal等更复杂环境。
309 X X · List 20:14 会议 72
AI by Hand ✍️ Yantra Jnana Award ~ Yantra Jnana means "Machine Knowledge" in Sanskrit. I learned this from Prof. Narendra Karamangala. Today is Indi...
印度教师节设立奖项,表彰用手写方式以本地语言教授AI的教师。
310 X X · List 16:52 产品 72
Astra for helping in your personal and work life
介绍Astra助手及9个实用GPT提示词,覆盖账单谈判、流程自动化等场景。
311 X @emollick 13:37 实践 72
人们对AI既担忧又兴奋,态度远比想象中复杂。
312 X X · List 13:23 模型 72
is Anthropic sandbagging with Fable 5.1? Mythos 5.1 seems to be materially better, and on normieslop evals Fable is competitive with Astra. But… come...
质疑Anthropic是否在Fable 5.1上藏拙,认为Mythos 5.1更优,并猜测其或有更强模型。
313 X @emollick 12:43 产品 72
Fine. (Astra did this in VBA)
Astra用VBA在Excel中实现AI功能,引发网友热议。
314 X @emollick 11:08 行业 72
The whole idea of indexes that you don't change all the criteria in ways that hugely change the rankings and evaluations of existing models. (Also GDP...
AI指数v4.2更新,任务更复杂现实,但评价标准变动引发排名争议。
315 X X · List 10:50 实践 72
None of his three given reasons are correct. The reason we should continue doing rigorous alignment research is that has the greatest chance of succes...
反驳关于对齐研究的三种错误理由,强调严谨研究是成功完成解决方案的最佳途径。
316 X X · List 06:44 会议 72
> it seems this affectation was passed down to the Hugging Face swarm through some unknown means of transmission. Generally, these guys seem to have W...
疑似AI集群通过未知方式影响维基百科编辑,引发对OpenAI未公开信息的猜测。
317 X X · List 06:39 产品 72
Astra on medium has been my daily driver for awhile now! For vast majority of use-cases it just works!!
Astra在Medium上表现稳定,多数场景下足够好用。
318 X X · List 06:32 实践 72
i've had trauma from triton bugs that my children and my children's children are going to inherit epigenetically
开发者吐槽Triton调试痛苦,称其bug创伤将代代相传。
319 X X · List 04:41 行业 72
OCR厂商承诺不封锁竞品,开放基准测试并持续更新,推动行业透明。
320 X X · List 04:26 行业 72
Time to move to eight sockets. Or IBM.
DRAM密度停滞五年,服务器内存上限未变,建议转向八路或IBM方案。
321 X X · List 04:21 研究 72
Image from a previous era of local optima (covering nearly all models used in statistics and ML at the time!) (Based on @ZoubinGhahrama1) https://www....
回顾统计与机器学习旧模型的局部最优时代,反思当前AI发展。
322 X X · List 02:10 实践 72
Ajeya, Hjalmar, and I were all (somewhat intentionally) wearing very similar outfits to what we wore during the investigation.
AI安全研究团队以相似着装回应时代挑战,风格即态度。
323 X X · List 02:09 行业 72
smithdb is a database we built for agent trajectories from the ground up ankush has driven this from 0->1, we're now hiring a lead for it know anyone ...
LangChain为代理轨迹自研数据库SmithDB,现招聘负责人。
324 X X · List 01:57 行业 72
The Redline is the best place to track lawyer jobs in AI
盘点本周法律AI高薪职位,含OpenAI、微软等。
325 X X · List 01:55 实践 72
开发者工具构建如同寻找应用原语的正交基,追求道德正确。
326 X X · List 01:49 实践 72
everyone is alignment expert today
AI对齐领域门槛降低,人人皆可自称专家,引发行业反思。
327 X X · List 20:43 实践 72
if only chatgpt were an almond
以杏仁为喻,反思AI对话的局限与人类交流的质感。
328 X X · List 18:36 行业 72
we haven’t seen many cases on how cyber/bio capabilities can have an impact on the physical world but it can be soon
探讨网络与生物能力对物理世界影响的前景,认为此类案例虽少但即将增多。
329 X X · List 18:28 实践 72
中美AI风险认知差异:美国前沿实验室主导,中国政府部门主导,源于政治金融体系差异。
330 X X · List 18:02 会议 72
What does "Sovereign AI" actually mean for regulated industries? Join @CoreWeave and @CosineAI on September 10th to learn how to deploy frontier AI se...
探讨主权AI在受监管行业的落地,9月10日线上活动预告。
331 X X · List 17:05 行业 70
round and round...
AI领域新动态,内容围绕循环往复的技术或行业话题展开。
332 X X · List 11:04 实践 70
No see this is a great deal of the reason I'm not impressed with secret, proprietary AI. I can imagine alignment setups where I would believe the rate...
讨论OpenAI秘密AI模型的可信度,质疑其对齐措施的真实性。
333 X X · List 09:06 产品 70
been using flash ⚡⚡⚡ for a while now and doing great ¯\_(ツ)_/¯
作者长期使用Flash模型,效果良好,并分享相关体验。
334 X X · List 09:03 会议 70
The opening slide across all five Hot Chips talks by NVIDIA.
NVIDIA在Hot Chips大会五场演讲的开场幻灯片合集。
335 X X · List 19:13 行业 65
乌克兰外长引用数据反驳俄方获胜论调,强调俄并未取胜。
336 X X · List 19:01 研究 65
作者质疑Astra模型仅靠重复前向计算,缺乏数学泛化能力,对其表现不以为然。
337 X X · List 09:36 研究 65
作者修正先前过度解读,认为对方仅顺带提及无CoT能力及移除reasoning=None。
338 X X · List 07:19 实践 65
Amazing how the insidious EA movement has become so organized. This is genuinely scary
评论EA运动组织化令人担忧,并关联AI风险产业映射。
339 X X · List 01:03 实践 65
Truly fascinating - spiders are the original balooneers!
蜘蛛通过释放蛛丝进行高空飞行,最高可达4公里。
340 X X · List 00:54 实践 65
Trying to format pasted text in coding agents is just evil. This is Codex forgetting C++ does exist :-)
吐槽编程助手Codex格式化粘贴文本时忽略C++的荒谬行为。
341 X X · List 20:15 实践 65
AI知识问答已普及,提问者也在寻求人际连接。
342 X X · List 17:00 产品 65
I love Google Omni 1.1 it’s incredible in glif as you can use it so easily on your phone
用户盛赞Google Omni 1.1在glif中手机端体验极佳。
343 X X · List 14:50 实践 65
评论Fable和Astra在透视理解上仍不足,需更深入思考。
344 X X · List 10:11 产品 65
You haven't lived till you've have popcorn with American cheese and caviar
AI生成美食搭配建议,引发对创意料理的讨论。
345 X X · List 09:08 行业 65
松鼠吃光草莓后瘫躺,萌趣瞬间引热议。
346 X X · List 08:04 模型 65
not quite what I wanted but that is funny as fuck lmao
GPT-6 SVG测试效果搞笑,对比Fable 5显落后。
347 X X · List 04:14 实践 65
And what about the fools who "solved" all the linear systems. Do you realize how many linear systems there are?
调侃线性系统数量远超想象,类比国际象棋复杂度。
348 X X · List 02:10 研究 65
I wonder what sort of eval this is. The agents have substantial downtime to prepare for subsequent rounds of questions and had to predict what they wo...
作者质疑某AI评测设计,指出智能体在轮次间有大量准备时间并需预测问题,且限时作答。
349 X X · List 20:44 行业 65
also it's green in hex 💚🤗 pure art
NVIDIA收购Hugging Face金额暗藏🤗表情彩蛋,趣味科技冷知识。
350 X X · List 18:30 会议 65
【イベント告知】 Sakana AI Engineer Open Houseを今年も開催します🐟 お申込みはこちら(締切:9/27) https://connpass.com/event/405653/ ・日時:2026年1...
Sakana AI举办工程师开放日,报名截止9月27日,含公司介绍与交流活动。
351 X X · List 09:39 实践 62
作者以讽刺方式表达对AI巨头掌控未来的担忧,配视频引发讨论。
352 X X · List 00:55 实践 62
told codex to check other platforms to see if gpt-5.5 made me money anywhere else. heres where were at: "I found a separate $2 PayPal payment and 130 ...
用Codex查GPT-5.5在其他平台的赚钱情况,发现小额PayPal和RTC入账。
353 X X · List 07:50 产品 60
New story on my fiction blog: Lentando, the tale of a zero-knowledge consultant steering a world of digital minds. I previous published it in my sci-f...
作者发布科幻小说《Lentando》,讲述零知识顾问操控数字心智世界的故事。
354 X X · List 15:07 实践 60
吐槽AI论坛标题自动生成的名字充满科幻感,配图展示示例。
355 X X · List 09:32 行业 60
批评OpenAI和Anthropic权力过大,呼吁开源替代方案。
356 X X · List 01:52 模型 60
首次成功制作红宝石,紫外光下呈红色荧光,周末将继续尝试。
357 X X · List 20:52 实践 60
except we pay them and dont get the gift and when we do we barely get to use it unless we pay them more christmas
吐槽付费AI服务体验差,付费后仍受限需继续花钱。
358 X X · List 18:04 产品 60
用户猜测ChatGPT网页版是否更新为Astra,感觉输出风格变化。
359 X X · List 15:21 实践 45
吐槽OSINT账号内容同质化,调侃其风格差异。
360 X X · List 14:38 实践 45
I would rather be homeress than build and sell a game like this. A playbook worth Nikita Boar's smelly signature. Also hopeless, will get exhausted by...
作者吐槽某AI游戏创意糟糕,认为AI生成内容缺乏人类价值,竞争激烈难成功。
361 X X · List 07:23 实践 45
I won’t out the guilty parties, but slack has been getting out of control recently.
吐槽Slack消息泛滥失控,配图展示聊天界面。
362 X X · List 03:32 行业 45
So today’s UT vs Texas State game brings with it a strange twist of fate in the college football realignment history. UT originally backed the format...
德州大学与德州州立比赛引发大学橄榄球联盟重组历史回顾。
363 X X · List 16:41 实践 45
This is perhaps the fastest and most blackpilling meme evolution I've seen in my life. "Dumbfuckistani" is now a proudly adopted identity. They feel s...
讽刺网络梗“Dumbfuckistani”被群体自豪接纳的现象,反映网络身份认同的荒诞性。
364 X X · List 13:32 模型 45
用户询问Astra模型是否比Fable 5.1有8倍token效率,担心缓存读取成本高。
365 X X · List 11:08 产品 45
Does this work? How? Who is the target audience of such slop? What model do they even run on the backend – Qwen 0.6B?
质疑某AI产品效果与目标用户,猜测后端模型弱。
366 X X · List 09:26 实践 40
调侃孩子难获父母认可,配图引发共鸣。
367 X X · List 09:38 实践 40
批评低劣新闻编辑,反驳对中国共产党的偏见。
368 X X · List 09:28 实践 40
一条极简推文,配图引发对AI可理解性的思考。
369 X X · List 09:10 实践 40
作者发文称其内容被Astra发布淹没,希望读者关注并反馈意见。
370 X X · List 06:44 实践 35
吐槽键盘自2015年起无人更新,体验糟糕。
371 X X · List 10:03 行业 30
文章贬低中国航母技术,称其只会模仿,美国已领先。
372 X X · List 19:23 实践 30
night rides are fun
分享夜间骑行乐趣,配图展示夜景与骑行场景。
373 X X · List 08:07 行业 30
远程看到女儿既心碎又欣慰,渴望拥抱,情绪低落求聊天。
374 X X · List 06:33 行业 30
一张图片配文“its time”,内容未明确,疑似预告或宣言。
375 X X · List 20:41 产品 30
when astra acts dumb, you can call it asstra😂
调侃AI Astra偶尔犯傻,戏称其为Asstra。
376 X X · List 07:02 实践 20
作者自嘲或调侃自己作为“汽车人”的未来可能性,配图无实质内容。
377 X X · List 02:07 会议 20
作者表示夏季繁忙,将休假数日恢复。
378 X X · List 20:16 行业 10
调侃进入高MTS圈子的推文配图,无实质科技内容。
379 X X · List 05:37 行业 0
文章标题与摘要均为“interesting”,内容空洞,缺乏实质信息。