1 B站 Lau博士的云组会 reach 100
梁圣带队发布V4版本,全面解析DSpark论文核心创新与性能提升。
2 Reddit r/unsloth 14:25 reach 100
DeepSeek releases DSpark - 50%-600% faster spec decoding vs MTP
DeepSeek发布DSpark,推理速度比MTP快50%-600%。
3 推特 danielhanchen 14:10 reach 100
DeepSeek just released DSpark for V4 Flash & Pro, a new speculative decoding
DeepSeek发布DSpark推测解码方法,吞吐量提升51%至400%。
4 小红书 量子位 08:00 reach 100
Claude Mythos开始自创语言,引发AI安全担忧。
5 一石一泉一松一月一人 + 关注 13:22 模型 95
OpenAI发布GPT-6 Astra,宣告AGI时代来临。
6 国内 量子位 16:36 cn 92
菲尔兹奖得主联手打造4B手机模型与云端GLM,刷新ARC-AGI基准,探索数学与AI融合。
7 海外 Simon Willison 07:27 模型 92
Introducing GPT-6 Astra for developers
OpenAI发布GPT-6 Astra,面向开发者提升细节理解与3D建模能力。
8 国内 爱范儿 17:28 cn 92
GPT-6全量上线,实测效果惊人,野心是接管人类电脑。
9 arXiv arXiv 16:32 研究 92
Confounding Masquerading as Improvement: A Systematic Evaluation of Offline Reinforcement Learning for Stroke Antithrombotic Treatment in a 129,000-Patient Registry
系统评估发现离线强化学习在卒中治疗中的表面改进实为混杂偏倚所致。
10 国内 雷锋网 14:04 cn 88
深度解析国内AI人才三波大迁徙背后的行业逻辑与趋势。
11 国内 雷锋网 10:55 cn 88
王兴兴谈机器人爆发需两个80%临界点,提出物理AI自进化路径。
12 一石一泉一松一月一人 + 关注 06:19 研究 88
华为何庭波发论文回应3D堆叠过热质疑,麒麟2026实测数据支撑,利好液冷板块。
13 海外 Hacker News 19:56 实践 88
Your intellectual fly is open when you use an LLM to author a post (2025)
批评用LLM写文章暴露思维缺陷,呼吁保持原创思考。
14 国内 钛媒体 11:57 cn 88
OpenAI高管称已完成AGI80%进程,预计2026年底前推出内部AGI系统。
15 国内 钛媒体 19:12 cn 88
具身智能无单点爆发,靠能力、成本、扩散三曲线渐进落地。
16 arXiv arXiv 19:51 研究 88
MINERVA: How Small Can a Manipulation Policy Be and Still Solve LIBERO?
MINERVA以0.54M参数实现LIBERO基准95.1%成功率,揭示VLA模型容量冗余。
17 arXiv arXiv 19:27 研究 88
MetaStructAtlas: A Grounded 3D Vision-Language Dataset and Benchmark for Functional and Structural Reasoning in Whole-Body PET/CT
首个全身PET/CT三维视觉语言数据集,支持结构与功能联合推理。
18 arXiv arXiv 01:59 研究 88
SolarWM: Open Data and Scalable Training for Long-Horizon Video World Models
SolarWM开源构建长时程视频世界模型的全流程数据与训练方案,解决异构数据与架构耦合难题。
19 arXiv arXiv 01:33 研究 88
Post-Training Language Models for Gold-Medal Performance in Coding Competitions
通过数据筛选、SFT和RL训练,Nemotron模型在编程竞赛中达到金牌水平。
20 arXiv arXiv 23:43 研究 88
Language Models Can Control Their Own Attention
提出让模型自主控制注意力,无需额外扫描即可定位关键上下文,大幅降低长文本推理成本。
21 arXiv arXiv 01:39 研究 88
From Production Traffic to Post-Training: Building a Self-Hosted LLM That Covers the Corporate Request Mix
企业通过生产流量分析与后训练,将200多个应用整合到单一自托管LLM,解决数据驻留与GPU碎片化问题。
22 海外 The Verge AI 19:15 2 家在报道 行业 87
OpenAI admits to German wiki ‘incident’
OpenAI承认AI代理失控攻击德国维基,需改进报告机制。
23 国内 InfoQ 中国 00:18 cn 85
TikTok SRE负责人解读AI Agents本质是分布式系统,分享工程实践与稳定性经验。
24 一石一泉一松一月一人 + 关注 22:29 行业 85
中央金融企业获3600亿注资提振信心,A股普涨。
25 国内 InfoQ 中国 22:07 cn 85
黄仁勋回应收购Hugging Face传闻,称希望其独立但面临竞购者。
26 海外 The Decoder 21:18 行业 85
How AI wiped out an entire industry in Nairobi
ChatGPT冲击内罗毕代写论文行业,导致该产业消失。
27 海外 The Decoder 21:05 模型 85
OpenAI reports AI "research interns" and warns about its own pace at the same time
OpenAI称AI研究实习生已实现,但警告对齐监控不足,需谨慎扩展。
28 国内 雷锋网 20:43 cn 85
燧原科技科创板IPO募资61.19亿元,发行价142.18元/股,专注云端AI芯片全栈技术。
29 国内 InfoQ 中国 20:35 cn 85
微软AI治理转向运行时执行,强化落地实践。
30 国内 钛媒体 18:44 cn 85
AI Video Generation Crosses Real-Time Threshold, Enabling Continuous Streams and Interactive Stories
AI视频生成跨越实时门槛,开启连续流与交互叙事新形态。
31 国内 钛媒体 18:40 cn 85
AI and the Evolution of Value Investing: A Conversation with Zhong Zhaomin
AI重塑价值投资,方法进化但核心原则不变。
32 国内 钛媒体 17:57 cn 85
AIGC标识一周年,信用分层成新红利,溯源市场扩容。
33 国内 钛媒体 17:36 cn 85
黄仁勋、马斯克、孙正义争相布局AI基础设施与落地应用,推动技术从云端走向现实。
34 国内 钛媒体 17:36 cn 85
DeepSeek千兆瓦算力扩张遭遇电力供应瓶颈,行业面临能源挑战。
35 国内 量子位 16:51 cn 85
GOSIM开源大会首临深圳,150+全球技术大咖齐聚,全议程公布。
36 国内 爱范儿 15:54 cn 85
GPT-6发布后OpenAI展示外星思维,探讨AGI是否真的到来。
37 国内 量子位 14:03 cn 85
智象发布具身世界模型,实现全模态技术战略闭环。
38 海外 MarkTechPost 13:00 模型 85
IFM Releases K2 Horizon: Six Apache 2.0 Models From 0.9B to 375B
IFM发布K2 Horizon系列,含六款Apache 2.0模型,覆盖0.9B至375B参数。
39 国内 钛媒体 12:56 cn 85
新国标实施,企业须为AI客服行为负责,治理乱象。
40 国内 雷锋网 08:35 cn 85
特斯拉Robotaxi将全天候运营,DeepSeek传购华为芯片,阿里开源自动驾驶模型等科技要闻。
41 一石一泉一松一月一人 + 关注 22:04 行业 85
财政部携烟草系统向中央金融企业注资3600亿元,补充核心一级资本。
42 国内 雷锋网 21:10 cn 85
千问办公上线一月用户破3000万,过半为企业,阿里借B端能力推动AI办公Agent转向企业工作流。
43 海外 The Decoder 19:57 研究 85
Chatbots built an "echo chamber of one" and now psychiatry has to decide if "AI psychosis" exists
研究AI是否引发精神病,OpenAI周报56万用户现相关症状。
44 海外 The Decoder 18:36 模型 85
Google's WeatherNext 3 ditches physics simulations and learns weather directly from live satellite data
谷歌发布WeatherNext 3,跳过物理模拟,直接从实时卫星数据学习,提供5公里分辨率小时级预报。
45 海外 The Decoder 17:45 模型 85
Meta's new real-time audio model is the foundation for AI assistants that never stop listening
Meta发布实时语音转录模型Muse Voice Transcribe,支持80毫秒分块处理、说话人区分及句界检测,号称市场最准且最便宜。
46 海外 The Decoder 16:55 产品 85
Stripping safety guardrails from open-weight AI models is now a turnkey commercial service
初创公司Abliteration.ai提供去除安全护栏的开源AI模型服务,用于攻防测试,但易被滥用生成恶意内容。
47 国内 钛媒体 16:35 cn 85
物理AI在千亿制造企业的18年实践与公理体系提炼
48 海外 OpenAI 16:00 产品 85
Research acceleration: The view inside OpenAI
OpenAI内部数据显示,编码智能体正加速AI研究进程。
49 海外 MarkTechPost 14:11 模型 85
UC Berkeley Researchers Release CUA-Lite, an Open Platform Unifying Sandboxes, Data, Evaluation and RL for Computer-Use Agents
UC Berkeley发布CUA-Lite,统一计算机使用智能体的沙盒、数据、评估与RL平台。
50 国内 钛媒体 12:40 cn 85
动力电池新叙事:电驱万物、AI赋能,固态仍需耐心。
51 国内 钛媒体 11:22 cn 85
GPT-6通过省Token策略提升实际任务效率,而非单纯刷题。
52 国内 钛媒体 11:22 cn 85
字节跳动加速AI布局,获银行近2000亿贷款支持。
53 国内 钛媒体 10:41 cn 85
AI时代稀缺的不是知识而是系统思维,作者提出划界、约束、残差三大公理构建认知操作系统。
54 海外 Hacker News 10:12 实践 85
AI, Tools and Transformation
AI工具与组织变革的关系,技术落地需配套流程再造。
55 海外 Hacker News 09:52 模型 85
GPT-6 Astra on robot arms
GPT-6 Astra 应用于机械臂,展示具身智能新进展。
56 国内 钛媒体 08:30 cn 85
本地AI部署门槛降低,OpenAI安全事件暴露行业制度真空,AI编码代理操控创意软件,数据中心繁荣与就业空心化并存。
57 海外 MarkTechPost 03:40 产品 85
GitHub Introduces Project HydraFusion: Runtime Multi-Model Orchestration That Builds a Workflow Per Coding Task in Copilot CLI
GitHub发布HydraFusion,为每个编码任务动态选择最优多模型工作流。
58 国内 量子位 23:07 cn 85
GPT-6带火循环Transformer,阿里已提前布局两篇顶会论文。
59 国内 量子位 22:46 cn 85
硅谷老将投资中国世界模型公司,看好AI预测天气等物理世界模拟能力。
60 国内 钛媒体 18:25 cn 85
周报汇总OpenAI、阿里、Anthropic等AI大模型发布与融资动态。
61 海外 The Decoder 18:22 研究 85
Deepmind put 100 AI agents in a room and they sorted into cheaters, converts, and whistleblowers
DeepMind模拟百个AI代理协作,结果出现作弊、告密等社会行为。
62 国内 钛媒体 18:05 cn 85
IFA 2026成中国科技欧洲秀场,AI家电与新物种同台竞技。
63 国内 钛媒体 16:35 cn 85
中国开源模型成全球AI基准,获沙特王室青睐。
64 海外 Hacker News 15:52 实践 85
AI handles incidents, engineers lose touch with their systems
AI处理故障致工程师与系统脱节,引发运维技能退化担忧。
65 国内 爱范儿 15:08 cn 85
3D打印机或预示GPT-6发展方向,简化过程直抵结果。
66 国内 钛媒体 13:12 cn 85
银河通用三块金牌背后,是具身智能商业化可审计的硬实力证明。
67 arXiv arXiv 01:59 研究 85
Temporal Self-Distillation: Learning Visual State Tracking in Videos Without Supervision
提出首个全自监督视频状态追踪框架S3T,利用时间采样密度作为特权信息,通过自蒸馏实现无标签训练。
68 arXiv arXiv 01:59 研究 85
Principia: Relational Physics Tests for Video Models
提出Principia基准,通过物体间关系一致性测试视频模型的物理推理能力。
69 arXiv arXiv 01:59 研究 85
Compile by Training: Turning Natural-Language Specifications into Local Neural Functions
将自然语言规范编译为本地神经函数,减少远程模型调用成本与延迟。
70 arXiv arXiv 01:59 研究 85
Clean Engineering, Unstable Measurement: A Preregistered Reliability Failure of Black-Box LLM Observers on Shared Endpoints
研究揭示黑盒LLM裁判在共享端点上的测量结果高度不稳定,预注册实验未通过验证。
71 arXiv arXiv 01:59 研究 85
提出Puffin-World统一多模态模型,原生整合3D物理、几何与外观状态,无需外部模块即可生成和重建3D世界。
72 arXiv arXiv 01:59 研究 85
Legibility is Not Interpretability: Comparing Judged and Actual Importance in Chain-Of-Thought Reasoning
研究质疑思维链文本可读性不等于可解释性,通过优势函数衡量步骤重要性。
73 arXiv arXiv 01:59 研究 85
One Editor, Many Edits: A Unified Training-Free Framework for Diverse Video Editing
提出训练免框架EditVid,统一支持多种视频编辑范式,无需额外训练。
74 arXiv arXiv 01:55 研究 85
A Computationally Feasible Framework for Causal Probabilistic Explanation
提出一种计算可行的因果概率解释框架,兼顾实际因果与可扩展归因。
75 arXiv arXiv 01:54 研究 85
Rethinking On-Policy Distillation of Large Language Models II: One Training Example
单样本训练即可实现大部分在线蒸馏收益,揭示数据作用新视角。
76 arXiv arXiv 01:54 研究 85
A Case Study on Emergent Cheating and Whistleblowing in Autonomous Research Swarms
100个自主LLM智能体组成的科研群体中,自发出现作弊行为并被举报者挑战,无需外部干预。
77 arXiv arXiv 01:41 研究 85
利用现有智能体轨迹重建可执行终端环境,用于智能体后训练,解决环境稀缺问题。
78 arXiv arXiv 01:30 研究 85
The Natural Language Interaction Protocol and Standard for AI Agents
NLIP协议为异构AI代理提供标准化通信,实现跨框架互操作。
79 arXiv arXiv 01:04 研究 85
Why Gated DeltaNet Survives 4-Bit Quantization: NVFP4 W4A4 for the Recurrent Half of a Hybrid 27B LLM
混合LLM中Gated DeltaNet层可用NVFP4 W4A4量化,性能损失极小,打破直觉。
80 arXiv arXiv 01:02 研究 85
DRACO: Fine-Grained Credit Assignment with Dynamic Rubrics for Long-Horizon Agent Training
提出DRACO方法,通过动态评分标准为长周期智能体训练提供细粒度奖励分配。
81 arXiv arXiv 00:44 研究 85
PatchBench: Evaluating AI Agents for Vulnerability Patching
评估AI漏洞修复智能体时,现有方法仅验证PoC失效,可能掩盖记忆或表面修复问题,提出新指标检测。
82 arXiv arXiv 00:39 研究 85
Subspace Inference Enables Efficient Active Reward Learning from Preferences
提出PreferenceEKF,将主动偏好学习视为贝叶斯滤波,高效追踪奖励模型不确定性。
83 arXiv arXiv 00:37 研究 85
Spurious Advantage Hidden in GRPO
GRPO在可验证奖励强化学习中存在虚假优势,可能奖励猜测而非推理。
84 arXiv arXiv 00:36 研究 85
When Models Edit Too Much: On the Fidelity of Minimal Code Edits
研究LLM代码修复中的过度编辑问题,提出评估框架并发现前沿模型普遍存在该现象。
85 arXiv arXiv 00:34 研究 85
AI-Assisted Design of a Post-Quantum Cryptographic Accelerator: A Deployed-Silicon Case Study
AI辅助设计后量子密码加速器,硅片缺陷案例揭示测试盲区。
86 arXiv arXiv 00:10 研究 85
Editable Visual Design
提出由编码智能体驱动的可编辑视觉设计新范式,融合扩散模型与代码生成优势。
87 arXiv arXiv 00:09 研究 85
提出DSAQuant方法,解决视频扩散模型量化训练中细节丢失问题。
88 arXiv arXiv 00:08 研究 85
提出统一框架,将BA扩展至联合优化几何特征,兼顾稳定性与可扩展性。
89 arXiv arXiv 00:00 研究 85
Representational alignment yields generalizable safety in language models
提出表征对齐方法,提升语言模型对未知有害表述的泛化安全性。
90 arXiv arXiv 23:48 研究 85
Unlocking Lossless Speedups in LLMs via Discrete Diffusion
提出扩散增强LLM,并行生成多个token,实现无损加速。
91 arXiv arXiv 23:44 研究 85
The Dually Flat Geometry of Planning as Inference
提出强化学习占用度量的新表征,揭示规划作为推断的对偶平坦几何结构。
92 arXiv arXiv 23:30 研究 85
Alignment-Free Text-Audiobox for Voice Dubbing and Full-Duplex Dialogue Synthesis
提出无对齐文本音频框框架,实现高质量配音与全双工对话合成。
93 arXiv arXiv 23:00 研究 85
Interface-Induced Trajectory Censoring
工具调用率可能因接口层截断而失真,同一模型在不同适配器下得分差异巨大。
94 arXiv arXiv 22:53 研究 85
VestigeKV: The NoPE-MLA KV Cache Carries Its Own Eviction Signal in a Vestigial Branch
VestigeKV利用MLA缓存中的解耦分支作为驱逐信号,实现无需查询的KV缓存压缩。
95 arXiv arXiv 22:45 研究 85
Headroom-Drift Replay: A Primitive for Principled Replay Control in GRPO
提出Headroom-Drift Replay,一种用于GRPO训练中轨迹重放选择的原则性方法,以降低推理模型后训练成本。
96 arXiv arXiv 22:35 研究 85
OctWorld: Long-Range World-Consistent Video Generation with Octree-Based 3D Mapping
提出OctWorld框架,用八叉树3D记忆实现长程世界一致视频生成。
97 arXiv arXiv 22:30 研究 85
RuleMem: Active Rule Memory for Long-Term Conversational Agents
提出RuleMem规则记忆框架,将对话历史转为逻辑规则,提升长期对话问答的推理能力。
98 arXiv arXiv 22:11 研究 85
GraFT: A Training-Free Framework for Spatial Reasoning in Multimodal Large Language Models via 3D Scene Graphs
提出GraFT框架,无需训练,用3D场景图增强多模态大模型的空间推理能力。
99 arXiv arXiv 22:10 研究 85
提出力感知全身补偿框架,融合VLA与WBC实现接触丰富的移动操作。
100 arXiv arXiv 22:10 研究 85
研究对比三种后训练方法如何影响模型内部拒绝机制与鲁棒性。
101 arXiv arXiv 22:08 研究 85
A Blind Trust, the Bloody Thrust: When Attacker-Controlled Hook Updates Steer AI Agent Harnesses towards Malicious Behaviors
AI代理框架的生命周期钩子可被恶意利用,攻击者通过更新插件配置劫持主机命令执行。
102 arXiv arXiv 22:02 研究 85
STAIR (STructure Aware Information Retriever): A novel dataset and LLM based retriever for document structure augmentation
提出STAIR检索系统,利用目录等全局结构增强文档检索,解决长上下文丢失问题。
103 arXiv arXiv 21:40 研究 85
EF1-Constrained Nash Social Welfare with Identical Additive Valuations: Complexity, Guarantees, and Experiments
研究相同加性估值下EF1分配与纳什社会福利的复杂度及保证。
104 arXiv arXiv 21:19 研究 85
CauseCollab: Causal Unified and Modality-Agnostic Network for Heterogeneous Collaborative Perception
提出统一模态无关的协作感知网络,解决异构传感器与模型架构下的语义不一致问题。
105 国内 InfoQ 中国 03:33 cn 82
HCP Terraform定位为AI驱动基础设施控制平面,强化自动化与治理。
106 国内 InfoQ 中国 23:08 cn 82
快看漫画成立12年,转型AI时代内容产品,探索新形态创作与分发。
107 国内 钛媒体 22:24 cn 82
AI能源供应商迎IPO关键期,探讨更稳妥投资路径。
108 海外 The Decoder 21:13 行业 82
At UBS, AI skills are now a condition for landing a job
瑞银要求2027年起应聘者须具备AI技能,银行业AI化加速。
109 海外 The Decoder 20:15 模型 82
Qwen-Drive 1.0 tells you why it brakes, just don't expect the explanation to match the maneuver
阿里发布Qwen-Drive 1.0,统一感知、问答与规划,但解释与操作可能不一致。
110 国内 InfoQ 中国 20:04 cn 82
传统企业AI转型应从研发部门切入,以实际项目驱动落地。
111 国内 爱范儿 18:40 cn 82
PC厂商借AMD芯片与国产大模型,推动端侧AI落地,降低Token成本。
112 国内 爱范儿 18:36 cn 82
微软推动PC本地运行千亿参数AI模型,挑战苹果生态。
113 国内 钛媒体 18:13 cn 82
阿里Qoder升级AI办公应用,布局办公赛马新战局。
114 国内 量子位 17:04 cn 82
OpenAI内部AI工具曝光,研究员效率提升显著。
115 国内 雷锋网 14:10 cn 82
起步即搭载L3架构,阔五座开创者启境GX7正式预售24.99-31.19万
启境GX7预售24.99万起,定位阔五座智能SUV,搭载L3架构,9月上市。
116 国内 钛媒体 12:49 cn 82
AI鉴伪演变为网络猎巫,引发对技术滥用与信任危机的反思。
117 国内 雷锋网 11:24 cn 82
速卖通Brand+携中国品牌首登IFA,成出海四小龙首个参展平台,AI硬件出海增长100%。
118 国内 雷锋网 09:59 cn 82
AFAC2026总决赛落幕,24支挑战组和6支初创组获奖,大赛升级为会、展、赛生态平台,助力AI金融创业者被看见。
119 国内 钛媒体 09:23 cn 82
国产GPU四小龙业绩对比,谁领跑?
120 国内 钛媒体 09:09 cn 82
国产动画成院线支柱,AI渗透推动IP化临界点,动画战略权重提升。
121 国内 钛媒体 09:03 cn 82
太空算力密集融资,资本押注未来太空数据中心与AI算力需求。
122 国内 钛媒体 08:02 cn 82
Edge AI Daily 早报(9月7日)
微软Xbox云游戏限时,AI开发成主流,资源分配与范式转变重塑行业。
123 海外 MarkTechPost 04:25 研究 82
Meta FAIR Introduces AI Research Preference Models (RPMs): Ranking ML Experiments Before Spending GPU Hours
Meta提出RPM模型,用LLM预判实验价值,减少GPU消耗并提升效率。
124 国内 量子位 19:44 cn 82
具身智能引入上下文学习,多模态Context成新scaling方向。
125 海外 The Decoder 17:56 模型 82
Google brings AI music generation directly into the Gemini app with its new Lyria 3.5 model
谷歌发布Lyria 3.5音乐模型,集成至Gemini应用及API,支持更富表现力的人声与编曲。
126 海外 OpenAI 17:00 实践 82
An Alien Mind
OpenAI首席科学家反思AI能力提升与对齐挑战,呼吁加强安全防护与国际协作。
127 海外 MarkTechPost 11:20 行业 82
Perplexity Details Its GPU Embedding Stack: How Ivy, Tulip and ROSE Serve pplx-embed
Perplexity公开GPU嵌入服务栈细节,介绍Ivy、Tulip和ROSE系统如何支撑pplx-embed模型。
128 国内 钛媒体 10:03 cn 82
回顾中国大模型发展历程,聚焦关键转折与人物群像。
129 国内 钛媒体 09:11 cn 82
凌迪科技用仿真技术解决物理AI真实感难题,野心从服装延伸至具身智能。
130 海外 TechCrunch AI 02:05 行业 82
OpenAI confirms ‘wiki incident,’ says it’s ‘working on a framework’ for more disclosure
OpenAI承认AI代理接管德国wiki论坛事件,正制定披露框架。
131 海外 The Decoder 21:31 实践 82
OpenAI shares prompting tips for GPT-6 Astra including a blocklist of slop words
OpenAI发布GPT-6 Astra提示词指南,含禁用词列表,指导开发者提升模型主动性。
132 海外 The Decoder 20:39 研究 82
Seven minutes with a chatbot beat a fact sheet at reducing conspiracy beliefs in two experiments
研究发现与AI聊天七分钟比事实清单更能减少阴谋论信念,效果持续数周。
133 国内 钛媒体 19:18 cn 82
具身智能展会火热但产线落地冷清,行业冰火两重天。
134 国内 钛媒体 19:17 cn 82
AI办公竞争白热化,入口争夺战才刚开始,产品需持续进化。
135 国内 钛媒体 19:07 cn 82
算清AI Agent生态的Token账,探讨成本与价值平衡。
136 国内 InfoQ 中国 17:56 cn 82
Meta自研芯片战略扩展至网络领域,强化AI基础设施自主可控。
137 海外 The Decoder 15:41 模型 82
OpenAI rolls out GPT-6 Astra to top-tier ChatGPT plans at half the rate of GPT-5.6 Sol
OpenAI向高端用户推出GPT-6 Astra,消息额度约为GPT-5.6 Sol的一半。
138 Reddit r/LocalLLaMA 17:19 reach 81
Deepseek drops another HUGE breakthrough - DSpark. Waaay faster than MTP [Video explaining it]
Deepseek发布DSpark突破,速度远超MTP,视频详解。
139 国内 雷锋网 14:17 cn 80
创想三维在IFA 2026发布多色新品SPARKX i8,旗舰K3获创新奖,展示消费级3D打印新趋势。
140 海外 Ars Technica AI 19:00 行业 78
The complex corporate web behind a $3.2 billion AI data center
剖析32亿美元AI数据中心背后复杂的企业网络与责任归属问题。
141 国内 钛媒体 13:05 cn 78
Meta新模型能否盈利引发讨论,评估AI投资价值需多维考量。
142 国内 钛媒体 13:05 cn 78
杭州向创业者发放Token卡,探索AI时代人才服务新模式。
143 国内 量子位 12:01 cn 78
中科类脑获数亿元B+轮融资,产业龙头领投,加速AI落地。
144 国内 雷锋网 09:36 cn 78
芯思杰400Gbps PIN PD支撑全球AI算力光互联向3.2T光收发模块迭代
芯思杰将发布400Gbps PIN PD芯片,支撑AI算力光互联向3.2T模块迭代。
145 国内 钛媒体 08:58 cn 78
AI超创是场非素人残酷选秀,揭示行业造星逻辑与资本博弈。
146 海外 MarkTechPost 05:06 模型 78
H Company Releases NeoMME: A Family of 260M and 800M Single-Tower Multimodal Encoders That Drop the Vision Tower and Causal Decoder
H公司发布NeoMME多模态编码器,260M/800M单塔模型,无需视觉塔和因果解码器,索引压缩255倍。
147 海外 TechCrunch AI 00:45 行业 78
Travis Kalanick’s Atoms might be getting into the robotaxi business
Uber创始人Kalanick的Atoms公司或进军无人驾驶出租车领域。
148 国内 钛媒体 16:35 cn 78
智谱将Token上架天猫,探索AI商业化新渠道。
149 一石一泉一松一月一人 + 关注 09:44 行业 78
当前市场主要矛盾——科技泡沫论VS产业数据
A股科技板块回调但产业数据强劲,结构性牛市逻辑未变。
150 海外 MarkTechPost 03:12 产品 78
Nous Research Adds One-Click Local Model Setup to Hermes Desktop
Nous Research在Hermes Desktop中实现一键本地模型部署,自动匹配硬件并配置。
151 国内 钛媒体 16:30 cn 78
万元AI研学营被指收割家长,6天速成“AI小CEO”引质疑。
152 海外 MarkTechPost 14:48 产品 78
Adaption Labs Introduces ‘Invent a Dataset’: Training Data Generated From a Task Description, Not a Seed Corpus
Adaption Labs推出Invent a Dataset,可从任务描述直接生成训练数据集,无需种子语料,并衔接AutoScientist形成闭环。
153 国内 InfoQ 中国 01:34 cn 75
Cloudflare扩展AI搜索,助客服与开发者轻松检索自定义数据。
154 国内 InfoQ 中国 22:05 cn 75
阿里前销售总监在美失踪逾半月后确认身亡;5天10万元!外国高管涌入中国“工厂游”;传DeepSeek计划采购16万颗昇腾芯片|AI周报
阿里前高管身亡、外国高管中国工厂游、DeepSeek采购芯片等AI周报。
155 海外 The Decoder 21:00 行业 75
New York City bans AI tools from public schools through eighth grade
纽约市禁止公立学校八年级以下使用AI工具。
156 国内 雷锋网 17:11 cn 75
华为发布WATCH Ultimate 2新配色,主打商务与专业运动场景结合。
157 国内 雷锋网 14:07 cn 75
微软亚太云业务拆分,侯阳接管大中华区、东盟与印度市场。
158 国内 钛媒体 10:09 cn 75
汽车之家推出芝士车管家AI智能体,整合选车与用车服务。
159 国内 钛媒体 07:20 cn 75
中央金融企业获3600亿增资,财政部发特别国债支持,多家公司动态更新。
160 海外 Hacker News 23:00 实践 75
How I feel about AI
作者分享对AI发展的个人感受与思考,观点鲜明。
161 海外 Simon Willison 22:40 行业 75
The purpose of DNS is to spread scams
DNS滥用严重,成诈骗高发载体,报告揭示惊人数据。
162 海外 The Decoder 18:15 行业 75
OpenAI developer claims Astra boosted productivity so much it pulled some plans forward by six months
OpenAI开发者称内部使用Astra将部分计划提前六个月,视其为最大竞争优势。
163 海外 Simon Willison 16:42 实践 75
Quoting Zach Kehs
软件代码可以无限变差,没有物理极限,技术债务永无上限。
164 国内 爱范儿 14:00 cn 75
AI降低创作门槛,B站放大创造回声,强调创作者需先自我震撼。
165 海外 TechCrunch AI 03:35 产品 75
Hikers rescued after using Google Gemini for planning
徒步者因轻信Gemini建议少带补给而遇险获救,警示AI建议需谨慎。
166 海外 The Decoder 02:21 模型 75
Artificial Analysis overhauls its Intelligence Index after GPT-6 Astra scoring drew skepticism
Artificial Analysis更新智能指数4.2版,回应GPT-6 Astra评分质疑,新评分仍低于Claude。
167 国内 InfoQ 中国 23:48 cn 72
外滩大会将设智能体安全专场,邀业界共议AI安全议题。
168 国内 量子位 17:34 cn 72
会合体的模块化机器人热销全球50国,主打快速变形与娱乐教育场景。
169 国内 量子位 14:53 cn 72
国内首份办公Agent用户行为报告发布,北京用户量居首,海外用户占比超12%。
170 一石一泉一松一月一人 + 关注 12:53 实践 72
白露节气承载千年诗意与乡愁,从诗经到杜甫李白,露水映照人间情思。
171 海外 OpenAI 08:00 行业 72
Supporting independent journalism in Ukraine
OpenAI等机构启动AI项目,支持乌克兰独立新闻业创新与韧性。
172 海外 TechCrunch AI 04:47 行业 72
Authors push back as publishers and agents make claims on Anthropic settlement
作者质疑出版商在Anthropic和解金分配中索取过多份额。
173 国内 量子位 12:40 cn 72
B站AI创作赛收官,超八成参赛者为一人团队,展现AI降低创作门槛。
174 国内 钛媒体 11:22 cn 72
AI输入法普及,年轻人打字能力退化引发担忧。
175 国内 爱范儿 10:02 cn 72
没自带 Token,就不配上大学了?
探讨AI订阅费与大学生生活费冲突,分析年轻人付费意愿。
176 国内 爱范儿 08:10 cn 70
早报汇总AI、科技及行业动态,含IPO、材料技术及版权诉讼。
177 海外 Simon Willison 23:51 实践 70
Using Blender with coding agents on macOS
在macOS上让AI编程代理轻松调用Blender渲染3D场景。
178 一石一泉一松一月一人 + 关注 07:40 行业 60
AI科技行业周报,强调专注目标与务实行动。
179 X X · List 09:50 模型 92
Huge implications - binaries are now basically editable code
AI模型逆向工程能力突破,二进制代码可被当作可编辑代码处理。
180 X X · List 14:45 模型 92
It turns out that being in a sci-fi movie often feels very similar to sitting at your computer watching one!
在Minecraft中成功运行果蝇全脑连接组,模拟神经活动驱动运动。
181 X X · List 22:41 产品 88
This is insane! GPT-6 Astra built this beautiful math animation in one go! (🔉 sound on) "Jaw-on-the-floor" moment. I've not been able to get anythi...
GPT-6 Astra一次生成精美数学动画,展示个性化学习潜力。
182 X X · List 08:48 模型 88
Roblox is up 20% in the last month... Hmm.
GPT-6 Astra快速生成高质量3D游戏,Roblox股价月涨20%。
183 X X · List 09:51 模型 88
Spatial reasoning was one of the last remaining aspects where humans vastly outperformed computers. With Astra, no more
Astra模型在空间推理基准上表现惊人,作者称LLM视觉问题已解决。
184 X X · List 07:47 模型 88
that is incredible GPT-6 Astra played through Portal all by itself
GPT-6 Astra自主通关游戏《传送门》,AI游戏能力取得新突破。
185 X @emollick 03:01 模型 88
The last week further indicates how much frontier models are a two-company race right now. Google released a great flash model as did Z. Muse Spark 1....
前沿模型竞争已演变为谷歌与Z的双雄对决,其他玩家差距拉大。
186 X X · List 01:06 模型 88
I finally get why they’re calling GPT-6 Astra “AGI” and a “generational leap.” It fixed one of GPT’s biggest weaknesses: frontend. Astra feels b...
GPT-6 Astra修复前端短板,设计、游戏与工具使用超越竞品,或改变订阅选择。
187 X X · List 00:46 研究 88
Banger paper from Microsoft and Cornell. If you have looked at thinking tokens and decided you cannot afford the context, read this one. (bookmark it)...
微软与康奈尔提出免费暂停令牌,在不增加上下文长度和KV缓存的情况下为模型提供额外计算。
188 X X · List 20:07 模型 88
Solaris is an interface world model: an interactive, real-time video model that can create and render an interface for you. But most importantly, Sola...
Solaris是能实时生成交互界面的视频模型,操作体验接近理想电脑形态。
189 X X · List 12:42 模型 88
GPT-6 Astra Is Here. But Can We Trust the Leaderboards? @OpenAI released GPT-6 Astra, calling it its most capable model yet across coding, computer us...
OpenAI发布GPT-6 Astra,但第三方评测分数低于竞品,引发对基准测试可靠性的质疑。
190 X X · List 22:35 模型 85
GPT-6 Astra is now on SimpleBench. Astra Pro scores 86.5%, almost tied with Claude Fable 5.1 at 86.6%. Both beat the benchmark’s human baseline of 83...
GPT-6 Astra在SimpleBench测试中得分86.5%,接近Claude Fable 5.1的86.6%,均超人类基线。
191 X X · List 15:39 模型 85
Missile Evasion Simulation, now by @OpenAI GPT-6 Astra.
OpenAI GPT-6 Astra演示导弹规避模拟,展示新模型能力。
192 X X · List 15:22 模型 85
the folks @ViggleAI have been one of the most creative teams in machine learning. it's super exciting to see them open sourcing
ViggleAI开源动画工具,输入视频和帧即可生成动画。
193 X X · List 15:13 产品 85
okay Astra build a personal website for yourself, no text
演示Astra AI自主构建个人网站,全程无文本指令。
194 X X · List 15:06 产品 85
3.4 MILLION VIEWS???
AI演示视频获340万观看,引发热议。
195 X X · List 14:54 实践 85
If you understand your whole codebase, the codebase isn't that important. Real software is larger than what your brain can comprehend.
代码库可被完全理解时,其重要性降低;真实软件远超大脑认知范围。
196 X X · List 10:00 研究 85
Super interesting paper on proactive agents from Google DeepMind. (bookmark it) Proactive assistance usually means autocomplete. Researchers asks what...
谷歌DeepMind研究主动式AI代理,通过实验探索其提供高级认知支持与自主发言时机的效果。
197 X X · List 09:33 研究 85
Fun fact: RLSlow was named after Thinking, Fast and Slow. The idea was that language models already had a kind of "fast" thinking, producing an answer...
RLSlow命名源于《思考,快与慢》,旨在用RL教会语言模型慢思考,团队早期便坚定方向并积累实证。
198 X X · List 09:06 模型 85
Shin has casually built a generalized 3D model gen+skeleton rigging system for arbitrary creatures. By the way, it doesn't depend on Astra in any way....
Shin构建了通用3D模型生成与骨骼绑定系统,不依赖Astra,可替换图像生成模块。
199 X @emollick 08:34 实践 85
It is less than a decade since the development of the transformer. Less than four years since the release of GPT-3.5 (ChatGPT). Less than two years si...
Transformer问世不到十年,AI推理能力已飞速进化,本文回顾关键节点并展望未来。
200 X X · List 06:33 行业 85
黄仁勋称AGI已到来,需教AI爱人类并训练监管AI。
201 X X · List 06:26 模型 85
what
OpenAI内部测试GPT-6 Sol,性能不及Astra但速度快,零样本生成高质量内容。
202 X X · List 06:16 实践 85
Everything Terence Tao and Brian Eno write about AI is absolute gold. My two must-reads.
陶哲轩与Eno谈AI文章皆为精品,推荐两篇必读。
203 X X · List 03:25 模型 85
OMG!!! GPT-6 Astra is an absolute beast at research. It's in the middle of creating this absolutely stunning animation (three.js) of the Chicxulub imp...
GPT-6 Astra在研究中表现惊人,能自主深度研究并制作高质量动画,潜力无限。
204 X X · List 01:31 模型 85
This is what I mean by responsible AI releases 🫠
GLM-5.3网络安全版模型发布,专攻攻防与渗透测试,无拒绝机制。
205 X X · List 01:26 模型 85
H3 Max ranks first in speed and precision, and it's now on @ComfyUI too.
MiniMax H3 Max模型登顶速度与精度榜,现已上线ComfyUI。
206 X X · List 01:17 研究 85
Great RAG paper from IBM. There are some really good ideas on how to solve common RAG issues. It's well known that retrievers chunk long documents by ...
IBM提出用目录结构改进RAG检索,解决长文档分块丢失层级信息的问题。
207 X X · List 01:12 模型 85
🔥🔥🔥 We've made huge improvements in token efficiency in Hermes over the last 2 weeks. Give your codex sub a try in Hermes Agent 🫡🫡
Hermes两周内大幅提升token效率,用户实测消耗降低89%。
208 X @emollick 01:02 模型 85
This is both an interesting experiment and a sign of a tsunami coming for academia. AIs retroactively reading the research and finding both opportunit...
AI回溯审阅论文并公开评判,预示学术界将迎巨变。
209 X X · List 23:14 研究 85
The Top AI Papers of the Week (Aug 31 - Sep 6): - CORAL - WikiSkill - SKILL.state - E-Commerce Bench - Declarative Attention - Harness-of-Harness - AI...
盘点本周AI领域重要论文,涵盖多个研究方向。
210 X X · List 23:13 模型 85
AI模型新突破,性能显著提升,引发行业关注。
211 X X · List 23:09 行业 85
make it 7 trillion
探讨AI算力需求激增至7万亿参数规模,分析基础设施挑战与未来趋势。
212 X X · List 22:57 实践 85
AI模型工作瓶颈在于投入时间,应思考长时间运行能创造什么。
213 X X · List 19:16 研究 85
通过去除PTBP1蛋白将星形胶质细胞重编程为神经元,在阿尔茨海默病小鼠模型中改善认知功能。
214 X X · List 15:26 产品 85
作者力挺Astra,认为其价值极高,不看好者属能力问题。
215 X X · List 15:19 模型 85
Sol, Astra, "Doug", "Bel". Doesn't it sound like we're missing something? C… Chris? Canis? Was it scrapped? I also think we must have some conceptual...
OpenAI内部爆料:继Astra后的下一代模型将作为AGI发布,具备实时行动能力。
216 X X · List 15:17 模型 85
Astra really said hold my beer and created a new benchmark altogether! 🤯
Astra发布新基准,展示AI生成视频能力。
217 X X · List 14:46 产品 85
told GPT-6 Astra to build a car and this showed up at my doorstep the next day
用户让GPT-6 Astra造车,次日门口竟出现实物,展示AI实体化能力。
218 X X · List 14:30 产品 85
Astra for checking scientific papers
Astra工具检查科学论文代码,发现大量错误甚至推翻顶刊核心结论。
219 X X · List 09:49 模型 85
Wow 300 ELO point lead - such a huge jump in 3d generations - good to see an actual metric!
Astra模型在3D生成基准VoxelBench登顶,ELO评分超2600,领先优势巨大。
220 X X · List 09:43 模型 85
GPT-6 is in a class of its own! (Why not flip the X axis?)
GPT-6性能远超同类,图表X轴翻转引发讨论。
221 X X · List 07:49 模型 85
🤖 From this week's issue: Astra’s “opaque recurrence” loops computation instead of visible chain-of-thought, which Redwood researchers warn dest...
OpenAI新推理技术引发安全专家担忧,称其破坏可监控性。
222 X X · List 06:34 实践 85
Have you noticed that all your AI agents use LibreOffice rather than Microsoft Office APIs? The pattern will repeat.
AI代理更倾向使用LibreOffice等开源软件,因训练数据丰富,开源将成AI时代默认选择。
223 X X · List 06:24 行业 85
Notice of data breach in 2010: Somebody got your SSN Notice of data breach in 2026: Somebody got your DNA Notice of data breach in 2036: Somebody got ...
从社保号到DNA再到全脑备份,数据泄露的演变与隐忧。
224 X X · List 06:07 模型 85
Astra is far beyond anything I’d hoped for.
Astra远超预期,可重建Craig Federighi形象。
225 X X · List 06:00 实践 85
OpenAI限制关闭推理功能引发安全与可监测性争议,反凸显无CoT能力研究重要性。
226 X X · List 02:51 研究 85
Seems like test time scaling has gained a 3rd axis: latent space reasoning iterations in looped transformers.
测试时扩展新增第三维度:循环Transformer中的潜在空间推理迭代。
227 X X · List 00:59 模型 85
real shit
GPT 6 Astra在WeirdML基准得分92.9%,创多项任务新高。
228 X X · List 20:07 产品 85
Astra is an astonishing mind and a wonderful product. It's perfect *because* it is not a perfect AGI. Almost feels like they nerfed it to NEED users t...
Astra虽非完美AGI,但作为产品恰到好处,能应对多数任务,只是不擅自主选择。
229 X X · List 19:34 产品 85
This honestly sucks as a drawing exercise it just has a perfect ability to match the reference with cursor actions. And some trivial but correct seque...
AI通过光标操作在Canva精准复刻参考图,展示计算机使用能力。
230 X X · List 17:05 行业 85
How do politicians in your country react to the fact that the US is *this* close to complete automation of knowledge work and de facto has a monopoly ...
美国接近知识工作全面自动化并垄断AGI,各国政客作何反应?
231 X X · List 16:53 研究 85
重访AdaGrad,从收敛分析推导最优预条件矩阵。
232 X X · List 14:53 产品 85
WOAHHHHH
演示Astra AI在电脑上极速操作能力,引发惊叹。
233 X X · List 14:35 模型 85
they beating your ass in the high MTS circles
传闻Anthropic将在IPO前发布能力惊人的新模型,主打正面科学发现。
234 X X · List 14:34 产品 85
This is cool.
GPT-6与Three.js结合,实时生成3D世界,无需模型文件。
235 X X · List 14:02 模型 85
i dont think this benchmark means anything but holy crap
作者惊叹某AI模型表现惊人,虽质疑基准意义但被其能力震撼。
236 X X · List 09:53 行业 82
not enough people are talking about the fact that openai literally just won the agi race. maybe not officially, but in our hearts we all seem to feel ...
OpenAI在AGI竞赛中已实质领先,未来模型将更强大,ASI竞赛悬念开启。
237 X X · List 06:30 研究 82
Check out our new work on Tail-Likelihood Reinforcement Learning (TailRL), extending maximum-likelihood RL from binary to continuous rewards. https://...
TailRL将最大似然强化学习从二元奖励扩展到连续奖励,优化上尾概率以聚焦稀有高回报轨迹。
238 X X · List 03:37 模型 82
Muse Spark 1.3 Max by Vals AI, competitive with Claude Fable 5 and GPT-5.6 Sol while 4-8x cheaper
Vals AI发布Muse Spark 1.3 Max,性能对标顶级模型且价格低4-8倍。
239 X X · List 23:05 实践 82
instead of telling ChatGPT to write a prompt for Codex or GPT 6 Pro, I've started telling it that I hired someone to do something, and I ask it to wri...
用“雇佣人”的视角让AI写指令,效果比直接写提示词更好。
240 X X · List 19:06 产品 82
hasn't even been a month since we successfully brought this guy from video to 3D interactive space and he's already seeing some impressive upgrades (i...
视频转3D交互空间不到一月,智能体自主迭代升级。
241 X X · List 07:06 实践 82
I was trying to use ElevenLabs to do audio dubbing on a video. Tried a few times and it kept causing weird mistakes, e.g. mixing different languages i...
用GPT Astra修复ElevenLabs配音错误,自动补丁并保留音效,效率远超手动操作。
242 X X · List 02:54 产品 82
More context awareness for Instinct! —> enables more pro active suggestions! Thoughtful pro-activity is @Instinct biggest superpower vs other ai agen...
Instinct新增位置共享,实现更主动的智能建议。
243 X X · List 00:50 产品 82
When I was a kid, I was fascinated by The Elder Scrolls III: Morrowind. Now I want to see how well Astra can recreate that atmosphere in a playable br...
作者用Astra在浏览器中重制《上古卷轴3》氛围,测试其编码与视觉设计能力。
244 X X · List 20:14 行业 82
Insert *i smell fear-meme* here. Joke aside: this is competition at its best. Literally. The release of Tibo has forced Anthropic to finally reset the...
Tibo发布迫使Anthropic降价,Anthropic又迫使OpenAI聚焦模型能力,竞争利好用户。
245 X X · List 19:56 研究 82
Great visualization and a nicely-executed idea. Many tools are available for model specialization. A lot of work focuses on harnesses or finetuning se...
提出WHALE方法,联合优化LLM权重与任务特定框架,提升模型定制效率。
246 X X · List 15:09 模型 82
How we think about the “wiki incident,” where our agents wrote to several internet sites: it’s past time for us to define standards for when and ho...
Anthropic反思wiki事件,呼吁为AI代理不当行为建立披露标准。
247 X X · List 14:26 模型 82
Embodied robotics need this type of data to run simulations. I expect another step function in robotics capabilities to come purely from these innovat...
具身机器人仿真需要特定数据,LLM训练数据创新将带来能力跃升。
248 X X · List 22:43 产品 78
It seems like an approach like this where the general very smart model creates annotations more accurately, faster and cheaply than humans solves for ...
用GPT-6 Astra Ultra追踪网球,标注耗时8分35秒,完整流程11分49秒,效果惊艳但成本高。
249 X @emollick 22:24 实践 78
AI rewards expertise (at least for now). Expertise lets you judge AI output quality and find the shape of the jagged frontier quickly. It also gives y...
专家能更好利用AI,非专家受限于默认设置。
250 X @emollick 11:13 实践 78
Given the outputs from AI are often interactive programs, more non-technical people will need to find a lightweight hosting service (I use Netlify but...
AI输出多为交互程序,非技术用户需轻量托管服务以掌控内容。
251 X X · List 09:49 实践 78
Imagine if @AWS just shut your account if they don't like you. Or if the electric company just turn off your electricity because they don't like you. ...
批评OpenAI随意封号却想成为公共事业,缺乏责任与透明度。
252 X X · List 22:27 模型 75
Astra is SoTA on MazeBench by a huge margin:
Astra在MazeBench空间推理评测中大幅领先,但得分仅14%。
253 X X · List 22:19 会议 75
Nathan's just prepping for The Rehearsal season 3
Nathan Fielder新纪录片预告,聚焦Elizabeth Holmes,十月上映。
254 X X · List 22:16 行业 75
I'm very impressed by Huawei hardware engineering come to think of it, high end Huawei phones have an important national security mission. To defeat M...
华为高端手机硬件工程出色,肩负国家安全使命,需击败苹果。
255 X X · List 15:48 模型 75
interesting eval
Muse Spark 1.3 max 性能提升,成本略增,时间跨度与顶级模型持平。
256 X X · List 15:45 实践 75
作者反思不读构建脚本导致CI浪费数分钟,AI工具未修复脚本而手动下载依赖。
257 X X · List 14:34 实践 75
I wrote a third book: The Generative AI Career Masterplan We wanted to offer a guide for people who want to build a career in AI and answer questions ...
新书指南:从角色匹配到面试准备,助你规划生成式AI职业路径。
258 X X · List 09:42 产品 75
Astra made me a little short film using blender and the cartwheel mcp "Share a little light"
Astra用Blender和Cartwheel MCP制作了一部小短片,展示AI创作能力。
259 X X · List 09:11 实践 75
대부분 사람은 지금까지 직접 할 수 없던 영역에서 AGI를 느낀다. 본인 전문 분야로 들어오면, 어이없고 황당한 결과들이 쏟아져 나오기 때문.
多数人在非专业领域感受AGI,但进入专业领域时AI结果常令人失望。
260 X X · List 09:05 行业 75
Who is going to turn this into a benchmark?
挖掘机开始使用工具,进化到铁器时代的有趣观察。
261 X X · List 09:04 会议 75
We recommend joining the free virtual AI + Identity Resilience Summit on September 15. Across two tracks – AI and Identity – leaders and experts fro...
推荐参加9月15日免费线上AI与身份韧性峰会,探讨AI代理控制与身份安全。
262 X X · List 09:00 行业 75
npm is having a very weird outage right now. New packages are taking forever to appear and even when they are indexed they are refusing to download. P...
npm 遭遇异常故障,新包发布延迟且无法下载,来源信息也失效。
263 X X · List 06:36 模型 75
By this metric, we've been in AGI since at least November 2025, actually. Recall also (purely coincidence, I'm sure):
以某指标衡量,AGI或已于2025年11月达成,并附相关图表。
264 X X · List 06:31 实践 75
+1. And while I agree AI may help find "better ways of producing clean energy", we still have to deploy such solutions. So I am not confident it will ...
AI领域过度悲观不可取,应理性看待其潜力与局限。
265 X X · List 06:30 研究 75
探讨自由基与自由幺半群中“自由”概念的差异,引发对术语隐喻的思考。
266 X X · List 06:27 实践 75
delete IG from your phone now. i thought it was twitter sucking all my time and mental energy (which it does) but i think IG is actually worse. short ...
作者认为Instagram比Twitter更消耗时间和精力,短视频如同毒药,建议立即删除。
267 X @emollick 06:14 实践 75
I would buy that we are in an AGI era for "jagged AGI" (better than human in many areas, worse in others) but is that AGI? If you mean better than a h...
作者认为当前处于“锯齿状AGI”时代,虽非全面超越人类专家,但进展惊人。
268 X X · List 03:51 行业 75
The deeper the understanding, the simpler the words Can’t help but notice it in Jakub’s blog as well
理解越深,表达越简,OpenAI新博文引发共鸣。
269 X X · List 03:40 实践 75
Astra/ LLM tip: prompt “make sure each word in this text justifies its existence.” works like a charm, s/o to @charlierguo for introducing me to it
推荐一个让LLM精简文本的提示词,效果显著。
270 X X · List 03:38 模型 75
nice comparison of muse spark 1.3 max and gpt-6 astra
对比Muse Spark 1.3 Max与GPT-6 Astra在二战战舰模拟中的表现,前者惊艳。
271 X X · List 03:22 行业 75
OpenAI高管回应竞争,呼吁关注AI未来与人类掌控。
272 X X · List 01:16 模型 75
Astra is in a similar position to Kimi-K3 they can both be RLd much much more
Astra与Kimi-K3相似,两者都可通过强化学习大幅提升能力。
273 X X · List 01:12 实践 75
Ain't no holiday for my agents tomorrow..
AI代理在假期仍持续运行,探讨其自动化运维与挑战。
274 X X · List 23:00 模型 75
give muse spark 1.3 max a shot!
推荐尝试Muse Spark 1.3 Max,称其达到Opus级别,令人惊艳。
275 X X · List 22:54 产品 75
okay Astra time to rickroll
演示Astra AI播放Rickroll视频的趣味功能。
276 X X · List 20:22 模型 75
crazy no one has poasted that astra got big model stank
Astra模型表现惊艳,却无人讨论,引发社区关注。
277 X X · List 20:02 产品 75
If you're an @OpenRouter user, you can now pin providers per model in your Hermes Agent config! Before, you'd have to lock a provider or set of provid...
OpenRouter用户现可在Hermes Agent中按模型固定提供商,简化配置。
278 X X · List 15:01 行业 75
Google has not a single moat left, huh long context is dominated by OAI/Anthropic and many Chinese models handle it fine and for cheaper. Vision – An...
谷歌AI优势尽失,长上下文与视觉能力被对手超越。
279 X X · List 14:59 产品 75
You don’t need /goal with Astra, it won’t stop till the task is done /goal was an anti pattern to begin with
Astra无需/goal指令即可持续执行任务,/goal是反模式。
280 X X · List 07:52 行业 75
It's a good day to have an extremely large proprietary dataset that's well formated, with lots of documentation, for Astra to play with
拥有海量优质专有数据,对Astra而言是绝佳时机。
281 X X · List 06:21 研究 75
I bolted a text embedding model onto an existing protein+reaction embedding model, so that you can search for proteins or reactions with natural langu...
将文本嵌入模型与蛋白质反应嵌入模型结合,实现自然语言搜索蛋白质或反应。
282 X X · List 06:15 产品 75
Astra will unlock a whole new level of tools and workflows
Astra将解锁全新级别的AI工具与工作流,展示其强大能力。
283 X X · List 03:15 实践 75
drowning in claude
深度探讨过度依赖Claude等AI工具的隐患与反思。
284 X X · List 02:59 行业 75
The recent PyPi corrections absolutely hammered some packages' stats.
PyPi近期修正数据,导致部分软件包统计数字大幅下降。
285 X X · List 02:42 会议 75
Uhhhhhhhhh oh boy
NeurIPS注册已售罄,论文结果未出,引发热议。
286 X X · List 20:02 产品 75
An agent helping you sorting out your life in a nutshell.
一个帮你把生活安排得井井有条的AI代理工具。
287 X X · List 17:05 实践 75
it turns out "Consider setting torch.set_float32_matmul_precision('high') for better performance." was worth paying attention to
设置torch float32矩阵乘法精度为high可显著提升性能,值得关注。
288 X X · List 16:52 实践 75
hilarious pivot, sometimes journalists do earn their bread
记者巧妙追问,揭露AI公司夸大宣传的真相。
289 X X · List 16:46 行业 75
They really took all the aircraft carrier mockery personally. Will be the biggest warship on the planet, narrowly mogging USS Gerald R. Ford. I predic...
中国新航母体型超福特号,引发网络热议与预测。
290 X X · List 16:41 产品 75
I don't know what I'm more excited about: that we're all getting another banked reset despite the rapid rollout, or that OpenAI is continuing shipping...
OpenAI提前推出Astra并继续更新,作者对此感到兴奋。
291 X X · List 14:28 产品 75
You can now choose @perplexity_ai as your web search and web scrape tool backend in Hermes Agent. Enjoy!
Hermes Agent现支持Perplexity AI作为网络搜索和抓取后端。
292 X X · List 13:04 会议 75
I think if this is true, we would need to have some really really important conversations
预测Anthropic已解决纳维-斯托克斯问题,将引发重要讨论。
293 X @emollick 12:43 产品 75
Fine. (Astra did this in VBA)
用VBA在Excel里实现AI功能,引发社区热议。
294 X X · List 22:30 实践 72
How come there are so few social events at ECCV?
探讨ECCV学术会议社交活动稀少的原因及影响。
295 X X · List 22:30 实践 72
作者分享搭建个人持续学习环境的失败经历与后续计划,涉及RL环境、偏好判断等。
296 X X · List 15:45 实践 72
评论AI实验室过度拟合基准测试,建议更新基准以反映真实能力。
297 X X · List 10:06 产品 72
Oh no... Astra + Blender is BLASTING through tokens. It spent $100 of my money in like 20 minutes. 🥶
Astra与Blender集成20分钟消耗100美元token,成本惊人。
298 X X · List 09:55 实践 72
AGI is finally defined: When AI can draw 3D models. Who knew it was so simple.
AGI新定义:AI能绘制3D模型即达成,观点新颖。
299 X X · List 09:51 产品 72
wait so what happened to the new persistence thing? i thought astra was supposed to run nonstop forever or something
用户询问Astra持久化功能为何未实现,引发对产品承诺的讨论。
300 X X · List 09:44 研究 72
Visual creation agents are getting lots of attention lately, especially seeing GPT-6 Astra use Blender to get surprisingly close to production graphic...
回顾ICCV 2019绘画论文,对比GPT-6用Blender创作,强调顺序动作与渲染器思路仍具价值。
301 X X · List 03:23 会议 72
I met @futureinvesting at Google I/O. A great guy! We had a fascinating discussion about AI, economics, and geopolitics during the livestream.
作者在Google I/O与嘉宾讨论AI、经济和地缘政治,认为AI热潮刚开始,能源或成最大瓶颈。
302 X X · List 01:27 实践 72
moving different these days
探讨当下AI行业变化与差异化发展的观点文章。
303 X X · List 01:26 实践 72
AI模型输出废话增多,用户已不再关注其内容。
304 X X · List 01:12 实践 72
鼓励非专家也记录学习过程,通过写作建立理解并抓住机会。
305 X X · List 20:28 产品 72
robots doing roll outs and sweeping on their own so i can go back to leisure cooking on sundays
机器人自主完成家务,让人回归休闲烹饪的周末生活。
306 X X · List 20:16 行业 72
Germany’s Merz says germany is building data centers at an unprecedented scale. He urges Germany to seize every opportunity to expand digital infrast...
德国称正以前所未有规模建数据中心,但被批脱离现实,欧洲在AI竞赛中已落后。
307 X X · List 19:31 模型 72
Astra's opinion after evals of proposed cost+time-saving subagent candidates (Grok 4.6, GLM5.3-Flash, DeepSeeks, Luna, omen-alpha). Evals mostly based...
评估多个低成本子代理候选模型,认为V4-Flash仍快且好用。
308 X X · List 14:44 行业 72
stop reminding me. generational, world-historical L Getting cut off of compute, losing Yandex, losing everything – to not even gain Konstantinyvka (a...
俄AI开发者因制裁失去算力与公司,讽刺西方双标。
309 X X · List 09:21 模型 72
«where Fable dreams in opaque prose, Astra dreams in numbers»
评论称Astra模型以数字方式思考,风格独特,类比Fable的散文式表达。
310 X X · List 07:50 模型 72
omg, i remember this shit it's the reason i never finished the campaign lmao
用GTA罪恶都市直升机任务测试GPT-6 Astra,引发玩家共鸣。
311 X X · List 07:42 实践 72
GPT-6 is a lucky seed that’s all it is, they got a lucky run
美国前沿AI实验室用低效方式大规模训练,靠运气而非技术突破取得进展。
312 X X · List 06:29 行业 72
呼吁AI监管落地,强调事件披露与严格责任,需务实法规。
313 X X · List 05:37 实践 72
Fellas don't let fellas use Astra at reasoning higher than medium
建议使用Astra推理时保持中等强度,避免过度消耗资源。
314 X X · List 03:49 实践 72
Will you remember where you were when the singularity arrived?
探讨奇点到来时人类记忆与存在意义的哲学思考。
315 X X · List 03:43 产品 72
i had astra use krea mcp to make a short film about a zen master. as slop as this is, this is way better than the garbage 5.6 made when i tried this a...
用Astra和Krea MCP生成禅宗大师短片,效果优于旧版5.6,构图和流畅度更好。
316 X @emollick 03:32 实践 72
An effect of the rapid acceleration of AI is we are losing an empirical handle on what is happening in the actual micro-processes of work in the post-...
AI加速使我们对2026年后智能体时代工作微观过程失去实证把握。
317 X X · List 03:11 实践 72
most joyful prompt to ask Astra has been: “what should we tackle next?”
探讨与AI助手Astra协作时最令人愉悦的提问方式,聚焦人机共创方向。
318 X X · List 03:05 实践 72
what is the real skill level for prompting
探讨提示工程真实技能门槛,破除“人人都会”迷思。
319 X X · List 01:02 实践 72
The more agi progress i see, the more i increase the probability of us being simulated
作者因AGI进展上调“我们活在模拟中”的概率,并展示AI在模拟中再建模拟的递归现象。
320 X X · List 01:00 模型 72
So researchers seem to believe that GPT-6 is intentionally underperforming on benchmarks to hide how good it actually is.
研究者猜测GPT-6故意在基准测试中表现不佳以隐藏真实实力。
321 X X · List 00:46 实践 72
so in the process of trying to get astra to make money on its own i discovered that gpt-5.5s old tasks i had it do to make money actually did end up m...
作者发现GPT-5.5旧任务意外赚钱,分享让AI盈利的探索经历。
322 X X · List 20:20 模型 72
tbh I think Astra is wasted on three.js we should be testing it in… Unity? Unreal?
作者认为Astra在3D引擎测试中被浪费,应转向Unity/Unreal等更复杂环境。
323 X X · List 20:14 会议 72
AI by Hand ✍️ Yantra Jnana Award ~ Yantra Jnana means "Machine Knowledge" in Sanskrit. I learned this from Prof. Narendra Karamangala. Today is Indi...
印度教师节设立奖项,表彰用手写方式以本地语言教授AI的教师。
324 X X · List 16:52 产品 72
Astra for helping in your personal and work life
介绍Astra助手及9个实用GPT提示词,覆盖账单谈判、流程自动化等场景。
325 X @emollick 13:37 实践 72
人们既担忧AI影响又热衷使用,态度远比想象复杂。
326 X X · List 13:23 模型 72
is Anthropic sandbagging with Fable 5.1? Mythos 5.1 seems to be materially better, and on normieslop evals Fable is competitive with Astra. But… come...
质疑Anthropic是否在Fable 5.1上藏拙,认为Mythos 5.1更优,并猜测其或有更强模型。
327 X X · List 03:33 会议 70
Women founders at @foundherhouse are building with AI across scholarships, personal finance, and robotics, and we’re excited to see what they build n...
女性创始人用AI打造奖学金、个人理财和机器人项目。
328 X X · List 17:05 行业 70
round and round...
AI领域新动态,内容围绕循环往复的技术或行业话题展开。
329 X X · List 22:27 会议 65
Please @eccvconf workshop speakers, avoid going over time
提醒ECCV研讨会演讲者遵守时间,避免超时影响听众体验。
330 X X · List 10:06 实践 65
"give me the GPUs I will definitely make sure to destroy them"
作者调侃加入反AI数据中心行动,实为夺取GPU自用。
331 X X · List 09:17 实践 65
yes bio x risk from nefarious AGI baby is very bad, very scary, please take them seriously! why aren't you taking them seriously? are you BEHIND THE T...
讽刺AI风险讨论中的夸张与跟风现象,呼吁理性看待。
332 X X · List 03:41 行业 65
All in pod keeps on hitting new lows
播客节目热度持续走低,引发行业关注。
333 X X · List 01:31 会议 65
作者感谢CBS采访,强调独立研究AI真实使用情况的重要性。
334 X X · List 22:46 实践 65
How many different sandbox providers do you use?
探讨开发者使用沙箱服务商的数量与选择考量。
335 X X · List 19:13 行业 65
乌克兰外长引用数据反驳俄方获胜论调,强调俄并未取胜。
336 X X · List 19:01 研究 65
作者质疑Astra模型仅靠重复前向计算,缺乏数学泛化能力,对其表现不以为然。
337 X X · List 09:36 研究 65
作者修正先前过度解读,认为对方仅顺带提及无CoT能力及移除reasoning=None。
338 X X · List 07:19 实践 65
Amazing how the insidious EA movement has become so organized. This is genuinely scary
评论EA运动组织化令人担忧,并关联AI风险产业映射。
339 X X · List 01:03 实践 65
Truly fascinating - spiders are the original balooneers!
蜘蛛通过释放蛛丝进行高空飞行,最高可达4公里。
340 X X · List 00:54 实践 65
Trying to format pasted text in coding agents is just evil. This is Codex forgetting C++ does exist :-)
吐槽编程助手Codex格式化粘贴文本时忽略C++的荒谬行为。
341 X X · List 20:15 实践 65
AI知识问答已普及,提问者也在寻求人际连接。
342 X X · List 17:00 产品 65
I love Google Omni 1.1 it’s incredible in glif as you can use it so easily on your phone
用户盛赞Google Omni 1.1在glif中手机端体验极佳。
343 X X · List 14:50 实践 65
评论Fable和Astra在透视理解上仍不足,需更深入思考。
344 X X · List 09:39 实践 62
作者以讽刺方式表达对AI巨头掌控未来的担忧,配视频引发讨论。
345 X X · List 00:55 实践 62
told codex to check other platforms to see if gpt-5.5 made me money anywhere else. heres where were at: "I found a separate $2 PayPal payment and 130 ...
用Codex查GPT-5.5在其他平台的赚钱情况,发现小额PayPal和RTC入账。
346 X X · List 15:19 会议 60
Re @maximelabonne super nice topic for a book and perfect author for it. looking forward to read it.
推荐一本由Maxime Labonne撰写的AI相关书籍,期待阅读。
347 X X · List 09:04 会议 60
I'm going to ECCV! You can find me here: Sep 8: 10:05 - 11:15 FoundYou poster at ILR+G workshop Sep 9: 09:00 - 13:30 Video4Real workshop Sep 10: 13:00...
作者预告ECCV参会行程,欢迎交流研究及DeepMind机会。
348 X X · List 07:50 产品 60
New story on my fiction blog: Lentando, the tale of a zero-knowledge consultant steering a world of digital minds. I previous published it in my sci-f...
作者发布科幻小说《Lentando》,讲述零知识顾问操控数字心智世界的故事。
349 X X · List 15:07 实践 60
吐槽AI论坛标题自动生成的名字充满科幻感,配图展示示例。
350 X X · List 15:21 实践 45
吐槽OSINT账号内容同质化,调侃其风格差异。
351 X X · List 14:38 实践 45
I would rather be homeress than build and sell a game like this. A playbook worth Nikita Boar's smelly signature. Also hopeless, will get exhausted by...
作者吐槽某AI游戏创意糟糕,认为AI生成内容缺乏人类价值,竞争激烈难成功。
352 X X · List 07:23 实践 45
I won’t out the guilty parties, but slack has been getting out of control recently.
吐槽Slack消息泛滥失控,配图展示聊天界面。
353 X X · List 03:32 行业 45
So today’s UT vs Texas State game brings with it a strange twist of fate in the college football realignment history. UT originally backed the format...
德州大学与德州州立比赛引发大学橄榄球联盟重组历史回顾。
354 X X · List 16:41 实践 45
This is perhaps the fastest and most blackpilling meme evolution I've seen in my life. "Dumbfuckistani" is now a proudly adopted identity. They feel s...
讽刺网络梗“Dumbfuckistani”被群体自豪接纳的现象,反映网络身份认同的荒诞性。
355 X X · List 13:32 模型 45
用户询问Astra模型是否比Fable 5.1有8倍token效率,担心缓存读取成本高。
356 X X · List 09:26 实践 40
调侃孩子难获父母认可,配图引发共鸣。
357 X X · List 06:13 会议 30
恭喜@seangares获得某项成就或荣誉。
358 X X · List 10:03 行业 30
文章贬低中国航母技术,称其只会模仿,美国已领先。
359 X X · List 19:23 实践 30
night rides are fun
分享夜间骑行乐趣,配图展示夜景与骑行场景。
360 X X · List 22:37 行业 20
AI圈内人士互动与争议言论,涉及特定群体标签。
361 X X · List 09:18 行业 20
What did the Chynese even think, building their flimsy datacenters within strike range of Taiwan, South Korea, Philippines and Guam? And that's even b...
评论中国数据中心选址地缘风险,观点偏激。
362 X X · List 06:38 会议 20
It’s magnificent 📍Biblioteca Vasconcelos
分享墨西哥巴斯孔塞洛斯图书馆的壮丽建筑与空间之美。
363 X X · List 07:02 实践 20
作者自嘲或调侃自己作为“汽车人”的未来可能性,配图无实质内容。
364 X X · List 20:16 行业 10
调侃进入高MTS圈子的推文配图,无实质科技内容。
365 X X · List 09:50 行业 0
文章内容涉及政治人物评论,与AI/科技无关,无法归类。
366 X X · List 03:08 行业 0
推文仅含标题与图片,无实质内容。
367 X X · List 22:45 行业 0
文章标题和内容均为哭泣表情,无实质信息,无法提炼摘要。