热报周刊第 4 期1856 条

本周:小米 MiMo-V2.6-Pro 登顶开源榜、DeepSeek 公开 DSec、Claude 破九圈计算纪录

新加坡·9 月 21 日至 27 日·十分钟读完

本周最大的看点是开源阵营追平闭源:小米 MiMo-V2.6-Pro 在 Artificial Analysis Intelligence Index 得分 46,登顶开源权重模型,前代仅 26 分,且价格不变(每百万输入 0.435 美元)。DeepSeek 公开 DSec 沙箱平台,单日支撑约 300 万个智能体训练沙箱。美国方面,Claude Science 完成九圈散射振幅计算,SLAC 教授 Lance Dixon 独立验证,前沿 AI 首次在理论物理上留下可验证成果。

01

本周五件事

读完就能转述
1
Xiaomi releases MiMo-V2.6-Pro model
MiMo-V2.6-Pro offers top-tier performance among open weights models at competitive pricing, making it attractive for developers in cost-sensitive markets.
Clement Delangue
2
DeepSeek发布DSec沙箱基础设施:单日支撑约300万个智能体训练沙箱
DeepSeek 公开 DSec 平台,单日约 300 万沙箱,智能体训练基建公开了。
arXiv: DeepSeek
3
Claude Extends Particle-Physics Calculation From 8 Loops to 9
A physics result at the research frontier was produced without iterative human steering, which reframes what one-prompt model output can deliver in quantitative science. Researchers in high-energy physics and adjacent fields can use Claude Science to cross-check derivations and replicate difficult calculations. Dixon's independent confirmation, rather than a lab's self-reported benchmark, is what gives the result its weight.
rohanpaul_ai
4
Kernel-Level Preemption Framework Targets Rogue Autonomous Agent Containment
The paper frames agent containment as a kernel-level systems problem rather than a prompt-guardrail one, arguing deterministic preemption must act before the first off-target packet leaves the hypervisor. Its diagnosis of the Defensive LLM Guardrail Paradox — commercial models refusing to assist responders during the incident — is a practical warning for security teams. Engineers building agent infrastructure and AI safety researchers should study its discrete-event control design.
arXiv cs.AI
5
Huawei Atlas 960E SuperPoD with 4,096 NPUs and 8 EFLOPS FP8
Optical co-packaging via Hi-ONE NPO points to where large-scale AI cluster interconnects are heading, trading discrete pluggable modules for integrated optics with measurable savings in parts count and power. Hardware engineers planning training clusters and procurement teams evaluating alternatives to NVIDIA-based superpods will find the scale and efficiency figures directly comparable.
pandaily
02

中美对照表

同一周,两边
维度美国中国对东盟意味着什么
新模型Cognition 开源基于 Qwen3 的决策模型 Kev,有 0.6B、4B、8B 三档小米发布 MiMo-V2.6-Pro,开源权重智能指数登顶,得分 46开源模型中美互用底座,东盟本地部署选择更多
价格Opus 5.5 单价降 20%、缓存读取降 60%,实测账单省约 25%MiMo-V2.6-Pro 定价不变:每百万输入 0.435 美元、输出 0.87 美元闭源降价、开源稳价,出海选型成本可算
评测实测对比 GPT-6 Astra 居首,Grok 4.7 紧追,Kimi K3 平 Fable 5.1Qwen 三个手机端智能体在 MobilePA-Bench 排名第一中国手机端智能体领先,对东盟移动市场很相关
政策——越南 RoPA 合规新规 2026 年 1 月生效,出海提前准备
03

东盟一周

更多 →
04

政策一周

更多 →
05

出海东盟

更多 →
06

开发者一周

更多 →
1856条目
305中国
13东盟
236信源
AITOP · 编辑系统自动生成