模型官方一手精选

中国移动云推出混合架构大模型推理栈

China Mobile Cloud Debuts GPU–Neuromorphic Heterogeneous LLM Inference Stack

精选理由

移动云搞了个新东西,把GPU和类脑芯片混搭起来跑大模型,在DeepSeek V4 Flash上测试效果不错,能省不少电和钱。

中国移动云在2026中国算力大会上发布了一套混合推理系统,将GPU与类脑芯片结合用于大模型推理。该系统在DeepSeek V4 Flash上测试,实现了约2倍的输出提升和能耗降低,同时将运营成本降低超过40%。

原文 · pandaily

China Mobile Cloud Debuts GPU–Neuromorphic Heterogeneous LLM Inference Stack

At the 2026 China Computing Power Conference, China Mobile Cloud and partners unveiled a domestic GPU plus neuromorphic mixed-inference system for large models, citing roughly 2× output and energy gains and over 40% lower opex on DeepSeek V4 Flash.