Cognition完成NVIDIA Vera Rubin NVL72基准测试
Cognition测出Vera Rubin比GB200快近5倍,能同时运行更多Devin会话,成本还更低。
Cognition工程师在NVIDIA Vera Rubin NVL72上完成了首个客户执行的基准测试。与NVIDIA GB200 NVL72相比,SWE-2推理吞吐量提升最高4.8倍,强化学习输出吞吐量提升3.8倍。该基准测试和负载均在CoreWeave平台上运行,支持训练、强化学习和用户推理服务。
ICYMI: @Cognition's engineers ran the first customer-executed benchmark on NVIDIA Vera Rubin NVL72.
They measured it against their own NVIDIA GB200 NVL72 baseline.
Up to 4.8x the total token throughput for SWE-2 inference. 3.8x the output token throughput for reinforcement learning.
Agentic coding is an unforgiving workload. Long contexts. High concurrency. Token volumes where cost per token decides what you can ship. For Cognition, those numbers mean more concurrent Devin sessions per GPU, drastically accelerated research loops, and lower cost per session, with no loss in generation speed.
The whole stack runs on CoreWeave: training, reinforcement learning, and the inference that serves users. All of it on one platform.
The benchmark is theirs. So is the workload. https://t.co/gCt0GF45bF