模型多源确认83°

Claude Sonnet 5.5、GPT-6.1 Sol和Gemini 4 Argon发布

精选理由

三大模型最新编程性能对比:谁第一谁最便宜,开发者选型参考

Claude Sonnet 5.5在Claude Code中以68分位居编程代理指数榜首,但每任务成本高达14.19美元。Gemini 4 Argon在Antigravity CLI中得64分,每任务成本5.84美元,使用谷歌促销价且尚未公开。GPT-6.1 Sol在Codex中得63分,每任务成本仅1.04美元,约为Argon的六分之一。

原文 · Artificial Analysis

This week Claude Sonnet 5.5, GPT-6.1 Sol and Gemini 4 Argon all launched near the top of the Coding Agent Index leaderboard, but each has a different balance of performance and cost

The Artificial Analysis Coding Agent Index measures agents (a combination of model and harness) across three agentic coding evaluations.

➤ Claude Sonnet 5.5 (max) in Claude Code takes the top spot at 68, but also has the highest measured cost per task: $14.19

➤ Gemini 4 Argon (high) in Antigravity CLI scores 64 at $5.84 per task - less than half of Sonnet 5.5’s cost. Note, this uses Google’s promotional pricing, and Argon is not yet publicly available

➤ GPT-6.1 Sol (xhigh) in Codex scores 63 at $1.04, roughly one sixth of Argon’s cost