模型多源确认83°

Claude Sonnet 5.5 在代码竞技场获第三名

精选理由

Anthropic 发布的 Claude Sonnet 5.5 在代码竞技场表现优异,性价比超高,比前代版本提升 159 分。

Claude Sonnet 5.5 在 Code Arena: WebDev 中以 1786 分排名第三,仅比第二名 GPT-6 Astra 少 2 分。该模型在游戏、参考设计和品牌营销领域排名第二,在模拟领域排名第三。Claude Sonnet 5.5 的价格为每百万代币 8 美元,比 GPT-6 Astra 便宜 80%。

原文 · lmarena.ai

Exciting update: Claude Sonnet 5.5 with xHigh reasoning has landed in the Code Arena: WebDev. With 1786 pts, its ranked #3 ! At a blended $8/M tokens, Claude Sonnet 5.5 remains on the Pareto frontier with xHigh reasoning. This release is just 2 pts from GPT-6 Astra in the #2 spot with 1788 pts, for 80% of the price. By domain, Claude Sonnet 5.5 (xHigh) landed: - #2 in Gaming, Reference-Based Design, and Brand & Marketing - #3 in Simulations - #4 in Content Creation Tools and Consumer Product - #6 in Data & Analytics Congrats again to @AnthropicAI on this release! Arena.ai @arena Real-world results are in for Claude Sonnet 5.5 (High) by @AnthropicAI . It just landed #4 in Code Arena: WebDev with 1699 pts, and has reshaped the Pareto frontier with its cost efficiency! Claude Sonnet 5.5 (High) delivers nearly top performance at a blended $8 per Mtoken, reshaping the Pareto frontier! This model is 80% cheaper than both Claude Fable 5.1 (Max) in the #3 spot overall, and GPT-6 Astra (Max) at #2 . See Pareto placement below. Overall, Claude Sonnet 5.5 (High) is a +159 pt improvement from Sonnet 5 (High) at #37 with 1540 pts. This gain compared to its previous variant also shows up across these key domains so far: - Reference-Based Design: #38 → #4 - Simulations: #37 → #4 - Gaming: #36 → #4 Congrats to @AnthropicAI on this release! 🔗 View Quoted Tweet 💬 22 🔄 9 ❤️ 271 👀 24195 📊 38 ⚡