Arena 公布 Claude Sonnet 5.5 (High) 的实测结果,以 1699 分位列 Code Arena: WebDev 第 4,比 Sonnet 5 (High) 的 1540 分提升 159 分。
Arena 的实测榜单数据显示该模型以约 1/5 的成本进入 WebDev 前四,读者可据此权衡性价比选型。
Real-world results are in for Claude Sonnet 5.5 (High) by @AnthropicAI. It just landed #4 in Code Arena: WebDev with 1699 pts, and has reshaped the Pareto frontier with its cost efficiency!
Claude Sonnet 5.5 (High) delivers nearly top performance at a blended $8 per Mtoken, reshaping the Pareto frontier! This model is 80% cheaper than both Claude Fable 5.1 (Max) in the #3 spot overall, and GPT-6 Astra (Max) at #2.
See Pareto placement below.
Overall, Claude Sonnet 5.5 (High) is a +159 pt improvement from Sonnet 5 (High) at #37 with 1540 pts. This gain compared to its previous variant also shows up across these key domains so far:
- Reference-Based Design: #38 → #4
- Simulations: #37 → #4
- Gaming: #36 → #4
Congrats to @AnthropicAI on this release!
Introducing Claude Sonnet 5.5, the second model in the Claude 5.5 family. It’s a clear upgrade over Sonnet 5, runs more than 30% faster, and costs up to 30% less for most work.在 X 查看被引用的帖子
来源:Arena.ai · x.com