跳到正文
原文
Arena.ai· @arena · X·· 1 天前精选AI 评分78
AI 导读

Arena 公布 Claude Sonnet 5.5 (High) 的实测结果,以 1699 分位列 Code Arena: WebDev 第 4,比 Sonnet 5 (High) 的 1540 分提升 159 分。

推荐理由

Arena 的实测榜单数据显示该模型以约 1/5 的成本进入 WebDev 前四,读者可据此权衡性价比选型。

正文

Real-world results are in for Claude Sonnet 5.5 (High) by @AnthropicAI. It just landed #4 in Code Arena: WebDev with 1699 pts, and has reshaped the Pareto frontier with its cost efficiency!

Claude Sonnet 5.5 (High) delivers nearly top performance at a blended $8 per Mtoken, reshaping the Pareto frontier! This model is 80% cheaper than both Claude Fable 5.1 (Max) in the #3 spot overall, and GPT-6 Astra (Max) at #2.

See Pareto placement below.

Overall, Claude Sonnet 5.5 (High) is a +159 pt improvement from Sonnet 5 (High) at #37 with 1540 pts. This gain compared to its previous variant also shows up across these key domains so far:
- Reference-Based Design: #38 → #4
- Simulations: #37 → #4
- Gaming: #36 → #4

Congrats to @AnthropicAI on this release!

引用Claude@claudeai
Introducing Claude Sonnet 5.5, the second model in the Claude 5.5 family. It’s a clear upgrade over Sonnet 5, runs more than 30% faster, and costs up to 30% less for most work.
在 X 查看被引用的帖子

来源:Arena.ai · x.com