Anthropic 发布 Claude Haiku 5.5,官方称是其史上最便宜、最快、能力最强的小模型,平均运行成本比 Haiku 4.5 低约 75%,10 万 token 以内的提示成本降低 90%。
Anthropic dropped Haiku 5.5
costs 90% less than Haiku 4.5 on prompts up to 100K tokens.
Haiku 5.5 adds the first Haiku effort setting, trading cost for accuracy, but Anthropic still recommends Sonnet 5.5 and Opus 5.5 for complex agentic coding such as Terminal-Bench 4.0 tasks.
Asana reported over 30% lower task-completion latency and up to 2.5x faster inference per agent turn than its current model.
On OSWorld 2.1, which tests whether an AI agent can operate a real computer to finish long multi-step tasks, Haiku 5.5 jumped from Haiku 4.5's 15.7% to 72.4%.
On Chartography, a visual reasoning test of reading and interpreting charts without tools, it rose from 6.4% to 46.4%.
Introducing Claude Haiku 5.5: the cheapest, fastest, and most capable small model we’ve ever released. On average, it costs around 75% less to run than Claude Haiku 4.5.在 X 查看被引用的帖子
来源:Rohan Paul · x.com