Arena 公布 Anthropic Claude Haiku 5.5 (High) 的真实评测结果,在 Code Arena: WebDev 以 1587 分首秀排名第 30,略在 Pareto 前沿之外。
榜单方给出 Haiku 5.5 在 Code Arena 的实际得分与各档 Claude 模型的成本分差,便于横向权衡性价比。
Real-world results are in for Claude Haiku 5.5 (High) by @AnthropicAI.
Debuting at #30 with 1587 pts in Code Arena: WebDev, this release sits just outside the Pareto frontier while still offering strong cost efficiency. It matches GPT‑6 Luna’s price at $0.10/$0.50 per 1M input/output tokens while scoring six pts higher.
Compared with other Claude models, Haiku 5.5 offers:
- 90% lower cost than Haiku 4.5 at +257 pts
- 95% lower cost than Sonnet 5.5 (High) at -128 pts
- 98% lower cost than Opus 5 (High) at -70 pts
- 99% lower cost than Fable 5.1 (Max) at -157 pts
Congrats to the team at @AnthropicAI on this release!
Introducing Claude Haiku 5.5: the cheapest, fastest, and most capable small model we’ve ever released. On average, it costs around 75% less to run than Claude Haiku 4.5.在 X 查看被引用的帖子
来源:Arena.ai · x.com