跳到正文
北京时间
原文
Arena.ai· @arena · X·· 4 小时前精选AI 评分71
AI 导读

Arena 公布 Anthropic Claude Haiku 5.5 (High) 的真实评测结果,在 Code Arena: WebDev 以 1587 分首秀排名第 30,略在 Pareto 前沿之外。

推荐理由

榜单方给出 Haiku 5.5 在 Code Arena 的实际得分与各档 Claude 模型的成本分差,便于横向权衡性价比。

正文 · 原文

Real-world results are in for Claude Haiku 5.5 (High) by @AnthropicAI.

Debuting at #30 with 1587 pts in Code Arena: WebDev, this release sits just outside the Pareto frontier while still offering strong cost efficiency. It matches GPT‑6 Luna’s price at $0.10/$0.50 per 1M input/output tokens while scoring six pts higher.

Compared with other Claude models, Haiku 5.5 offers:
- 90% lower cost than Haiku 4.5 at +257 pts
- 95% lower cost than Sonnet 5.5 (High) at -128 pts
- 98% lower cost than Opus 5 (High) at -70 pts
- 99% lower cost than Fable 5.1 (Max) at -157 pts

Congrats to the team at @AnthropicAI on this release!

引用Claude@claudeai
Introducing Claude Haiku 5.5: the cheapest, fastest, and most capable small model we’ve ever released. On average, it costs around 75% less to run than Claude Haiku 4.5.
在 X 查看被引用的帖子

来源:Arena.ai · x.com