Anthropic 发布 Claude Haiku 5.5,成本较 Haiku 4.5 降低 90%
Anthropic 发布 Claude Haiku 5.5,在最高 100K token 的提示词上成本比 Haiku 4.5 低 90%,并新增首个 Haiku effort 设置,可在成本与准确率之间取舍;官方仍推荐 Sonnet 5.5 和 Opus 5.5 处理 Terminal-Bench 4.0 这类复杂 agentic coding 任务。
Anthropic dropped Haiku 5.5
> costs 90% less than Haiku 4.5 on prompts up to 100K tokens.
> Haiku 5.5 adds the first Haiku effort setting, trading cost for accuracy, but Anthropic still recommends Sonnet 5.5 and Opus 5.5 for complex agentic coding such as Terminal-Bench 4.0 tasks.
> Asana reported over 30% lower task-completion latency and up to 2.5x faster inference per agent turn than its current model.
> On OSWorld 2.1, which tests whether an AI agent can operate a real computer to finish long multi-step tasks, Haiku 5.5 jumped from Haiku 4.5's 15.7% to 72.4%.
On Chartography, a visual reasoning test of reading and interpreting charts without tools, it rose from 6.4% to 46.4%.
来源:rohanpaul_ai · x.com