Claude Haiku 5.5
Mirrored from Hacker News — AI on Front Page for archival readability. Support the source by reading on the original site.
Introducing Claude Haiku 5.5: the cheapest, fastest, and most capable small model we’ve ever released.
Claude Haiku 5.5 is designed for high-volume, cost-sensitive tasks. It reliably handles quick and repetitive workloads (like summaries, compactions, database queries, and classification requests). It pairs well with Opus 5.5 and Sonnet 5.5 as a subagent on coding work. And, since it’s also our fastest model to date, it works especially well for speed-sensitive tasks like live customer support and browser use.¹
Haiku 5.5 is available at a much lower price than Haiku 4.5. On average, it now costs around 75% less to run.²
Along with this launch, we’re making improvements to the value of our model range. We’re halving the price of Claude Sonnet 5.5’s cache reads, which means Sonnet 5.5 now runs around 20% cheaper on most agentic work. And we’re introducing a new monthly API credit for our Claude Max and Team subscribers, designed to support our users in building new agents and applications that run on the Claude Platform.
Performance
Here’s how Claude Haiku 5.5 performs across a range of benchmarks:
| Haiku 5.5 | Haiku 4.5 | GPT-6 Luna | Sonnet 5.5For reference | ||
|---|---|---|---|---|---|
| Knowledge workGDPval-AA v2.1 | |||||
| Knowledge workGDPval-AA v2.1 | 1620 | 735 | 1437 | 1840 | |
| Knowledge workAA-Briefcase v1.1 | |||||
| Knowledge workAA-Briefcase v1.1 | 1578 | 614 | 1336 | 1824 | |
| Computer useOSWorld 2.1 | |||||
| Computer useOSWorld 2.1 | 72.4%Offline subset | 15.7%Offline subset | 48.9%Offline subset | 83.9%Offline subset | |
| Multidisciplinary reasoningHumanity’s Last Exam | |||||
| Multidisciplinary reasoningHumanity’s Last Exam | 45.9%no tools | 10.2%no tools | — | 56.9%no tools | |
| 57.4%with tools | 18.7%with tools | — | 64.5%with tools | ||
| Agentic codingTerminal-Bench 4.0 | |||||
| Agentic codingTerminal-Bench 4.0 | 39.2% | 0.0% | 16.4% | 70.6% | |
| Agentic codingFrontierCode 1.1 (Main) | |||||
| Agentic codingFrontierCode 1.1 (Main) | 46.4% | — | 42.4% | 52.1%Xhigh | |
| Visual reasoningChartography | |||||
| Visual reasoningChartography | 46.4%no tools | 6.4%no tools | 29.1%no tools | 61.6%no tools |
For details on how we run our evaluations, see the Haiku 5.5 System Card.
Haiku 5.5 is our first Haiku-class model to come with an adjustable effort setting. This means that, as with our other models, users can decide whether to optimize for cost or intelligence. The charts below show how Haiku 5.5 performs on three benchmarks at each effort setting:
OSWorld 2.1 measures how well agents can operate a real computer to finish long, multi-step tasks.
Artificial Analysis’s GDPval-AA v2.1 evaluates agents on real-world professional work across 44 occupations.
Humanity’s Last Exam (HLE) is a test of expert-level academic knowledge and reasoning.
In early testing, our customers reported results consistent with the performance and cost improvements shown above. Here’s what they told us about the new model:
“We’re very impressed with Claude Haiku 5.5, particularly its speed. We ran it through our eval suite for AI Teammates, our AI agent product, covering use cases like triaging bugs, setting up projects, and searching large portfolios to surface high-risk or overdue work. Compared with the model we use today, we saw over a 30% reduction in latency for task completions and up to 2.5x faster inference per agent turn. It’s a noticeably snappier experience.”
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.