Hacker News — AI on Front Page · · 3 min read

Claude Haiku 5.5

Mirrored from Hacker News — AI on Front Page for archival readability. Support the source by reading on the original site.

207 pts · 91 comments on Hacker News

Introducing Claude Haiku 5.5: the cheapest, fastest, and most capable small model we’ve ever released.

Claude Haiku 5.5 is designed for high-volume, cost-sensitive tasks. It reliably handles quick and repetitive workloads (like summaries, compactions, database queries, and classification requests). It pairs well with Opus 5.5 and Sonnet 5.5 as a subagent on coding work. And, since it’s also our fastest model to date, it works especially well for speed-sensitive tasks like live customer support and browser use.¹

Haiku 5.5 is available at a much lower price than Haiku 4.5. On average, it now costs around 75% less to run.²

Along with this launch, we’re making improvements to the value of our model range. We’re halving the price of Claude Sonnet 5.5’s cache reads, which means Sonnet 5.5 now runs around 20% cheaper on most agentic work. And we’re introducing a new monthly API credit for our Claude Max and Team subscribers, designed to support our users in building new agents and applications that run on the Claude Platform.

Performance

Here’s how Claude Haiku 5.5 performs across a range of benchmarks:

Haiku 5.5Haiku 4.5GPT-6 LunaSonnet 5.5For reference
Knowledge workGDPval-AA v2.1
Knowledge workGDPval-AA v2.1162073514371840
Knowledge workAA-Briefcase v1.1
Knowledge workAA-Briefcase v1.1157861413361824
Computer useOSWorld 2.1
Computer useOSWorld 2.172.4%Offline subset15.7%Offline subset48.9%Offline subset83.9%Offline subset
Multidisciplinary reasoningHumanity’s Last Exam
Multidisciplinary reasoningHumanity’s Last Exam45.9%no tools10.2%no tools—56.9%no tools
57.4%with tools18.7%with tools—64.5%with tools
Agentic codingTerminal-Bench 4.0
Agentic codingTerminal-Bench 4.039.2%0.0%16.4%70.6%
Agentic codingFrontierCode 1.1 (Main)
Agentic codingFrontierCode 1.1 (Main)46.4%—42.4%52.1%Xhigh
Visual reasoningChartography
Visual reasoningChartography46.4%no tools6.4%no tools29.1%no tools61.6%no tools

For details on how we run our evaluations, see the Haiku 5.5 System Card.

Haiku 5.5 is our first Haiku-class model to come with an adjustable effort setting. This means that, as with our other models, users can decide whether to optimize for cost or intelligence. The charts below show how Haiku 5.5 performs on three benchmarks at each effort setting:

Computer use: OSWorldKnowledge work: GDPval-AAMultidisciplinary reasoning: Humanity’s Last Exam
OSWorld 2.1 (offline subset)Accuracy vs. cost

OSWorld 2.1 measures how well agents can operate a real computer to finish long, multi-step tasks.

GDPval-AA v2.1Accuracy vs. cost

Artificial Analysis’s GDPval-AA v2.1 evaluates agents on real-world professional work across 44 occupations.

Humanity’s Last Exam (no tools)Accuracy vs. cost

Humanity’s Last Exam (HLE) is a test of expert-level academic knowledge and reasoning.

In early testing, our customers reported results consistent with the performance and cost improvements shown above. Here’s what they told us about the new model:

AsanaHubSpotAlphaSenseBoxRogoCognition
Quote

“We’re very impressed with Claude Haiku 5.5, particularly its speed. We ran it through our eval suite for AI Teammates, our AI agent product, covering use cases like triaging bugs, setting up projects, and searching large portfolios to surface high-risk or overdue work. Compared with the model we use today, we saw over a 30% reduction in latency for task completions and up to 2.5x faster inference per agent turn. It’s a noticeably snappier experience.”

CompanyAsana
AuthorAaron Vinh, Staff Software Engineer

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from Hacker News — AI on Front Page