Claude Haiku 5.5 is now available in CRHQ

Anthropic's new small model is far more capable than Haiku 4.5 and is priced 50% to 90% lower. It replaces Haiku 4.5 on your CRHQ server.

Claude Haiku 5.5 is now in every model chooser on your CRHQ server. Anthropic calls it the cheapest, fastest and most capable small model it has released. It replaces Haiku 4.5: anything that used Haiku 4.5 now uses Haiku 5.5 automatically, with nothing for you to change.

Claude Haiku 5.5 selected at the end of the Claude list in the CRHQ model picker, below Fable 5.1, Fable 5 and Haiku (latest).

A big step up from Haiku 4.5

These are Anthropic's reported results. The chart compares Haiku 5.5 with the Haiku it replaces and with GPT-6 Luna, OpenAI's small model. Sonnet 5.5 is shown for reference.

Haiku 5.5 against the previous Haiku and GPT-6 Luna (higher is better)
Haiku 5.5Haiku 4.5GPT-6 LunaSonnet 5.5, for reference
GDPval-AA v2.1
Knowledge work (Elo rating, scale 0 to 2000)
Haiku 5.51620
Haiku 4.5735
GPT-6 Luna1437
Sonnet 5.51840
OSWorld 2.1
Computer use, offline subset
Haiku 5.572.4%
Haiku 4.515.7%
GPT-6 Luna48.9%
Sonnet 5.583.9%
Humanity's Last Exam (with tools)
Multidisciplinary reasoning, with tools
Haiku 5.557.4%
Haiku 4.518.7%
GPT-6 Lunan/a
Sonnet 5.564.5%
Chartography
Visual reasoning, no tools
Haiku 5.546.4%
Haiku 4.56.4%
GPT-6 Luna29.1%
Sonnet 5.561.6%
FrontierCode 1.1 (Main)
Agentic coding
Haiku 5.546.4%
Haiku 4.5n/a
GPT-6 Luna42.4%
Sonnet 5.552.1%
Terminal-Bench 4.0
Agentic coding in a terminal
Haiku 5.539.2%
Haiku 4.50.0%
GPT-6 Luna16.4%
Sonnet 5.570.6%
A big step up. On every benchmark where Anthropic reports both scores, Haiku 5.5 is far ahead of Haiku 4.5 and ahead of GPT-6 Luna. Sonnet 5.5 is still clearly stronger, most of all on agentic coding.
Source: Anthropic, Introducing Claude Haiku 5.5, October 7, 2026. These are Anthropic's reported results. n/a means Anthropic reports no score. The Sonnet 5.5 FrontierCode score is at Xhigh effort.
BenchmarkWhat it measuresHaiku 5.5Haiku 4.5GPT-6 LunaSonnet 5.5
GDPval-AA v2.1Knowledge work162073514371840
AA-Briefcase v1.1Knowledge work157861413361824
OSWorld 2.1Computer use72.4%15.7%48.9%83.9%
Humanity's Last Exam (no tools)Multidisciplinary reasoning45.9%10.2%n/a56.9%
Humanity's Last Exam (with tools)Multidisciplinary reasoning57.4%18.7%n/a64.5%
Terminal-Bench 4.0Agentic coding39.2%0.0%16.4%70.6%
FrontierCode 1.1 (Main)Agentic coding46.4%n/a42.4%52.1%
ChartographyVisual reasoning46.4%6.4%29.1%61.6%

The jump over Haiku 4.5 is large everywhere. On knowledge work the GDPval-AA rating more than doubles, from 735 to 1620, and on computer use the OSWorld score goes from 15.7% to 72.4%. Haiku 5.5 is also the first Haiku with an adjustable effort setting, so it can be tuned for cost or for intelligence.

Anthropic is also clear about where the bigger models still win: Sonnet 5.5 and Opus 5.5 remain the better choice for complex agentic coding. Haiku 5.5 is best at narrowly scoped work. Our suggestion is to use it for quick, high-volume jobs such as summaries, lookups, classification and helper agents, and to keep Sonnet 5.5 or Opus 5.5 for the heavy lifting.

Pricing

Price per 1M tokens: Haiku 5.5 has two prices, by prompt length (lower is better)
Haiku 5.5, prompts up to 100kHaiku 5.5, prompts over 100kHaiku 4.5Sonnet 5.5, for reference
Input
Per 1M input tokens
Haiku 5.5$0.10prompts up to 100k
Haiku 5.5$0.50prompts over 100k
Haiku 4.5$1.00any length
Sonnet 5.5$2.00any length
Output
Per 1M output tokens
Haiku 5.5$0.50prompts up to 100k
Haiku 5.5$2.50prompts over 100k
Haiku 4.5$5.00any length
Sonnet 5.5$10.00any length
Cache reads
Per 1M cached tokens read
Haiku 5.5$0.01prompts up to 100k
Haiku 5.5$0.05prompts over 100k
Haiku 4.5$0.10any length
Sonnet 5.5$0.10any length
50% to 90% lower. Haiku 5.5 is priced 90% lower than Haiku 4.5 for prompts up to 100,000 tokens, and 50% lower above that. Anthropic puts the average saving at around 75%.
Source: Anthropic, Introducing Claude Haiku 5.5 and Anthropic pricing. Each panel is scaled to its own highest price. Cache write prices are in the table below.
Per 1M tokensHaiku 5.5, prompts up to 100kHaiku 5.5, prompts over 100kHaiku 4.5Sonnet 5.5
Input$0.10$0.50$1.00$2.00
Output$0.50$2.50$5.00$10.00
Cache reads$0.01$0.05$0.10$0.10
Cache writes$0.125$0.625$1.25$2.50

Haiku 5.5 is priced by prompt length. Prompts up to 100,000 tokens get the lower price, and Anthropic says those made up around 90% of requests to the previous Haiku. Longer prompts still cost half of what Haiku 4.5 did.

One thing to know for long sessions: the 100,000 tokens are counted for each request, and the whole prompt counts, including context read from or written to the cache. In long agent sessions most requests are over 100,000 tokens, so the saving there is about half. Short, focused jobs get the full saving, about a tenth of the old price.

Haiku 5.5 has a 1 million token context window and can write up to 128,000 tokens in one reply. With this launch Anthropic also halved the price of Sonnet 5.5 cache reads, from $0.20 to $0.10 per million tokens.

What changes on your CRHQ server

  • Two Haiku options in every model chooser. Haiku 5.5 is pinned, so your choice stays on exactly this version. Haiku (latest) always follows the newest Haiku, which is now 5.5.
  • Haiku 4.5 is retired in CRHQ. Haiku 5.5 is far better at a lower price, so CRHQ no longer offers the old one.
  • The switch is automatic. Everything that was set to Haiku, including agents, sessions and scheduled jobs, now runs Haiku 5.5. You do not need to change anything. Older chats keep the model and cost they were recorded with.
  • Effort control now works for Haiku. Turn it down for cheap, fast answers or up for harder work. The default is medium.

Pricing is tracked per turn in each session's Costs tab, and Claude subscription rotation works exactly as before.

Also in this release

Agents can run longer commands. A command an agent runs can now take up to 10 minutes by default, and up to 18 minutes when the agent asks for more. A command that passes its limit is stopped and the agent is told. Work that needs longer than that is best set up as a scheduled job, which runs on its own and reports back when it is done.

Availability

FactValue
AvailabilityEvery CRHQ server
Live since2026-10-08
ProviderClaude
PricingFrom $0.10 in / $0.50 out per 1M tokens
Client action requiredNone