ModelsPublished Top storyAll news from that day
Ongoing case: Anthropic's Claude models New model · 9 storiesAnthropic cuts small-model prices with Claude Haiku 5.5
Anthropic says Claude Haiku 5.5, released on 7 October, costs around 75% less to run on average than Haiku 4.5, with input at $0.10 per million tokens up to 100,000.
Check our sources · 12 facts from 2 sources
Image: Anthropic
Key points
- For prompts up to 100,000 tokens Anthropic lists $0.10 per million input tokens and $0.50 for output, against $1.00 and $5.00 for Haiku 4.5.
- Anthropic's documentation says the same text counts as approximately 30% more tokens than on Haiku 4.5, so token budgets tuned for the older model may need recounting.
- On Anthropic's own table Haiku 5.5 scores 72.4% on the offline subset of OSWorld 2.1, against 15.7% for Haiku 4.5 and 83.9% for Sonnet 5.5.
What happened
Anthropic released Claude Haiku 5.5 on 7 October 2026, with the model ID claude-haiku-5-5. Anthropic calls it the cheapest, fastest and most capable small model it has released, and says it is the first Haiku-class model with an adjustable effort setting.
For prompts up to 100,000 tokens Anthropic lists $0.10 per million input tokens and $0.50 per million output tokens, against $1.00 and $5.00 for Haiku 4.5. For prompts over 100,000 tokens the price is $0.50 input and $2.50 output. Cache reads are $0.01 per million tokens up to 100,000 tokens and $0.05 above that.
Anthropic says the model costs around 75% less to run on average than Haiku 4.5, and that the cut is 90% for requests up to 100,000 tokens and 50% for longer ones.
Anthropic's documentation gives a 1M-token context window and up to 128k output tokens. The same text counts as about 30% more tokens than on Haiku 4.5, so token budgets tuned for the older model need recounting. Anthropic says the model is available on the Claude Platform, Amazon Web Services, Google Cloud and Microsoft Azure.
On Anthropic's own table, Haiku 5.5 scores 72.4% on the offline subset of OSWorld 2.1, against 15.7% for Haiku 4.5, 48.9% for GPT-6 Luna and 83.9% for Sonnet 5.5. These are Anthropic's figures, not an outside test.
What it means for you
Our viewThe price cut is large, but the 75% average and the 90% figure for short prompts are Anthropic's own. The 30% higher token count for the same text bears on the 90% per-token price cut, and the page does not state whether the 75% already counts it.
Prompts over 100,000 tokens cost $0.50 for input and $2.50 for output, and Anthropic puts the cut there at 50%, so long inputs save less. The gap to Sonnet 5.5 on OSWorld 2.1 is wide, so a small model fits simple, high-volume steps.
If you run Haiku 4.5, rerun a sample of your real prompts, compare the cost at the new token count, and check quality against your current results before switching.
It adds no new facts.
Your reaction
We count reactions per story and day, never who reacted. The counts help us choose what goes in the monthly issue. If you are signed in, your own page shows yours too.
Check our sources
Every sentence above is checked against these 2 sources.
1 Introducing Claude Haiku 5.5
Open the source archived copy-
Anthropic published Introducing Claude Haiku 5.5 dated October 7, 2026, with the model ID claude-haiku-5-5. Quote: "October 7, 2026"
citeOctober 7, 2026
-
Anthropic describes Claude Haiku 5.5 as the cheapest, fastest and most capable small model it has released. Quote: "Introducing Claude Haiku 5.5: the cheapest, fastest, and most capable small model"
citeIntroducing Claude Haiku 5.5: the cheapest, fastest, and most capable small model we’ve ever released.
-
Anthropic says Haiku 5.5 is its first Haiku-class model with an adjustable effort setting. Quote: "Haiku 5.5 is our first Haiku-class model to come with an adjustable effort setting."
citeHaiku 5.5 is our first Haiku-class model to come with an adjustable effort setting.
-
For prompts up to 100,000 tokens the page lists Haiku 5.5 input at $0.10 per million tokens against $1.00 for Haiku 4.5; over 100,000 tokens it lists $0.50. Quote: "Input tokens $0.10 / $0.50 $1.00 $2.00"
citeInput tokens $0.10 / $0.50 $1.00 $2.00
-
For prompts up to 100,000 tokens the page lists Haiku 5.5 output at $0.50 per million tokens against $5.00 for Haiku 4.5; over 100,000 tokens it lists $2.50. Quote: "Output tokens $0.50 / $2.50 $5.00 $10.00"
citeOutput tokens $0.50 / $2.50 $5.00 $10.00
-
The page lists Haiku 5.5 cache reads at $0.01 per million tokens for prompts up to 100,000 tokens and $0.05 above that. Quote: "Cache reads $0.01 / $0.05 $0.10 $0.10"
citeCache reads $0.01 / $0.05 $0.10 $0.10
-
Anthropic says Haiku 5.5 now costs around 75% less to run on average than Haiku 4.5. Quote: "On average, it now costs around 75% less to run."
citeHaiku 5.5 is available at a much lower price than Haiku 4.5. On average, it now costs around 75% less to run.
-
Anthropic says Haiku 5.5 is priced 90% lower than Haiku 4.5 for requests up to 100,000 tokens and 50% lower for requests over 100,000 tokens. Quote: "Claude Haiku 5.5 is priced 90% lower than Claude Haiku 4.5 for requests up to 100,000 tokens, and 50% lower for requests over 100,000 tokens."
citeClaude Haiku 5.5 is priced 90% lower than Claude Haiku 4.5 for requests up to 100,000 tokens, and 50% lower for requests over 100,000 tokens.
-
Anthropic says Haiku 5.5 is available on Amazon Web Services, Google Cloud and Microsoft Azure, and on the Claude Platform as claude-haiku-5-5. Quote: "Claude Haiku 5.5 is available now on all platforms, including Amazon Web Services, Google Cloud, and Microsoft Azure."
citeClaude Haiku 5.5 is available now on all platforms, including Amazon Web Services, Google Cloud, and Microsoft Azure. On the Claude Platform, developers can get started with claude-haiku-5-5 .
-
On Anthropic's benchmark table Haiku 5.5 scores 72.4% on OSWorld 2.1 (offline subset), Haiku 4.5 15.7%, GPT-6 Luna 48.9% and Sonnet 5.5 83.9%. Quote: "72.4% Offline subset 15.7% Offline subset 48.9% Offline subset 83.9% Offline subset"
citeComputer use OSWorld 2.1 72.4% Offline subset 15.7% Offline subset 48.9% Offline subset 83.9% Offline subset
2 Claude Haiku 5.5 overview
Open the source archived copy-
Anthropic documentation gives Haiku 5.5 a 1M token context window and up to 128k output tokens. Quote: "a 1M token context window, and up to 128k output tokens"
citeIt supports adaptive thinking with the effort parameter, a 1M token context window, and up to 128k output tokens.
-
The documentation says the same text counts as approximately 30% more tokens on Haiku 5.5 than on Haiku 4.5. Quote: "the same text counts as approximately 30% more tokens than on Claude Haiku 4.5"
citeIt uses the same newer tokenizer as Claude 4.7 and later models, so the same text counts as approximately 30% more tokens than on Claude Haiku 4.5.
Topics
The morning email
On the mornings we publish: the three top stories and up to four short ones. Free.
We email you a link to confirm. An issue may include one sponsor, always labelled Sponsored · Advertisement. Our emails count opens and clicks, not who made them. Unsubscribe in one click. What we keep