Today for AI

The Decoder · 10/7/2026, 6:49:23 PM

Anthropic Launches Claude Haiku 5.5 with 90% Input Cost Cuts

By Matthias Bastian
78AI Score
Executive Summary

Anthropic released Claude Haiku 5.5, cutting average costs by 75% and input prices by up to 90% for short prompts. The company also slashed pricing for Sonnet 5.5 and updated SDKs to support computer and browser use.

SOURCE COVERAGEOriginal coverage

Contents3 sections

Anthropic has released Claude Haiku 5.5, the company's fastest and most affordable small model to date. Benchmark results show a major performance jump over its predecessor, and Anthropic is also cutting prices for Sonnet 5.5.

Haiku 5.5 is designed for high-volume, cost-sensitive tasks like summarization, database queries, classification, and live customer support, according to Anthropic. On average, the model costs about 75 percent less than Haiku 4.5. For requests with prompts up to 100,000 tokens, which Anthropic says account for roughly 90 percent of all previous Haiku requests, prices drop by up to 90 percent. Prompts longer than 100,000 tokens cost five times as much.

Price per 1 million tokensHaiku 5.5 (prompts up to / over 100k)Haiku 4.5Sonnet 5.5
Cache Reads0.01/0.01 / 0.05$0.10$0.10
Cache Writes0.125/0.125 / 0.625$1.25$2.50
Input Tokens0.10/0.10 / 0.50$1.00$2.00
Output tokens0.50/0.50 / 2.50$5.00$10.00

Anthropic points out that Haiku 5.5 uses an updated tokenizer that consumes slightly more tokens per task than its predecessor. The same thing happened with the Opus 4.x models, where token usage jumped about 30 percent from the tokenizer change alone. Real-world savings are likely smaller than the per-token prices suggest.

Benchmarks show a big leap over Haiku 4.5

Haiku 5.5 scores 1,620 on the knowledge benchmark GDPval-AA v2.1, more than double the 735 its predecessor managed. On Humanity's Last Exam, it hits 45.9 percent without tools and 57.4 percent with tools, up from 10.2 and 18.7 percent.

The biggest jump is in computer use, where the model operates a computer on its own. Since computer use burns through large amounts of tokens, a cheap, fast model like Haiku 5.5 is a natural fit. It scores 72.4 percent on OSWorld-2.1, up from 15.7 percent. On the agentic coding benchmark Terminal-Bench 4.0, it reaches 39.2 percent while Haiku 4.5 scored zero.

Anthropic also lists OpenAI's budget model GPT-6 Luna as a comparison, and Haiku 5.5 leads across every tested category. Sonnet 5.5 reference scores show, however, that Haiku 5.5 still falls well behind Anthropic's larger model.

Haiku 5.5Haiku 4.5GPT-6 LunaSonnet 5.5
Knowledge work / GDPval-AA v2.11,6207351,4371,840
Knowledge Work / AA-Briefcase v1.11,5786141,3361824
Computer Use / OSWorld 2.172.4% (offline subset)15.7% (offline subset)48.9% (offline subset)83.9% (offline subset)
Multidisciplinary reasoning / Humanity's Last Exam45.9% (no tools)10.2% (no tools)—56.9% (no tools)
57.4% (with tools)18.7% (with tools)—64.5% (with tools)
Agentic coding / Terminal-Bench 4.039.2%0.0%16.4%70.6%
Agentic coding / FrontierCode 1.1 (Main)46.4%—42.4%52.1% (Xhigh)
Visual reasoning / Chartography46.4% (no tools)6.4% (no tools)29.1% (no tools)61.6% (no tools)

First Haiku model lets users trade cost for performance

Haiku 5.5 is the first Haiku-class model with adjustable reasoning levels, letting users balance cost against quality. Anthropic says the model works best for narrowly scoped tasks like compaction, summarization, or sub-agent work. For complex agentic coding, Sonnet 5.5 and Opus 5.5 remain the better picks.

Cybersecurity safeguards are tighter than on the predecessor but allow a broader range of defensive tasks than Sonnet 5.5, partly because the model is less capable overall. Penetration testing stays blocked. Organizations with broader needs can apply for Anthropic's verification programs for life sciences and cybersecurity.

Haiku 5.5 is available now across all platforms, including Amazon Web Services, Google Cloud, and Microsoft Azure.

Sonnet 5.5 gets cheaper too, and Anthropic hands out API credits

Alongside the Haiku launch, Anthropic is cutting cache read costs for Sonnet 5.5 by 50 percent, from 0.20to0.20 to 0.10 per million tokens. The company says this should reduce costs for most agentic tasks by about 20 percent. The move is almost certainly a reaction to OpenAI's new GPT-6.1 series, showing that AI price wars are being fought harder than ever.

Anthropic is also rolling out monthly API credits. Max-5x subscribers get 100,Max−20xsubscribersget100, Max-20x subscribers get 200, and Team subscribers receive up to $500 per month. Users can spend the credits to experiment with tools, apps, and agents through the API.

The company is updating its Python and TypeScript SDKs as well, adding beta support for computer use and browser use.