Breaking
DEDriver crashes into oncoming mopeds – four people dieCNUnraveling the Centennial Mystery of "Homochirality" French and Japanese Scholars Win Nobel Prize in ChemistryBRCar rolls over and ends up with its wheels up after driver loses control in Presidente PrudenteTR16-year-old boy stabbed in Gaziantep diedRUThe Ministry of Emergency Situations named the main mistake in the threat of a UAV attackRUTemporary restrictions have been introduced at Nizhny Novgorod airportBRUnified Competition will have tests on Saturday (10) and Sunday (11) in Piauí; salaries reach R$13.5 thousandTROffice 2021 Support Ends: Here are 5 Options You Can MakeRUThe UN is concerned about the fate of Russians in the face of Ukrainian attacksRUDemocratic Party sues Trump administration over adDEDriver crashes into oncoming mopeds – four people dieCNUnraveling the Centennial Mystery of "Homochirality" French and Japanese Scholars Win Nobel Prize in ChemistryBRCar rolls over and ends up with its wheels up after driver loses control in Presidente PrudenteTR16-year-old boy stabbed in Gaziantep diedRUThe Ministry of Emergency Situations named the main mistake in the threat of a UAV attackRUTemporary restrictions have been introduced at Nizhny Novgorod airportBRUnified Competition will have tests on Saturday (10) and Sunday (11) in Piauí; salaries reach R$13.5 thousandTROffice 2021 Support Ends: Here are 5 Options You Can MakeRUThe UN is concerned about the fate of Russians in the face of Ukrainian attacksRUDemocratic Party sues Trump administration over ad
BackAnthropic Launches Haiku 5.5: Its Cheapest and Fastest Claude Model Yet
Anthropic Launches Haiku 5.5: Its Cheapest and Fastest Claude Model Yet
Tech
Decrypt1 hour agoTech2 min read

Anthropic Launches Haiku 5.5: Its Cheapest and Fastest Claude Model Yet

Quick Look

  • Anthropic released Claude Haiku 5.5, positioning it as the cheapest, fastest, and most capable small model for high-volume tasks like document summarization and live customer support.
  • Priced at $0.10 per million input and $0.50 per million output tokens, it offers up to 75% average savings over Haiku 4.5 and matches OpenAI's GPT-6 Luna pricing.
  • Benchmarks show Haiku 5.5 outperforms Luna on OSWorld 2.1 (72.4% vs 48.9%) and Terminal-Bench 4.0 (39.2% vs 16.4%), while scoring 1620 on GDPval-AA v2.1.

AI-generated summary

Why It Matters

Anthropic released Claude Haiku 5.5 as part of its Claude 5.5 model series, following Opus 5.5 on September 22 and Sonnet 5.5 on September 28. The model is positioned as a cost-effective, high-speed option for enterprise AI tasks.

Font size

Anthropic released Claude Haiku 5.5 on Wednesday, calling it the cheapest, fastest and most capable small model it has made. It targets high-volume chores such as summarizing documents and querying databases, plus speed-sensitive jobs like live customer support and operating a web browser for the user.

Developers who plug Claude into their own products pay $0.10 per million input tokens and $0.50 per million output tokens for prompts up to 100,000 tokens. Tokens are the small chunks of text, roughly three-quarters of a word, that AI models read and write—and they’re what big AI companies bill by. That matches the rate OpenAI set for GPT-6 Luna, its rival small model, when it launched on September 22.

Haiku 4.5 charged $1 and $5, so the new rate is 90% lower, while prompts above 100,000 tokens get a 50% cut. Since about 90% of requests to the old model fell under that line, and Haiku 5.5 splits text into slightly more tokens, Anthropic puts the average saving near 75%.

Customer-support chats and long-email summaries are the jobs Anthropic built Haiku for. So essentially repetitive, easy, non creative tasks are the best fit for this model.

That said, benchmarks show this is not a dumb model by any means. For example, on OSWorld 2.1, a test of whether an AI can operate a real computer through long, multi-step tasks, scored as a partial-credit percentage, Haiku 5.5 beats OpenAI’s GPT Luna scoring 72.4% success rate vs 48.9%.

Terminal-Bench 4.0 hands AI agents professional tasks to finish by typing commands on their own, and scores the share completed correctly on the first attempt. Haiku 5.5 landed at 39.2%, against 16.4% for OpenAI’s Luna and 0% for Haiku 4.5

For comparison, Anthropic’s Claude Sonnet 5.5 scored 70.6%.

We tried the model on a simple logic question and it was extremely fast to respond, almost instantly. It also gave the wrong answer, which is not what you want from a model, so be careful about blindly trusting it.

On GDPval-AA v2.1, which rates models on real professional work across 44 occupations using an Elo scale, the head-to-head rating system borrowed from chess, Haiku 5.5 scored 1620. Luna got 1437 and Haiku 4.5 got 735.

This is also the first Haiku model with an adjustable effort setting, a dial that lets users trade cost for smarter answers. Anthropic also halved Sonnet 5.5's cache-read price to $0.10 per million tokens, the discounted rate for text a model has already processed.

Haiku 5.5 comes 15 days after Opus 5.5 on September 22 and nine days after Sonnet 5.5 on September 28, the last of the three Claude 5.5 models Anthropic had promised. Opus 5.5 was the first release since CEO Dario Amodei published an essay urging the industry to slow gains in AI capabilities.

Haiku 5.5 is available now on the Claude website, Amazon Web Services, Google Cloud and Microsoft Azure under the name claude-haiku-5-5. Anthropic is also rolling out monthly API credits this week: $100 for subscribers on the Max 5x plan, $200 for Max 20x, and up to $500, pooled across users, for Team plans.

What to Watch

AI outlook — possibilities, not facts

  • Anthropic will gain enterprise market share in AI model usage due to Haiku 5.5's pricing and performance advantages.

    Likely · Within months

  • OpenAI may respond with a price adjustment or performance update to GPT-6 Luna to maintain competitiveness.

    Possible · Within months

Open Questions

  • How will Haiku 5.5's adjustable effort setting be adopted by developers in practice?
  • What is the long-term impact of Anthropic's monthly API credit program on user retention and usage?

Related Topics

This article was originally published by Decrypt.

Related Stories

Google Lets You Build a Video Game by Typing a Description
Tech·

Google Lets You Build a Video Game by Typing a Description

Google launched Playground, a web platform allowing users to create playable games by describing them in plain language, with refinement via chat and sharing via link. Available to U.S. users 18+, access tiers depend on Google AI subscription. Games run in browsers, support privacy or community publishing with safety screening, and may integrate Unity Spark for advanced features. The launch follows Google's Genie AI projects and coincided with a drop in game stocks as investors worry about AI disrupting traditional game development tools.

Decrypt
2 min read
Mistral AI Drops 'Le Chonk': A Massive AI Model Named After a Cat Meme
Tech·

Mistral AI Drops 'Le Chonk': A Massive AI Model Named After a Cat Meme

Mistral AI launched Mistral Large 4, a 1-trillion-parameter AI model using a mixture-of-experts design with 49 billion active parameters per query. The model is positioned as a competitive open-weight alternative to Claude Opus 5.5 and GPT-6 Astra, with pricing at $1.36 per million input tokens and $4.18 per million output tokens. Benchmarks show strong performance in coding and automation tasks, trailing only Kimi K3 and Gemini 4 Argon in some tests. Mistral plans to release the model weights by end of October.

Decrypt
2 min read
More on this topicclaude haiku 5.5