AIToday
Large Language ModelsAI Business & IndustryImpress WatchPublished: Oct 8, 2026, 13:00 JST

Anthropic ships Haiku 5.5 with 75% lower running cost

Anthropic ships Haiku 5.5 with 75% lower running cost

3 Key Points

  1. What happened

    Anthropic began offering Claude Haiku 5.5, its smallest and fastest model, with average operating costs roughly 75% lower than Haiku 4.5, priced at $0.1 per million input tokens and $0.5 for output.

  2. Why it matters

    For businesses running high-volume tasks such as summarization, data compression, database queries and classification, the lower cost per task could make large-scale AI operations substantially cheaper — provided the model's quality holds up in production.

  3. What to watch

    The pricing edge hinges on whether the newer model's quality claims translate into real-world workloads, since its benchmark scores still trail the more expensive Sonnet 5.5. Also note that once a request exceeds 100,000 tokens, input pricing rises to $0.5 and output to $2.5 per million.

WHO IT HITSCost-sensitive engineering and data teams running high-volume AI workloads — such as summarization, classification and database queries — stand to gain from the reduced token pricing, and the model's availability on AWS, Google Cloud and Microsoft Azure may make it easier for enterprise buyers already using those platforms.

Not sure about something? Ask the AI

Questions and answers are published on this page.

Summaries like this, in your inbox every morning.

Context & Analysis

Anthropic's Claude line has seen a steady cadence of updates to higher-end models such as Fable, Opus and Sonnet, while Haiku — the smallest and fastest tier — had gone roughly a year without a refresh. The arrival of Haiku 5.5 therefore fills a gap at the cheap end of the lineup, where customers doing repetitive, high-volume work had been running an older generation.

The model is designed for tasks where speed and cost matter more than peak reasoning: summarization, data compression, database queries and classification. In coding, it can act as a sub-agent alongside Opus 5.5 or Sonnet 5.5, handling cheaper steps so the expensive models are reserved for harder work. Its benchmark scores show a large jump over Haiku 4.5 and even a lead over GPT-6 Luna, though they remain below Sonnet 5.5.

Haiku 5.5 also introduces effort-level settings to the Haiku line for the first time, letting users trade cost against capability. The real test is whether the price advantage holds for large production workloads, where token volumes can push requests past the 100,000-token threshold — a point at which the cost per token rises, though it is still half of Haiku 4.5. Separately, Sonnet 5.5 now has its cache-read fee cut in half, which Anthropic says makes it about 20% cheaper to run for most agent work.

FAQ
How does Haiku 5.5’s pricing compare to Haiku 4.5?
Anthropic says average operating costs are about 75% lower. Standard pricing is $0.1 per million input tokens and $0.5 for output; above 100,000 tokens input rises to $0.5 and output to $2.5, still half of Haiku 4.5.
Where can I use Claude Haiku 5.5?
It launched immediately on Amazon Web Services, Google Cloud and Microsoft Azure. The API model name on Claude is claude-haiku-5-5.
What benchmarks did Haiku 5.5 score?
It scored 1,620 on GDPval-AA v2.1 and 1,578 on AA-Briefcase v1.1. On OSWorld 2.1 it reached 72.4%, compared with 15.7% for Haiku 4.5 and 48.9% for GPT-6 Luna.

Also reported by AI Watch (Impress), GIGAZINE AI, SiliconANGLE AI, Simon Willison's Weblog, THE DECODER

AI news that matters for your work, delivered every morning.

Pick your industry and the AI tools you use, and get news related to your work every day.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.

Questions and answers are published on this page.

Related Articles

Next articleAnthropic's Claude Haiku 5.5 beats GPT-6 Luna, costs 75% less