AIToday
Large Language ModelsAI Business & IndustryLatent SpacePublished: Oct 8, 2026, 19:01 JST

Anthropic ships Claude Haiku 5.5, priced to match GPT-6 Luna

Anthropic ships Claude Haiku 5.5, priced to match GPT-6 Luna

3 Key Points

  1. What happened

    Anthropic shipped Claude Haiku 5.5, its first Haiku-tier update in about a year. It is priced to match OpenAI's GPT-6 Luna and scores 43 on the Artificial Analysis Intelligence Index, up 26 points.

  2. Why it matters

    The launch is widely read as aimed at OpenAI's GPT-6 Luna, with observers such as @kimmonismus calling it "way better than GPT-6-Luna." Anthropic also cut Sonnet 5.5 cache-read prices the same day.

WHO IT HITSDevelopers building high-volume, cost-sensitive agent workflows — subagents, summaries, compactions and database queries — are the direct audience, since Anthropic positions Haiku 5.5 as a cheap worker under an Opus 5.5 or Sonnet 5.5 lead.

Not sure about something? Ask the AI

Questions and answers are published on this page.

Summaries like this, in your inbox every morning.

Context & Analysis

Haiku 4.5 had been Anthropic's small model since October 2025, while OpenAI's GPT-6 Luna, Gemini 3.8 Flash, GLM-5.3 Flash and DeepSeek V4.1 Flash competed at the low-cost end. The Haiku 5.5 launch came with coordinated moves: the same-day Sonnet 5.5 cache-read cut, which Anthropic says makes Sonnet 5.5 about 20% cheaper on most long-running or agentic work, and monthly Claude Platform API credits for Max 5x ($100), Max 20x ($200) and Team (up to $500, pooled) subscribers.

Anthropic positions the model as a subagent paired with Opus 5.5 or Sonnet 5.5, for high-volume, cost-sensitive work such as summaries, compactions and database queries. Partner rollouts on day one included Cursor, GitHub Copilot in VS Code, Devin and Agent Arena. Devin reported 58.4% on FrontierCode 1.1, ahead of Sonnet 5 at roughly one-eighth the cost per task, and recommends it as a "sidekick" under an Opus 5.5 lead in Fusion.

Artificial Analysis's numbers come with caveats. At max effort Haiku 5.5 uses about 162k output tokens per Index task, roughly 3x GPT-6 Luna at max (~50k), so heavy token use offsets part of the per-token discount, and prompts over 100K tokens pay 5x more. That likely explains why Cursor's "10x less than Haiku 4.5" on shorter requests differs from Anthropic's "75% less" on average. Factual recall is weaker than Gemini 3.8 Flash and Luna, and a pre-release safety bug caused over-refusal on AutomationBench-AA, which Anthropic is working to fix.

FAQ
What is new in Claude Haiku 5.5 compared with Haiku 4.5?
It is the first Haiku with Anthropic's effort settings and adaptive thinking, and its context window grew to 1M tokens from 200k. Anthropic says it costs about 75% less to run than Haiku 4.5 on average.
Where can I use Claude Haiku 5.5?
It is live on the Claude Platform and in Claude Code. It was also added on day one to Cursor, GitHub Copilot in VS Code, Devin and Agent Arena.

Also reported by AI Watch (Impress), GIGAZINE AI, Impress Watch, SiliconANGLE AI, Simon Willison's Weblog

AI news that matters for your work, delivered every morning.

Pick your industry and the AI tools you use, and get news related to your work every day.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.

Questions and answers are published on this page.

Related Articles

Next articleLumistar shows TERO Pro AI tennis robot in New York