AIToday
Large Language ModelsAI Business & IndustryAI Watch (Impress)Published: Sep 29, 2026, 13:00 JST

Anthropic's Claude Sonnet 5.5: 30% faster, clears Pokémon Red

Anthropic's Claude Sonnet 5.5: 30% faster, clears Pokémon Red

3 Key Points

  1. What happened

    Anthropic announced Claude Sonnet 5.5 on September 28. It generates over 30% faster and cuts costs by up to 30% versus Claude Sonnet 5, scoring 70.6% on Terminal-Bench 4.0, up from 10.3%.

  2. Why it matters

    Anthropic is claiming major speed and efficiency gains at the same price per million tokens as Claude Sonnet 5, suggesting the model needs fewer tokens per task and may lower operating costs for teams running it at scale.

  3. What to watch

    The gains rest on efficiency the article ties to fewer tokens per task rather than a price cut, so real-world savings depend on how workloads use the model. Anthropic has also added safeguards for cybersecurity and anti-distillation without affecting general software or biology research.

WHO IT HITSEnterprise AI teams and developers running high-volume workloads on Claude Sonnet 5 stand to see faster responses and lower token costs if the claimed 30% efficiency gains hold, though actual savings will depend on how their specific tasks consume tokens.

Not sure about something? Ask the AI

Questions and answers are published on this page.

Summaries like this, in your inbox every morning.

Context & Analysis

Claude Sonnet 5.5 arrives as part of Anthropic's Claude 5.5 family, positioned as the faster and more efficient option compared with the Claude Sonnet 5 it replaces. The pricing is unchanged at $2 per million input tokens and $10 per million output tokens, but Anthropic attributes the cost improvement to using fewer tokens for the same task rather than to a price reduction.

The performance jump on Terminal-Bench 4.0, from 10.3% to 70.6%, is unusually large for a model refresh, and the article says the model also approached Opus 5.5 on GDPval-AA. Anthropic is presenting the model's long-horizon task ability and image recognition as strengths, pointing to it being the first Sonnet model to clear Pokémon Red using only screenshots. That claim speaks less to typical business workloads and more to the model's ability to act over extended sessions with visual input.

The security additions are worth noting. Anthropic introduced safeguards for high-risk tasks as its cybersecurity capabilities improved, and added safety classifiers aimed at distillation attacks that try to improperly extract a model's capabilities. The article says these restrictions are applied in a way that does not affect general software development or biology research. Whether the efficiency claims translate into real savings will likely hinge on how much customers' own task patterns resemble the token-saving behavior Anthropic describes.

FAQ
How much does Claude Sonnet 5.5 cost?
It is priced at $2 per million input tokens and $10 per million output tokens, the same as Claude Sonnet 5. Anthropic says cost efficiency improves because fewer tokens are needed for the same work.
What is Claude Sonnet 5.5's performance on Terminal-Bench 4.0?
It scored 70.6% on Terminal-Bench 4.0, up from 10.3% for the previous model. It also scored close to Opus 5.5 on GDPval-AA, which evaluates practical work ability across professions.
What safety measures were added to Claude Sonnet 5.5?
Anthropic added safeguards for high-risk tasks alongside improved cybersecurity capabilities. It also implemented safety classifiers to prevent distillation attacks that extract model capabilities, with limits that do not affect general software development or biology research.
AI Watch (Impress)Read Original Article

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • Okta's Blueprint Alliance takes on agent runtime securitySiliconANGLE AI · 10m ago
  • Omdia: 400 security leaders name confusion top AI agent identity blockerSiliconANGLE AI · 10m ago
  • Meta launches Meta Enterprise Platform, taps MongoDB CEO DesaiITmedia AI+ · 10m ago

AI-summarized, only the topics you pick: one digest a day via Email, LINE, or Slack.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.

Questions and answers are published on this page.

Related Articles

Next articleAnthropic's Claude Sonnet 5.5 scores 70.6% on Terminal-Bench 4.0, cuts task cost up to 30%