
What happened
A migration writeup lists seven breakages for code moved to claude-opus-5 or claude-sonnet-5, all in the official migration guide as of October 2026: 400 errors on temperature/top_p/top_k, fixed-budget budget_tokens thinking, and assistant prefill.
Why it matters
Those 400s and silent behavior changes make a one-line model swap unreliable, so teams running existing Claude API integrations are likely to see failed calls or truncated answers until they remove sampling parameters, move to adaptive thinking, and use structured outputs.
What to watch
Whether you hit the seventh item — Sonnet 5's new tokenizer making the same text about 1.3× more tokens — hinges on re-measuring with count_tokens before adjusting max_tokens or cost calculations, since pricing moved to $2/$10 from $3/$15.
WHO IT HITSEngineering and platform teams maintaining existing Claude API integrations, plus anyone responsible for prompt behavior and cost estimates, face code changes rather than a simple model-string swap.
Summaries like this, in your inbox every morning.
The article frames this as a version-history problem rather than a single bad release. Sampling parameters, a fixed thinking budget, and assistant prefill were all accepted in earlier Claude models, and the guide now marks them as removed or discouraged. Correct fixes have also flipped back and forth across 4.7, 4.8, and Opus 5, which is why the author suggests starting by suspecting the sentences you added for the previous model rather than the model itself.
The behaviors that don't raise an error are subtler. Opus 5 keeps responses long even at lower effort, verifies its own answers, and delegates to subagents more often, so leftover instructions written for 4.8 — asking for more delegation or for final verification — can push it into over-verification. Sonnet 5 reads instructions literally, which makes carried-over style notes like "be concise" bite harder, and a "report only critical issues" instruction can lower code-review recall. The author also notes that moving to the top-tier Claude Fable 5.1 adds three more constraints around disabled thinking, tool_choice, and edited conversation history.
Taken together, the practical stakes for teams already running Claude in production lie in prompt hygiene and measurement discipline. Whether a migration goes smoothly appears to hinge on auditing inherited prompts and re-checking token counts with count_tokens before trusting old budget assumptions, since the pricing change and the tokenizer change pull costs in opposite directions.
Pick your industry and the AI tools you use, and get news related to your work every day.
Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.
NetApp and Iterate.ai are packaging the AIPod Mini with Iterate's Generate platform and an embedded LLM, so en…
Mercor had 12 licensed CPAs work through simplified APEX Accounting Benchmark tasks

Testing Azure API Management's llm-token-limit policy at 800 tokens per hour, actual consumption hit 1,472 tok…

Qwen released Qwen3.8-Flash-Next on August 27, 2026, calling it a preview of the architecture planned for Qwen…

A Zenn article floated a hackathon where participants get the theme on the day, use no PC, internet, smartphon…

Anthropic's Message Batches API offers a 50% off rate, takes up to 10,000 requests per batch, and returns resu…
