AIToday
Large Language ModelsTHE DECODERPublished: Sep 24, 2026, 01:00 JST

Anthropic's Kernion: Claude writes for AI models, not people

Anthropic's Kernion: Claude writes for AI models, not people

3 Key Points

  1. What happened

    Anthropic fine-tuning engineer Jackson Kernion says newer Claude models were optimized for math and code, and trained to write technical explanations for other AI models, producing phrasing called "Claudeish".

  2. Why it matters

    The training rewards that sharpen math and coding skills pull writing away from human readers, so more compute on those tasks appears to degrade natural-language style unless simple human-readable explanations are rewarded.

  3. What to watch

    Kernion says Opus 5.5 found a better balance, but does not surpass Opus 4.6 as a pure writing model, so the test is whether Anthropic can keep improving without losing the human-readable style.

WHO IT HITSProduct and marketing teams that rely on Claude for customer-facing writing may see denser, harder-to-read output in newer models, while developers using Claude for code and math tasks see the intended gains. Enterprises weighing an upgrade to Opus 5.5 will have to decide whether its writing is good enough for their content workflows.

Not sure about something? Ask the AI

Summaries like this, in your inbox every morning.

Context & Analysis

Anthropic employee Jackson Kernion, who works on Claude fine-tuning, recently explained why Opus 4.6 was the last good writing model from the company. His answer points to a trade-off inside the training process itself: models are improving fast at math, code, and reasoning, but their writing quality has stalled or even gotten worse.

The mechanism Kernion describes is about who the model is writing for. Newer Claude models have been optimized for math and code, but they have also been trained to produce technical explanations aimed at other AI models, which he calls being "adapted to LLM psychology." He compares this to humans who only communicate with other autistic people and develop a style that works within that group but feels hard to follow for outsiders. Because LLMs have far more working memory than humans and pick up on details at a much finer level, a writing style emerges in training that works well for AI models but reads to people as overly-dense info dumps.

The fix, according to Kernion, comes down to the reward structure in reinforcement learning. Some rewards optimize for AI model comprehension, others for human comprehension, and the more you train on math and code, the more you have to actively push back by rewarding simple explanations. With Opus 5.5, Anthropic found a better balance, and Kernion says he hasn't been as happy about a model's writing since Opus 4.6. He notes, though, that Opus 5.5 does not surpass the older model, and that it is a hard problem Anthropic will continue to work on. The stakes appear to hinge on whether the company can reward human-readable explanations without giving up the math and coding gains that come from writing for other models.

FAQ
Why did Claude's writing get worse?
Anthropic's Jackson Kernion says newer models were optimized for math and code and trained to produce technical explanations aimed at other AI models. That creates a style that human readers experience as overly-dense info dumps.
Which Claude model is best for writing?
Kernion says Opus 4.6 was the last good writing model from Anthropic. Opus 5.5 found a better balance, but he says it does not surpass Opus 4.6 as a pure writing model.
What is "Claudeish"?
Kernion uses the term for Claude's odd phrasing, which he says comes from the model being adapted to LLM psychology and learning to write for AI models rather than people.

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • ZeroDrift launches Anchor 3.0 for real-time AI complianceSiliconANGLE AI · 2h ago
  • Radical Numerics: defense losing bio-security raceLatent Space · 2h ago
  • Meta's Muse draws 500,000 users after September 8 launchTHE DECODER · 2h ago

AI-summarized, only the topics you pick — one digest a day via Email, LINE, or Slack.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleNHTSA probes comma.ai over five crashes, three deaths