AIToday
Large Language ModelsHacker NewsPublished: Apr 25, 2026, 10:00 JST1 min read

OpenAI's GPT 5.5 tops error-detection benchmark, beating previous AI proofreading records

OpenAI's GPT 5.5 tops error-detection benchmark, beating previous AI proofreading records

3 Key Points

  1. OpenAI released GPT 5.5, which achieved the highest score ever recorded on Errata-Bench—a standardized test that measures how well AI systems catch grammar, style, and factual errors in written text.

  2. The model catches more types of mistakes than prior versions: not just typos and grammar, but also inconsistencies (like a character's age changing mid-document) and unclear phrasing that confuses readers. This means writers using GPT 5.5 for editing get feedback on problems that simpler spell-check tools miss.

  3. For students, journalists, and business writers who rely on AI editing tools, GPT 5.5 becomes a more reliable second pair of eyes—reducing the chance that embarrassing errors slip into final documents. For content teams at publishers and companies, it could reduce the number of manual proofreading passes needed before publication.

Ask the AI about this article →

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • DeepMind chief: frontier AI leadership is all that mattersTHE DECODER · 23m ago
  • John Deere launches AI chatbot for farmersThe Verge AI · 23m ago
  • Google Pics launches with AI image editing for WorkspaceThe Verge AI · 23m ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleGoogle commits up to $40 billion to AI startup Anthropic, escalating competition for AI talent and computing power