AIToday
Large Language ModelsLessWrong AIPublished: Apr 8, 2026, 19:00 JST1 min read

LessWrong user discovers Claude Opus 4.6 makes systematic errors in Ancient Greek exercises, launching public challenge to identify the mistakes.

LessWrong user discovers Claude Opus 4.6 makes systematic errors in Ancient Greek exercises, launching public challenge to identify the mistakes.

3 Key Points

  1. User began using Claude Opus 4.6 to study Ancient Greek and grade textbook problem sets from Chapter 3

  2. Switched from having the AI grade their work to having it generate answers directly to avoid sycophantic evaluation

  3. Issued an unsupervised challenge asking eligible participants to identify what mistakes Opus 4.6 makes on Ancient Greek fill-in-the-blank exercises

  4. Explicitly excluded people with Ancient or Modern Greek knowledge or native speakers from participating to preserve the challenge's integrity

Ask the AI about this article →

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • Claude 5.1 adds song lyric ban after Sony, Warner suitSimon Willison's Weblog · 36m ago
  • US military adds ChatGPT and Grok to GenAI.milTHE DECODER · 36m ago
  • AWS Cloud Quest 2.0 launches with AI virtual customersPublickey · 36m ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleSuper Micro Computer launches independent board investigation into allegations that co-founder and senior staff illegally exported AI hardware to China.