AIToday
WIRED AIPublished: May 26, 2026, 22:00 JST1 min read

AI Search and Chatbots Are Wrong About Half the Time, According to Fact-Checker Research

AI Search and Chatbots Are Wrong About Half the Time, According to Fact-Checker Research

3 Key Points

  1. A fact-checker at WIRED tested AI models (ChatGPT, Claude, Gemini, and Grok) on a professional fact-checking task. None of the models actually performed the fact-checking work; they all provided plans but stopped short of executing them.

  2. Recent benchmarks show stark accuracy gaps: Claude led RealFactBench with 73 percent accuracy, but OpenAI's SimpleQA found that none of the tested models exceeded 50 percent accuracy. A March 2025 Tow Center study found more than 60 percent of AI-powered search engine responses were inaccurate, while a BBC study puts chatbot error rates closer to 45 percent.

  3. WIRED's fact-checking desk encounters AI Overviews in Google search as their main interaction with AI for verification work, and the author assesses these as wrong about a third of the time. A 2025 Association for the Advancement of Artificial Intelligence report found that 60 percent of surveyed researchers doubted the 'factuality' problem would be solved anytime soon.

Ask the AI about this article →

Get AI news like this every morning

For example, today's edition would include:

  • Aranya raises $11M to turn bare-metal servers into AI clusters in 48 hoursSiliconANGLE AI · 2h ago
  • CBTS launches Forge Agents for custom AI agentsSiliconANGLE AI · 2h ago
  • Phonely launches Alma, voice AI trained on 10M callsSiliconANGLE AI · 2h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Next articleAnthropic seeks to block Pentagon use of its AI for autonomous weapons, as AI warfare moves from theory to practice