Claude Opus 4.7 achieves the highest Elo score of 1753 on GDPVal-AA knowledge work evaluation, surpassing GPT-5.4 (1674) and Gemini 3.1 Pro (1314).
The lead is minimal—Opus 4.7 only outperforms GPT-5.4 by 7-4 on directly comparable benchmarks, indicating an increasingly competitive LLM market.
Anthropic is keeping its more powerful successor model, Mythos, restricted to enterprise partners for cybersecurity testing and vulnerability patching.
Opus 4.7 excels in agentic coding, scaled tool-use, agentic computer use, and financial analysis compared to GPT-5.4 (released March 2026) and Gemini 3.1 Pro (February 2026).
Ask the AI about this article →
For example, today's edition would include:
AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.
Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. Q&As are published on this page for other readers too.
Israeli startup DataAgent Ltd
Taoyuan is positioning itself as a northern hub for AI data centers (AIDC), citing the Tatan area and an LNG c…

SK Hynix presented a custom HBM concept at SEMICON Taiwan 2026, where compute functions are placed in the base…

The U.S. Department of Defense announced on August 31 that it has deployed ChatGPT Mil, a customized version o…

Nvidia reported earnings that were both remarkable and boring, reflecting its focus on avoiding a consolidated…

Anthropic has agreed to a $35bn cloud-computing contract with Lambda, a Nvidia-backed cloud provider
