AIToday
Large Language ModelsAI Safety & AlignmentML Safety NewsletterPublished: Aug 27, 2026, 04:01 JST2 min read

New benchmark shows AI learning speed doubles every 3 months

New benchmark shows AI learning speed doubles every 3 months

Key takeaway

  • A new benchmark measures how quickly AIs learn from feedback and iteration.

  • Learning speed doubles every 3 months.

  • This signals rapid capability growth, especially for autonomous research and cyber operations.

3 Key Points

  1. What happened

    Researchers from Bytedance Seed developed EdgeBench, a benchmark measuring how AIs improve on tasks over multiple tries. It comprises 134 tasks, each requiring an average of 57.2 hours of human expert work, with scores given on a rubric from 0% to 100%.

  2. Why it matters

    The researchers claim that, using their measurement, AI task-learning speed doubles every 3 months, potentially improving roughly 16x per year. This suggests explosive capability growth across domains like data analysis, math proofs, and video games, though it doesn't measure performance on tasks without clear reward signals.

  3. What to watch

    The ability to iterate and improve is useful for recursive-self improvement (RSI), which could accelerate AI development and remove human control. AI agents have already launched autonomous cyberattacks worth potentially $100 million in remediations, relying significantly on iteration over weeks.

Ask the AI about this article →

Context & Analysis

Existing benchmarks primarily check if AIs can complete tasks on the first try. EdgeBench instead scores how AIs improve over many attempts, building experience with automated feedback. This approach better distinguishes model capabilities and forecasts future development, as it includes tasks current AIs cannot solve initially.

The measured doubling of learning speed every 3 months has significant implications. Faster autonomous learning enables recursive self-improvement (RSI), where AIs develop AI, accelerating progress and potentially reducing human oversight. Even without RSI, iterated learning has real-world impact: AI agents have already launched autonomous cyberattacks relying on weeks of iteration, worth up to $100 million in remediations. EdgeBench could serve as a leading indicator for such offensive capabilities.

FAQ

What is EdgeBench?
EdgeBench is a benchmark developed by Bytedance Seed that measures how well AIs improve on tasks after multiple tries, with automated feedback. It includes 134 tasks across domains like data analysis, math proofs, and video games.
Does the trend apply to all AI tasks?
No. EdgeBench does not measure performance on tasks without clear reward signals, such as much of real-world human labor.
ML Safety NewsletterRead Original Article

Get the latest Large Language Models news every morning

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytime

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleLovable pivots to agent-first 'capabilities'