AIToday
Large Language ModelsAI Business & IndustryLessWrong AIPublished: Sep 28, 2026, 13:00 JST

AI crosses METR limit as tasks hit months of human work

AI crosses METR limit as tasks hit months of human work

A LessWrong essay says METR's time-horizon graph went from seven-minute tasks at 80% reliability in early 2025 to over an hour by year-end, and Astra and Fable 5.1 now exceed what METR's task suite can measure.

Not sure about something? Ask the AI

Questions and answers are published on this page.

Summaries like this, in your inbox every morning.

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • Autoheal raises $7.9 million for self-fixing AI agentsSiliconANGLE AI · 1h ago
  • Paul Cheek: 30% of S&P 500 execs AI-literate, 78% gapFortune AI · 1h ago
  • Agent cost per successful task: a Zenn design-variable argumentZenn AI/ML · 1h ago

AI-summarized, only the topics you pick: one digest a day via Email, LINE, or Slack.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.

Questions and answers are published on this page.

Related Articles

Next articleOpenAI agents linked to 16,000 UNCTADstat scans