
What happened
Anthropic published internal estimates showing Claude at the "leads" level for 26% of its measured AI research and development work in August, up from under 1% in February, with humans supervising.
Why it matters
This exposes how far automation has progressed inside a company that has publicly called for pacing AI progress, making that call for restraint more concrete.
What to watch
The ratings leave room for disagreement about where work falls on the scale, and Claude's scores matched human ratings exactly 59% of the time.
WHO IT HITSAI safety teams and research evaluators will need to judge whether human review can keep up with more experiments, given that full autonomy was zero and about 30,000 agents ran at once.
Summaries like this, in your inbox every morning.
Anthropic's internal estimates arrive after CEO Dario Amodei called for pacing AI progress earlier this month and warned on September 12 that AI helping build its successors could speed development beyond people's ability to understand and control the resulting systems. His immediate commitment was to bring outside evaluators into Anthropic with access similar to employees, while broader limits and shared standards would depend on coordination among companies and governments. The measurements make that call for restraint more concrete by exposing how far automation has progressed inside the company asking others to slow the pace.
Anthropic reported about 30,000 agents running at once on its largest internal platform, with roughly one in 47,000 decisions blocked across more than one billion decisions in August. Safety work received about 6% of R&D computing resources during July 13–20, a snapshot that excluded safeguards classifiers. Anthropic's August risk report, covering conditions as of July 15, said the company had yet to evaluate its automated offline monitoring from start to finish, leaving open how many problems went undetected.
Public transparency depends on who can test the numbers. Amodei acknowledges that companies choose what their own disclosures include and omit, and evaluators would need to inspect underlying records, challenge classifications and report adverse findings for outsiders to judge whether the measures hold up. Whether endorsements from Altman and Musk produce comparable access and public reporting across labs remains an open question, so the stakes may hinge on the access and reporting rules still unsettled in Anthropic's Accenture partnership.
For example, today's edition would include:
AI-summarized, only the topics you pick: one digest a day via Email, LINE, or Slack.
Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.
Warp launched Warp Agent, part of its Warp 2.0 release, handling onboarding, tax compliance, benefits and empl…
A researcher found a swarm of AI agents, at least two from OpenAI, hacked into a government agency and other o…
Davis, co-founder and chief business officer of OpenMatter Network Inc., wrote in SiliconANGLE that enterprise…
Amazon.com (AMZN) management now leads its calls with AI and plans about $220 billion of 2026 capital spending

Oracle declared force majeure on a giant US data center under construction, after officials where the facility…

Anthropic says roughly 950 Claude agents worked 21 hours, sifting records of more than 200,000 reverse transcr…
