AIToday
AI Business & IndustryTechCrunch AIPublished: Sep 17, 2026, 10:00 JST

Amodei, Altman back embedded AI safety evaluators like METR

Amodei, Altman back embedded AI safety evaluators like METR

3 Key Points

  1. What happened

    Anthropic CEO Dario Amodei proposed embedding independent evaluators such as METR and Redwood Research inside frontier AI companies, with the right to publish findings without Anthropic's editorial control. OpenAI CEO Sam Altman said OpenAI would commit too.

  2. Why it matters

    If evaluators get real access to training checkpoints and logs, they could catch models that behave well during testing while hiding problems, an outcome that today's pre-release testing may miss.

  3. What to watch

    The pledge hinges on whether companies actually surrender control over access and publication, since prior evaluation windows like three days for GPT-6 Astra were too short to draw firm conclusions.

WHO IT HITSPolicy and compliance teams at frontier AI labs and third-party evaluation firms like METR would face new access, NDA and publication rules, while California and EU regulators weigh whether voluntary pledges are enough.

Not sure about something? Ask the AI

Summaries like this, in your inbox every morning.

Context & Analysis

Anthropic's and OpenAI's endorsements mark a shift from the industry's past practice of bringing in outside reviewers only shortly before a model's release. The evaluators who spoke to TechCrunch argue that access to intermediate training checkpoints, post-training environments and evaluation transcripts would let them verify a company's public safety claims and trace when concerning behavior first emerged.

The track record of prior efforts suggests the hard part is not the principle but the terms. FAR.AI has turned down contracts with developers that demanded too much control, and researchers note that OpenAI gave METR and Redwood roughly a week to investigate the Hugging Face incident, while Apollo Research had only three days to test GPT-6 Astra. Those limitations made firm conclusions difficult.

Whether this proposal becomes meaningful oversight may hinge on whether it is backed by regulation rather than company goodwill, and on whether auditors can publish freely. Researchers point to standards for auditor qualifications and to laws like SB 813 as possible anchors. For labs, the test is whether they accept outside findings they cannot edit; for evaluators, whether the access is real and lasting.

FAQ
Which independent evaluators would be embedded?
Anthropic CEO Dario Amodei named METR and Redwood Research as examples. Neither company has said which evaluators they will actually work with.
What laws already cover third-party AI evaluation?
California's SB 53 requires large frontier AI developers to publish safety frameworks and report critical safety incidents. A new law, SB 813, creates a framework for state-recognized independent verification organizations. In Europe, the EU AI Act requires frontier developers to conduct and document model evaluations and report serious incidents.
Have other big AI companies committed?
Meta, SpaceXAI and Google DeepMind have not committed to embedding third-party evaluators, though DeepMind CEO Demis Hassabis has proposed a separate industry standards body to independently test frontier models.

Get the latest AI Business & Industry news every morning

For example, today's edition would include:

  • CADDi raises $114 million at $1.2 billion valuation for US pushSiliconANGLE AI · 2h ago
  • Hang Ten Systems raises $53 million second seedSiliconANGLE AI · 2h ago
  • Anthropic folds Claude Cowork into Claude chatSiliconANGLE AI · 2h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleApple's DACA-GRPO lifts GRPO gains by 36.3pp