AIToday
AI Safety & AlignmentAI Business & IndustryFortune AIPublished: Oct 10, 2026, 04:00 JST

Fired OpenAI safety staff deny leaking, defend Astra monitoring

Fired OpenAI safety staff deny leaking, defend Astra monitoring

3 Key Points

  1. What happened

    Tomek Korbak, Mikita Balesni and Jasmine Wang, fired Oct. 1, published a four-page letter on Oct. 8 denying they leaked Astra monitoring concerns to The Information or mishandled METR talks. OpenAI said it found a significant breach of trust.

  2. Why it matters

    The dispute pits former staff against OpenAI over whether raising safety concerns or working with outside auditors led to dismissal, a claim OpenAI denies. The letter warns firings may be used to justify ending work with METR.

WHO IT HITSOpenAI's current safety and research staff face uncertainty over how internal dissent and outside-auditor contact are treated. Third-party safety assessors like METR may see their access and scope narrowed as OpenAI finalizes new contracts.

Not sure about something? Ask the AI

Questions and answers are published on this page.

Summaries like this, in your inbox every morning.

Context & Analysis

The dispute centers on two sensitive areas: OpenAI's ability to monitor its advanced models, and its engagement with outside safety organizations such as METR. Korbak served as technical point of contact for METR, the third-party research firm OpenAI asked to conduct an independent audit of the Hugging Face incident. Balesni and Korbak worked on OpenAI's investigation of that incident.

The letter also responds to internal versions of events. The former employees say ongoing internal and external communications are making current employees afraid to speak up, and they deny leaking concerns about monitoring OpenAI's latest Astra model to The Information, which published a Sept. 1 article about it.

OpenAI has not disclosed specific reasons for the firings and did not link them to METR communications or the Hugging Face incident. The company told Fortune it found a pattern of misconduct, including violations of its information-handling policies. Korbak said OpenAI told him verbally he was fired over how he communicated with METR. Wang said she was told she accessed an unnamed executive's email, which she says the company had given her for recruiting and that IT did not remove when she asked. Balesni said the researchers were fired for prioritizing safety over OpenAI's near-term corporate interests.

FAQ
Why were the OpenAI safety team members fired?
OpenAI said they mishandled sensitive information outside established procedures and that an internal investigation found a significant breach of trust. The former employees deny leaking Astra monitoring concerns to The Information and deny any foul play in their METR communications.
What did the fired employees say about METR?
Korbak said he was told verbally he was fired because of how he communicated with METR. The letter warns the firings may be used to justify ending OpenAI's work with METR or limiting external auditors' access.
Did OpenAI say anything about future third-party safety audits?
OpenAI said it remains committed to working with third-party safety assessors and is finalizing contracts it will announce in the coming weeks.

AI news that matters for your work, delivered every morning.

Pick your industry and the AI tools you use, and get news related to your work every day.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.

Questions and answers are published on this page.

Related Articles

Next articleRillet's 2-person team ships 3× faster with eve agents on Vercel