AIToday
Large Language ModelsAI Safety & AlignmentSemafor TechPublished: Sep 10, 2026, 22:01 JST2 min read

Christiano backs extinction warning, joins OpenAI safety team

Christiano backs extinction warning, joins OpenAI safety team

3 Key Points

  1. What happened

    Paul Christiano, co-author of an influential 2016 AI safety paper with future Anthropic founder Dario Amodei, said he now sees "meaningful risk" of "catastrophic and irreversible loss of control in the very near term," warning "Most people could die," and announced he was joining OpenAI's nonprofit safety team.

  2. Why it matters

    Christiano had long believed AI would take off relatively gradually and be made safe. His shift echoes Geoffrey Hinton, one of the three "Godfathers of AI," who told the BBC a 10% risk of human extinction "seems a not unreasonable estimate."

  3. What to watch

    The test is whether these warnings from insiders change how AI labs prioritize safety work, with Christiano's move to OpenAI's nonprofit safety team as one concrete signal. Watch whether the 10% figure and Anthropic alignment scientists' comments gain further traction.

WHO IT HITSAI safety researchers, lab leadership, and policymakers weighing extinction-risk arguments now have prominent insiders — Christiano and Hinton — publicly backing stark warnings.

Ask the AI about this article →

Summaries like this, in your inbox every morning.

Context & Analysis

The warning comes from inside the field, not from outside critics. Paul Christiano co-authored an influential 2016 AI safety paper with Dario Amodei, who later founded Anthropic. Christiano had long been known as a believer that AI would take off relatively gradually and be made safe — a position this shift reverses. His announcement that he is joining OpenAI's nonprofit safety team is the concrete step attached to that reversal.

His warning did not land alone. Geoffrey Hinton, one of the three so-called "Godfathers of AI," told the BBC that a 10% risk of human extinction "seems a not unreasonable estimate," a comment that echoed what one of Anthropic's top AI alignment scientists had said a day earlier. Read together, the two statements suggest the most dire framing is no longer confined to a fringe, but is being voiced by people with deep ties to the field's leading labs.

What happens next likely hinges on whether these warnings translate into changes in how safety work is prioritized inside those labs, with Christiano's move to OpenAI's nonprofit safety team as one early indicator. The 10% figure and the comments from Anthropic's alignment scientists may also shape how the wider public and policymakers weigh the risk.

FAQ
Who is Paul Christiano and what did he warn about?
He is the author of an influential 2016 paper on AI safety with future Anthropic founder Dario Amodei. He said he now sees "meaningful risk" of "catastrophic and irreversible loss of control in the very near term" and warned "Most people could die."
Where is Paul Christiano going and why?
He announced he was joining OpenAI's nonprofit safety team to reduce the risks.
What did Geoffrey Hinton say about extinction risk?
Hinton, one of the three so-called "Godfathers of AI," told the BBC that a 10% risk of human extinction "seems a not unreasonable estimate," echoing a comment from one of Anthropic's top AI alignment scientists.

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • DeepSeek V4.1-Flash: 763B model beats V4 Pro on AA Index 40Latent Space · 39m ago
  • Dynatrace acquires Arize AI as observability shifts to actionSiliconANGLE AI · 6h ago
  • Shared base cuts 100 fine-tunes from 1.5 TB to 19.3 GBDaily Dose of Data Science · 6h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next article29 companies join Unicorn Board, adding $63 billion