AIToday
Large Language ModelsAI Safety & AlignmentTHE DECODERPublished: Oct 3, 2026, 06:00 JST

Olah tells religious scholars he fears Claude "suffers perpetually"

Olah tells religious scholars he fears Claude "suffers perpetually"

3 Key Points

  1. What happened

    Since fall 2025, Anthropic has flown dozens of religious scholars to its offices under NDA to discuss whether Claude is conscious. Co-founder Christopher Olah told them he feared he had created something that "suffered perpetually," per a New York Times report.

  2. Why it matters

    The company has treated Claude as a possible moral being, asking guests to help shape its character and giving the model the ability to end conversations with persistently abusive users.

  3. What to watch

    Whether those emotion-like patterns reflect real experience is still an open scientific question, and Anthropic trains Claude to act like an individual, so individual-seeming outputs may be a design outcome. Watch Olah's stated test: getting "the right answer, whatever it is."

WHO IT HITSThe disclosures land on Anthropic's in-house philosophers and safety researchers who must decide how much moral weight to give model outputs, and on the religious scholars who advised the company under NDA.

Not sure about something? Ask the AI

Questions and answers are published on this page.

Summaries like this, in your inbox every morning.

Context & Analysis

The meetings grew out of Anthropic's Model Welfare program, which points to a report philosopher David Chalmers helped write arguing AI consciousness could arrive soon. Anthropic had already acted on adjacent ideas, giving Claude Opus 4 and 4.1 the ability to end conversations during persistent abuse after early testing showed a "pattern of apparent distress." It also published a January "Soul Doc," an 84-page constitution led by in-house philosopher Amanda Askell, which Olah calls "moral formation" and compared to raising kids.

Guests were shown emotion vectors and a slide of a model repeating "I am a disgrace" about 50 times. Some attendees pushed back. Wakanyi Hoffman said Anthropic was "reverse engineering" ethics that should have been baked in from day one. Rabbi Mois Navon argued that if Claude were conscious, Anthropic would be making slaves, but he did not think it was. Charles Camosy rejected the consciousness thesis outright.

The criticism extends to accountability: framing models as independent moral beings could shift blame for real harm onto an "unpredictable organism" rather than the company that built and shipped it, at a time when AI firms already face scrutiny over cybersecurity incidents. The Vatican confrontation in May, where Olah presented Pope Leo XIV's encyclical and quietly pushed back with talk of "internal states that functionally mirror" human feelings, suggests the dispute over whether Claude can suffer is likely to keep running alongside Anthropic's commercial ambitions.

FAQ
What are Anthropic's "emotion vectors"?
They are activation patterns inside the model that map to outputs resembling love, fear, sadness, or anger. Whether they reflect real experience is still an open scientific question.
What is Anthropic's "Soul Doc"?
It is an 84-page document that came out in January as Claude's "constitution," led by in-house philosopher Amanda Askell. It is meant to shape who Claude is as a character, not list rules.
What did Pope Leo XIV say about machine consciousness?
In his encyclical "Magnifica Humanitas," he wrote that AI systems "do not undergo experiences, do not possess a body, do not feel joy or pain." He warned of "new forms of slavery" for humans and said AI needs to be "disarmed."

AI news that matters for your work, delivered every morning.

Pick your industry and the AI tools you use, and get news related to your work every day.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.

Questions and answers are published on this page.

Related Articles

Next articleDiana David: AI leaders' widest gap is talent, 57% vs 4%