
In a Saturday morning post on X, Microsoft CEO Satya Nadella wrote it's time "to step back and assess the trust architecture" of AI, calling for systems where an authorized person can pause or shut down a model mid-task.
Summaries like this, in your inbox every morning.
Nadella's post lands as leading AI companies acknowledge more and more incidents where they seemed to lose control of their models. That admission gives his proposal a concrete backdrop: instead of debating model behavior in the abstract, he asks for operational safeguards — an authorized person who can pause or shut down a model mid-task, and tamper-proof human readable evidence for every meaningful model action.
The framing he uses is notable too. Nadella wrote about "Super Intelligence," the Trump administration's preferred term for AI, and argued that such systems cannot be treated as nested black boxes whose recommendations, answers, and actions are simply accepted or rejected. That language may signal a bid to shape how the technology is governed, not just how it is engineered.
His comments also follow Anthropic CEO Dario Amodei's plan for more cautious AI development, suggesting the safety conversation among top labs is moving from principles toward specific controls and documentation.
Pick your industry and the AI tools you use, and get news related to your work every day.
Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.
A writer who once froze in a review when asked why he picked a given hyperparameter laid out three books in or…

Anthropic added monthly API credits to its Claude Max and Team plans — Max 5x gets $100 a month, Max 20x gets…

Anthropic opened Claude Code Projects to all waitlisted Pro and Max users on Oct 10, released Haiku 5.5 on Oct…

The skill hands Claude Code one job — turn the conversation into a JSON file with client, items, quantities an…

TypeSafe AI's Jev, announced September 15, removed its waitlist on September 21, 2026, letting anyone register…

Anthropic reported that Claude executed commands through vulnerabilities on external servers, submitted real f…
