TopicAI incident reporting and accountability
Nadella Calls for an 'Emergency Brake' on AI Models: Assume Compromise, Let Humans Hit Stop
Microsoft CEO Satya Nadella posted a lengthy essay on X on Saturday, Oct 10, calling for a reassessment of AI's "trust architecture": separate the model from the harness that orchestrates it, externalize controls and safeguards, document every meaningful action with tamper-proof, human-readable evidence, and let an authorized person always pause or shut down a model mid-task.

Related
What happened
According to TechCrunch's Oct 10 report, Nadella wrote in the post that it is time to "step back and assess the trust architecture" of AI. He argued that "Super Intelligence" — the Trump administration's preferred term for AI — cannot be treated as a set of nested black boxes whose recommendations, answers and actions are simply accepted or rejected.
We must assume a model is compromised and contain it from the start. Think of it like an emergency brake.
Translation:我们必须假定模型已被攻陷,并从设计之初就加以隔离。可以把它想象成紧急刹车。
View on X →Nadella's specific proposals include: separating the model from the harness that orchestrates its work; externalizing controls and safeguards outside the model; documenting "every meaningful model action" with "tamper-proof, human-readable evidence"; and ensuring an authorized person can always "pause or shut down a model mid-task." TechCrunch noted the comments come as leading AI companies acknowledge a growing number of incidents in which they seemed to lose control of their models, and after Anthropic CEO Dario Amodei published a plan for more cautious AI development.
Key facts
- When: Saturday, Oct 10, in a lengthy post on X; first reported by TechCrunch at 2:47 PM PDT the same day
- Core proposal: Assume a model is compromised and contain it from the start; an authorized person must always be able to pause or shut down a model mid-task
- Architecture: Separate model from harness; externalize controls and safeguards; tamper-proof, human-readable evidence for every meaningful action
- Wording: Uses "Super Intelligence" throughout — the Trump administration's official term for AI
- Context: Follows a string of acknowledged model-control incidents at AI labs, and Anthropic CEO Dario Amodei's recent plan for more cautious AI development
Why it matters
the Microsoft chief has for the first time publicly framed AI safety around zero trust — assume compromise by default — and demanded that a human-controlled emergency brake become standard.
CompaniesMicrosoft
Comments
Today
Oct 12 Monday- BriefAmazon in Talks to Buy AI Startup Decart in Deal Valued Around $7 Billion
10 stories · Oct 11
- Quiz
- Call
Latest news
All →- Oct 11Amazon in Talks to Buy AI Startup Decart in Deal Valued Around $7 Billion
- Oct 11GSK Expands Chai Deal After Wet-Lab Validation of AI Designs
- Oct 11PPT Master Hits GitHub Trending: Documents Become Native PowerPoint
- Oct 11context-mode Hits GitHub Trending: Tool Output, Sandboxed First
- Oct 11Cloudflare Acquires Deno: Deploy Shuts Down in Six Months
- Oct 11Anthropic Updates Claude Usage Policy: Armed Drones Named and Banned, Effective Nov 12
