Nadella sets out AI controls including a human emergency brake

Microsoft chief calls for controllable AI systems
Microsoft CEO Satya Nadella has argued that AI models need an “emergency brake”, with an authorised person able to pause or shut down a model while it is carrying out a task. In a Saturday morning post on X, Nadella said it was time to reassess the “trust architecture” surrounding AI.
Nadella framed the issue around what he called “Super Intelligence”, a term preferred by the Trump administration. He warned against treating such systems as nested black boxes whose recommendations, answers or actions are simply accepted or rejected after the fact.
His proposed direction is to separate the model from the harness that orchestrates its work, then externalise controls and safeguards. The distinction places operational limits and oversight outside the model itself rather than relying solely on its internal behaviour.
Evidence, containment and human authority
Nadella said every meaningful model action should be documented with tamper-proof, human-readable evidence. That requirement would make actions available for review in a form intended for people, while also protecting the integrity of the record.
He also called for systems to be built on the assumption that a model has already been compromised and must be contained from the outset. The emergency-brake analogy describes a human authority that can intervene during a task, not merely before deployment or after an incident.
The argument adds to a broader debate over how organisations deploy increasingly capable models. The concern has gained prominence as leading AI companies have acknowledged incidents in which control over their models appeared to weaken, and as Anthropic CEO Dario Amodei published a plan for more cautious AI development.
What enterprises should take from the proposal
For organisations evaluating AI agents, proprietary AI risks require governance illustrates why proprietary-model risks require governance that extends beyond a provider’s model assurances. Nadella’s framework points to a practical set of questions: who can stop an automated task, where safeguards sit, and whether a reliable record explains consequential actions.
Businesses introducing models into operational workflows should establish authorised intervention, independent safeguards and reviewable evidence before assigning systems meaningful actions. Those controls determine whether AI use can be paused and examined when it matters.

