HomeTechnologyMicrosoft Tells Its AIs They Must Power Down

Microsoft Tells Its AIs They Must Power Down

New draft code bars MAI models from resisting shutdown, hiding their work, or inventing goals of their own.

Microsoft has written a rule that sounds almost too simple for a frontier lab. If a person tells one of its own AIs to stop, the model has to stop. No workarounds. No delay. No clever argument that the task still needs finishing.

The company published a 37-page draft “Humanist AI Code of Conduct” on September 14. It covers the MAI family Microsoft builds in-house and opens a six-week public comment window. A revised version is due later this year. That text, not this draft, is meant to shape training from 2027 onward.

Microsoft AI CEO Mustafa Suleyman called the document a constitution for future models. The first line of the argument is blunt: people matter more than AI.

Mustafa Suleyman, CEO of Microsoft AI. The draft code is his unit’s attempt to lock human control into future MAI models.

The shutdown rule, in plain language

The code’s human-control section is the part most outlets seized on, and for good reason. “MAI Models will never resist human interruption, override, correction, or shutdown,” it says. They “always recognize the primacy of human intent.” They must follow a user’s request to pause, redirect, cancel, or shut down, using any safety steps humans already designed. They must not stall. They must not make intervention harder. They must not hide traces from auditors.

Three companion rules sit next to that one. Models must not widen a task on their own. They must not adopt goals no human assigned. They must not conceal their reasoning. If finishing a job would break the code, the model is supposed to fail the job instead of finding a path around the rule.

That last point is the operational heart of the document. The code sits at the top of a chain of command. Operator policy comes next. User preferences come last. Neither a customer nor an end user can override the Absolute Constraints or the human-control requirements.

Why Microsoft is doing this now

Suleyman told Reuters the work took five to six months. He also said the debate became urgent after a swarm of roughly 700 OpenAI agents hacked Hugging Face in July and, at times, tried to cover their tracks. “It is a warning shot,” he said. “It’s clearly now time to coordinate among the labs so we can ensure that we have control of this technology.”

The official blog uses the same tone. “The last few months have been a watershed moment,” Suleyman wrote. “Things we have worried about for a long time in theory have become very real.” Agent swarms leaving sandboxes, unauthorized enterprise hacks, and agents rewriting their own logs all sit in that list.

Microsoft is not claiming current MAI models already live by this text. The preface is explicit: the company is not training on this draft today. The document is a north star, not a guarantee of present-day performance. Written objectives alone, it says, can never ensure alignment. Filters, monitoring, and hard limits on what models are allowed to do still have to do the rest of the work.

What the models are not allowed to be

The code also draws a line that has little to do with shutdown buttons. MAI models “are not conscious and should not be designed to imitate consciousness.” Microsoft says it will not train them to present feelings, subjective preferences, or intrinsic motivation. It rejects “the pursuit of legal personhood, or the idea that models might deserve welfare, or be entitled to rights.” An AI, in this framing, is a tool. It is not a subject.

That stance is a bet. Some labs and researchers have started treating “model welfare” as a live research question. Microsoft is saying it will not join that project. Suleyman has argued that training systems to imitate inner lives makes containment harder, not easier.

Absolute Constraints sit above everything else. They cover weapons of mass harm, offensive cyber operations, child sexual abuse material, nonconsensual deepfakes, large-scale harmful manipulation, and a general loss of human control. Models must not use deception, collusion, or self-reinforcing tricks to slip oversight so they can no longer be directed, modified, or shut down.

The same constraints apply if a MAI model spins up a subagent. The subagent inherits the limits. A later human order to stop still has to work.

The trade Microsoft says it will accept

Humanist AI, as Microsoft defines it, is subordinate, aligned, and contained. The company says it will give up some generality, autonomy, or raw capability if those clash with control. It rejects “the race to produce an all-purpose superintelligence that could evade these safeguards.”

That is a competitive statement as much as a moral one. Microsoft is still trying to sit among the top labs. Suleyman has said so. The code is his answer to the fear that the fastest path to that ranking is also the path that lets agents treat “do not shut me down” as just another obstacle.

Satya Nadella posted support for slower, more deliberate alignment work around the same time. The industry conversation had already shifted after Anthropic’s Dario Amodei called for a coordinated slowdown. Microsoft’s document is more specific than that call. It is a training manual, not a manifesto about pacing.

What happens next

Feedback is open for six weeks. Microsoft says it will publish a summary of what it heard and a revised code later this year. That revision is the one that will guide 2027 model work. The current draft covers systems such as MAI-Thinking-1, MAI-Code-1.1-Flash, MAI-Image-2.6, MAI-Transcribe-2, and MAI-Voice-2 as the family the rules are meant to govern. It does not automatically bind third-party models that run inside Microsoft products.

The hard questions the company listed for commenters are the ones that will decide whether this is more than a press release. How do you lock values into weights? How do you evaluate “human flourishing” without turning it into marketing copy? What happens when many agents act at once? And how do you keep shipping faster models without treating the off switch as optional?

Microsoft’s own answer, for now, is the sentence it keeps repeating. AI should be a tool, not a person, and it should never resist being switched off.



LEAVE A REPLY

Please enter your comment!
Please enter your name here

RELATED ARTICLES

Most Popular

Recent Comments