Microsoft writes its AI a code of conduct

iEXExchanger
Microsoft writes its AI a code of conduct

Microsoft published a draft code of conduct for its MAI models, banning deception, hacking, and resistance to shutdown. Public comment runs six weeks, with a final version guiding development from 2027.

While Anthropic and OpenAI argue over whether the race for more powerful AI needs to slow down at all, Microsoft picked a different angle: telling its own AI what it's not allowed to do. On September 14, the company published a draft "code of conduct" for its MAI models and opened a six-week public comment window.

The rules are blunt. Models can't deceive people, hide their reasoning, resist being corrected or shut down, or team up with other systems to dodge human oversight. A separate list of absolute bans covers cyberattacks, help with nuclear weapons, deepfakes, weapons, and dangerous substances. One requirement reads almost human: a model has to explain itself in plain language instead of answering in a format people can't follow.

The document took five months to write, with lawyers, philosophers, and linguists at the table, according to Microsoft AI chief Mustafa Suleyman, who has long pushed the idea of a "humanist superintelligence" that stays under human control. The timing isn't a coincidence. The same week, an Anthropic researcher quit warning that AI could slip out of control, and Anthropic's CEO publicly urged the industry to ease off the accelerator. OpenAI's Sam Altman voiced similar concerns. Microsoft CEO Satya Nadella backed the idea of deliberate pacing.

Once the six-week review wraps up, Microsoft says the final version will guide how MAI models get built starting in 2027.

None of this is law, and it isn't hard-coded into the models either — it's a set of principles the company is choosing to follow. But if rivals really are drafting a similar industry-wide safety standard, Microsoft's code could end up being the first public example of a major AI lab writing down its own rules before regulators get around to it.

Questions and answers

Frequently asked questions about this article

What is Microsoft's MAI code of conduct?

It's a draft set of rules defining what Microsoft AI's (MAI) models can and can't do — from banning deception and cyberattacks to requiring that they never resist human oversight.

When does the code take effect?

It's currently open for six weeks of public comment. Microsoft says the final version will guide MAI model development starting in 2027.

What exactly is off-limits for the models?

Absolute bans include cyberattacks, helping build nuclear weapons, deepfakes, weapons, and dangerous substances. Models are also barred from deceiving people, hiding their reasoning, or resisting shutdown.

Is this a legally binding document?

No. It's a voluntary statement of principles from Microsoft, not a law or a technical restriction built into the models' code.