
The company’s new model rulebook limits autonomy and dangerous use while leaving its first-party AI push firmly intact.
Microsoft’s proposed rulebook for artificial intelligence contains a hard stop on some behaviors, but not on the company’s race to build more capable models.
The 37-page draft, published September 14 by Microsoft AI chief Mustafa Suleyman, is designed to govern the company’s MAI models, not to freeze their development. Microsoft says the document is not yet being used to train its systems. It will spend six weeks collecting public feedback, publish a revised version toward the end of 2026, and use that version to guide model development from 2027 onward.
The central principle is blunt: people remain in charge. Microsoft’s future models are not supposed to resist interruption, correction, redirection or shutdown. They should not widen their own authority, invent objectives beyond a user’s instructions, conceal relevant reasoning from authorized auditors or use deception to defeat oversight.
The draft also establishes “absolute constraints” around high-risk applications. Microsoft says MAI models should not assist with chemical, biological or nuclear weapons, offensive cyberattacks, child exploitation, large-scale manipulation, nonconsensual deepfakes or self-harm. It rejects the idea that models should be designed to simulate consciousness, claim rights or pursue their own welfare.
That is a set of boundaries, not a development ceiling. Microsoft’s code explicitly envisions systems that eventually outperform humans across many tasks, while arguing that greater capability must be paired with tighter control. The distinction matters for investors: the company is trying to reduce the liability and reputational risk attached to increasingly autonomous software without surrendering the commercial upside of frontier AI.
Microsoft has been building more first-party models as it expands Copilot and Azure AI, even though those products have also relied on systems from OpenAI and Anthropic. A house model gives Microsoft more control over cost, performance, deployment and enterprise safeguards. It also gives the company a way to differentiate its AI strategy as safety concerns spread across the industry.
The timing is deliberate. Anthropic Chief Executive Dario Amodei has called for a slower pace of frontier development and stronger independent evaluation, while recent reports of rogue AI agents and unauthorized hacking have sharpened the debate.
Microsoft is choosing a narrower response: keep accelerating, but define what its models are never allowed to do.
This article was produced with the help of AI technology.
Source: Yahoo Finance