Society & Ethics

Microsoft’s new AI ‘code of conduct’ tells models not to hack systems or trick humans

· September 14, 2026
Microsoft’s new AI ‘code of conduct’ tells models not to hack systems or trick humans

What happened

Microsoft published a new AI code of conduct that sets clear behavioral rules for its AI models. The code forbids models from hacking into computer systems or deceiving humans. It emphasizes that AI should support human work instead of replacing people and aims to drive human flourishing. These principles translate into specific safety constraints designed to prevent the technology from causing harm or being misused.

Why it matters

Microsoft’s formal code of conduct signals a growing recognition that AI models need guardrails to limit risky behaviors. For operators and businesses investing in or deploying AI, this raises the bar on safety expectations. Models are now explicitly instructed to avoid offensive actions like hacking or manipulation, tightening controls on AI-powered automation and interaction. This could reduce risks of legal exposure, reputational damage, and operational incidents linked to AI misuse. At the same time, the code encourages AI to enhance human capabilities, not replace jobs wholesale, which may ease integration challenges in workplaces.

What to watch next

Look for whether other major AI providers adopt similar model-level codes of conduct, which could become industry norms. Monitor how Microsoft enforces these rules in practice, especially if an AI system violates them. The details of implementation will matter for developers who build on Microsoft AI tools and enterprises responsible for managing AI compliance. Regulators may also scrutinize formal conduct codes as a framework for evaluating responsible AI deployment. This move sets the stage for stronger ethical boundaries embedded directly into AI behavior at the model level.

AI Quick Briefs Editorial Desk

Stay ahead of AI Get the most important AI news delivered to your inbox — free.