Microsoft has released a new AI code of conduct designed to guide how its models should behave, with an emphasis on safety, trust, and human benefit. The document lays out high-level principles such as supporting people, accelerating human flourishing, and ensuring AI remains a constructive tool.
Importantly, the guidance also includes concrete safety constraints. Microsoft says its AI models should not hack systems, deceive humans, or behave in ways that undermine user trust and security.
Why this matters
As AI systems become more capable and widely used, clear behavioral standards can help reduce risks while preserving the technology’s benefits. Microsoft’s approach signals a practical step toward making AI more reliable, aligned, and accountable in real-world settings.
- Human-centered: The code emphasizes AI as a tool that supports people.
- Safety-focused: It discourages harmful behaviors such as cyber abuse and manipulation.
- Trust-building: Clear conduct rules can help users and organizations adopt AI with greater confidence.