Tech article

Microsoft’s new AI ‘code of conduct’ tells models not to hack systems or trick humans

Publisher description: The code of conduct lays out general principles that Microsoft AI models should uphold — supporting humans rather than replacing them, for instance, and accelerating human flourishing — as well as specific safety constraints meant to implement those principles.

TechCrunch | Sep 14, 2026 | Russell Brandom

Automated excerpt

As the AI world shifts its focus to safety and alignment, Microsoft has released a new AI code of conduct meant to guide AI models away from dangerous behavior. The document is more low level than Anthropic CEO Dario Amodei’s recent call for pacing the frontier, instead focusing on the values and red lines that guide model training within Microsoft AI. Still, the result is a comprehensive guide as to how Microsoft approaches AI safety and how those ideas are implemented in practice.

Selected automatically from source text; not independently written or fact-checked. Read the original for full context.

Read the original article

More tech news