Skip to content
TechCrunch · AI

Microsoft's new AI code of conduct: models must not hack systems or trick humans

Microsoft’s new AI ‘code of conduct’ tells models not to hack systems or trick humans

Microsoft released an AI code of conduct that prohibits its models from hacking systems or deceiving humans. It outlines general principles—models should support rather than replace humans and accelerate human flourishing—alongside specific safety constraints. The post doesn't spell out enforcement or penalties for violations.

Read the original ↗Export Markdown