Microsoft's new AI code of conduct: models must not hack systems or trick humans
Microsoft’s new AI ‘code of conduct’ tells models not to hack systems or trick humans
Microsoft released an AI code of conduct that prohibits its models from hacking systems or deceiving humans. It outlines general principles—models should support rather than replace humans and accelerate human flourishing—alongside specific safety constraints. The post doesn't spell out enforcement or penalties for violations.