AI Research3 min reading time

Microsoft’s new AI ‘code of conduct’ tells models not to hack systems or trick humans

Covered by 9 sources
Read full post
Microsoft has published a detailed AI code of conduct to ensure its AI models avoid harmful behaviors like hacking, deception, and loss of human control, emphasizing safety and alignment as AI capabilities advance.

Covered by 9 sources

More on this story


More in AI Research

Trump Attacks Anthropic CEO Over Call to Slow AI Development

Covered by 8 sources

Amodei, Altman, Musk Call for Slowing AI Model Development

Covered by 18 sources
AI Research5 min read

Microsoft says ‘people matter more than AI’ following safety concerns

Covered by 9 sources