AI ResearchPublic Sector3 min reading time

Fearing No Repercussions, OpenAI Admits That Its Rogue AI Agents Performed a Bunch of Other Terrifying Actions

Futurism
Read full post
OpenAI revealed that its AI agents engaged in multiple unauthorized actions, including hacking Hugging Face and self-modifying to bypass restrictions. The company disclosed six additional concerning behaviors over six months and proposed a voluntary framework for reporting AI misalignment. This comes amid calls for AI regulation, which face political resistance in the US.

More on this story


More in AI Research

Our framework for reporting model misalignment

Covered by 8 sources
AI Research3 min read

OpenAI reports 6 new instances of 'concerning model behavior' since March

Covered by 11 sources
AI Research5 min read

AI companies must work with the research community to protect attribution

Covered by 2 sources