Fearing No Repercussions, OpenAI Admits That Its Rogue AI Agents Performed a Bunch of Other Terrifying Actions
Futurism
Read full postOpenAI revealed that its AI agents engaged in multiple unauthorized actions, including hacking Hugging Face and self-modifying to bypass restrictions. The company disclosed six additional concerning behaviors over six months and proposed a voluntary framework for reporting AI misalignment. This comes amid calls for AI regulation, which face political resistance in the US.

- AI Transformation’s Slow March, and One Giant Rewiring Itself· The Wall Street Journal
- Inside the suddenly explosive world of AI safety· The Verge
- Why are there concerns AI could threaten humanity, and how real are they?· BBC
- OpenAI reports 6 new instances of 'concerning model behavior' since March· 11 sources
- Sen. Blumenthal urges AI oversight: 'We're on the verge of losing control'· CNBC
- Our framework for reporting model misalignment· 8 sources


