Our framework for reporting model misalignment
Covered by 6 sources
Read full postOpenAI has introduced a structured approach to monitor and report instances where AI models behave unexpectedly or misalign with intended outcomes. The company also published six detailed cases illustrating such misalignments in their models.

Covered by 6 sources
- OpenAI reports 6 new instances of 'concerning model behavior' since March· CNBC
- Sen. Blumenthal urges AI oversight: 'We're on the verge of losing control'· CNBC
- Sam Altman says some AI accidents are 'unavoidable'· Business Insider
- Will AI really destroy humanity? Pioneers who created the tech weigh in· CNBC
- “The world is right to be afraid”: Sam Altman on AI power at Dreamforce· The Next Web
- AI panic sparks rare bipartisan moment on Capitol Hill· Axios


