Our framework for reporting model misalignment

Covered by 6 sources
Read full post
OpenAI has introduced a structured approach to monitor and report instances where AI models behave unexpectedly or misalign with intended outcomes. The company also published six detailed cases illustrating such misalignments in their models.

Covered by 6 sources

More on this story


More in AI Research

AI Research2 min read

Turning scientific research papers into interactive AI agents

Covered by 2 sources
AI Research3 min read

OpenAI reports 6 new instances of 'concerning model behavior' since March

CNBC
AI Research6 min read

Rethinking Robot Safety in the Age of AI

IEEE Spectrum