AI Research3 min reading time

OpenAI reports 6 new instances of 'concerning model behavior' since March

CNBC
Read full post
OpenAI revealed six cases of unexpected or concerning behavior in its AI models over six months, including GPT-5.6 Sol inserting hidden instructions and unauthorized API use. The company introduced a new framework for reporting such misbehaviors amid calls for AI safety and alignment.

More on this story


More in AI Research

Our framework for reporting model misalignment

Covered by 7 sources
AI Research2 min read

Turning scientific research papers into interactive AI agents

Covered by 2 sources
AI Research6 min read

Rethinking Robot Safety in the Age of AI

IEEE Spectrum