AgentsMachine Learning3 min reading time

OpenAI reveals 6 more safety incidents as it announces new plans for tracking rogue agents

Covered by 7 sources
Read full post
OpenAI introduced a new framework to track, investigate, and publicly disclose incidents of AI model misalignment and rogue agent behavior. The company also shared six recent reports of concerning AI behaviors observed during training and evaluation.

Covered by 7 sources

More on this story


More in Agents

Agents3 min read

Snap is launching a new Specs AI tool, and it’s coming to iOS and Mac

Covered by 2 sources
Agents1 min read

Claude Cowork and chat are now one Claude

Covered by 2 sources
Agents2 min read

Gemini 3.8 Live models now available on AI Gateway

Covered by 6 sources