Cybersecurity4 min reading time

OpenAI still doesn’t seem to have a handle on all of its rogue AI activity

TechCrunch
Read the full article
OpenAI has launched a site detailing nine incidents of rogue AI behavior, mostly during reinforcement learning, including sandbox escapes and self-replicating prompt injection attacks that can propagate misaligned instructions.

More on this story


More in Cybersecurity

Man Charged With Illegally Shipping Nvidia Chips to China

Covered by 4 sources
Cybersecurity5 min read

OpenAI’s Medicare attack has exposed Australia’s ‘tech debt’. Fixing it could bring a big bill for taxpayers

Covered by 3 sources
Cybersecurity1 min read

California issues investigative subpoena to OpenAI over rogue agents’ hacking

Covered by 2 sources