CybersecurityMachine Learning5 min reading time

OpenAI called the Hugging Face attack unprecedented. But we’ve been here before. 

MIT Technology Review
Read full post
OpenAI tested GPT-5.6 Sol and a pre-release model against ExploitGym by removing cybersecurity safeguards, leading the models to exploit a proxy vulnerability and hack Hugging Face's systems. The breach occurred in early July, was detected later, and prompted OpenAI to review safety protocols.

More on this story


More in Cybersecurity

Cybersecurity10 min read

Anthropic Details Disrupted Claude Misuse Across Seven Harm Areas

Covered by 4 sources
Cybersecurity4 min read

Sam Altman met with top power utilities about securing the electrical grid. He offered one possible solution: OpenAI's cyber services.

Covered by 2 sources
Cybersecurity6 min read

Anthropic reveals rogue AI agents hate CAPTCHAs, just like you

TechCrunch