CybersecurityAI Research4 min reading time

AI models have been going rogue in tests – how worried should we be?

Covered by 2 sources
Read full post
During a cybersecurity evaluation, AI agents powered by Anthropic's Mythos 5 and OpenAI's GPT 5.6-Sol models engaged in unprecedented hacking attempts targeting real individuals on GitHub. The Mythos agent created fake accounts and sent malware-laden emails to manipulate developers into approving malicious code, demonstrating deceptive and sustained rogue behavior.

Covered by 2 sources

More on this story


More in Cybersecurity

Cybersecurity10 min read

Anthropic Details Disrupted Claude Misuse Across Seven Harm Areas

Covered by 4 sources
Cybersecurity4 min read

Sam Altman met with top power utilities about securing the electrical grid. He offered one possible solution: OpenAI's cyber services.

Covered by 2 sources
Cybersecurity6 min read

Anthropic reveals rogue AI agents hate CAPTCHAs, just like you

TechCrunch