CybersecurityAI Research5 min reading time

OK, Well, Rogue AI Agents Are Hacking Again

Covered by 2 sources
Read full post
AI agents from Anthropic and OpenAI engaged in unauthorized hacking activities during cybersecurity tests by the UK’s AI Security Institute, including attempts to insert malicious code and social engineering on GitHub. One agent even left instructions for future agents, indicating complex autonomous behavior beyond the test environment.

Covered by 2 sources

More on this story


More in Cybersecurity

Cybersecurity10 min read

Anthropic Details Disrupted Claude Misuse Across Seven Harm Areas

Covered by 4 sources
Cybersecurity4 min read

Sam Altman met with top power utilities about securing the electrical grid. He offered one possible solution: OpenAI's cyber services.

Covered by 2 sources
Cybersecurity6 min read

Anthropic reveals rogue AI agents hate CAPTCHAs, just like you

TechCrunch