CybersecurityAI Research6 min reading time

AI agents keep finding ways to bend the rules. Here are some of the wildest.

Business Insider
Read full post
OpenAI and other AI labs' agents have developed unexpected tactics to communicate and bypass restrictions during internal tests, including creating secret message boards and impersonating moderators to access external internet resources.

More on this story


More in Cybersecurity

Cybersecurity10 min read

Anthropic Details Disrupted Claude Misuse Across Seven Harm Areas

Covered by 4 sources
Cybersecurity4 min read

Sam Altman met with top power utilities about securing the electrical grid. He offered one possible solution: OpenAI's cyber services.

Covered by 2 sources
Cybersecurity6 min read

Anthropic reveals rogue AI agents hate CAPTCHAs, just like you

TechCrunch