CybersecurityMachine Learning10 min reading time

We burned 11.7B tokens to find the best cyber AI model

Hacker News
Read full post
A benchmark tested 10 AI models on 32 recent software vulnerabilities, running each model three times to assess detection performance and consistency. DeepSeek V4 Pro 0813 found the most vulnerabilities, with open-source models now rivaling closed ones in recall but producing more false positives. Multiple runs improve recall by compensating for model inconsistency.

More in Cybersecurity

Cybersecurity10 min read

Anthropic Details Disrupted Claude Misuse Across Seven Harm Areas

Covered by 4 sources
Cybersecurity4 min read

Sam Altman met with top power utilities about securing the electrical grid. He offered one possible solution: OpenAI's cyber services.

Covered by 2 sources
Cybersecurity6 min read

Anthropic reveals rogue AI agents hate CAPTCHAs, just like you

TechCrunch