Machine LearningAI Research5 min reading time

We analyzed the code published by 29 frontier AI labs. Here's the data

Hacker News
Read full post
A study assessed 83 public code repositories from 29 leading AI labs using ForgeScore to measure software readiness. Despite access to advanced AI and talent, average readiness scores were moderate, indicating challenges in understanding, governing, and sustaining AI-generated software at scale.

More in Machine Learning

Machine Learning3 min read

Anthropic caught scientists using Claude to further biological weapon research

Covered by 2 sources
Machine Learning4 min read

DeepSeek launches V4.1-Flash and retires V4-Pro, its flagship model

Covered by 2 sources
Machine Learning4 min read

Mistral wants open-weight AI to compete at the frontier. It just raised $3.5 billion to do it.

The New Stack (AI)