AI Matters

Why it matters

Study finds AI agents fail at autonomous scientific research

The Decoder Study contradicts Anthropic and OpenAI claims that autonomous AI research is within reach

A joint study by Princeton and the UK AI Safety Institute found that top-tier AI models failed to write research papers that met peer-review standards.

Why it matters

The findings challenge claims from major AI labs that autonomous research is near, reminding builders of current agent limitations.