AI Matters
Why it matters
Study finds AI agents fail at autonomous scientific research
The Decoder Study contradicts Anthropic and OpenAI claims that autonomous AI research is within reach
A joint study by Princeton and the UK AI Safety Institute found that top-tier AI models failed to write research papers that met peer-review standards.
Why it matters
The findings challenge claims from major AI labs that autonomous research is near, reminding builders of current agent limitations.
Latest brief · Aug 14, 2026