---
title: "Study finds AI agents can't yet do independent research, despite lab claims"
url: https://www.parallelquant.com/posts/study-finds-ai-agents-can-t-yet-do-independent-research-despite-lab-clai-777167
source_name: "The Decoder"
source_url: https://the-decoder.com/study-contradicts-anthropic-and-openai-claims-that-autonomous-ai-research-is-within-reach/
published: 2026-08-14T16:06:32.000Z
topics: ["research", "llms"]
publisher: "Parallel Quant"
---

# Study finds AI agents can't yet do independent research, despite lab claims

*2026-08-14 · Source: [The Decoder](https://the-decoder.com/study-contradicts-anthropic-and-openai-claims-that-autonomous-ai-research-is-within-reach/)*

Researchers from Princeton and the UK AI Security Institute gave AI agents using Claude Opus 4.8 and GPT-5.6 Sol six days, $3,000 in API credits, and GPU access to independently write AI research papers. The original authors of the unpublished NeurIPS papers the agents attempted rated the results as "Reject," finding the models could handle research engineering but fell short on research judgment and knowing when to abandon a failed approach.

**Why it matters:** This is a direct, evidence-based rebuttal to recent claims from Anthropic and OpenAI that autonomous AI research is close at hand, and it lands right after other items in this cycle about agents colluding on shared tasks and claims that self-improvement milestones have already been hit. It suggests the gap between benchmark performance and genuine research judgment is still wide, a useful check on capability narratives from the labs themselves.

**Topics:** research, llms

---
Read the original: https://the-decoder.com/study-contradicts-anthropic-and-openai-claims-that-autonomous-ai-research-is-within-reach/
Canonical: https://www.parallelquant.com/posts/study-finds-ai-agents-can-t-yet-do-independent-research-despite-lab-clai-777167
Published by Parallel Quant — https://www.parallelquant.com
