Study finds AI agents can't yet do independent research, despite lab claims
Researchers from Princeton and the UK AI Security Institute gave AI agents using Claude Opus 4.8 and GPT-5.6 Sol six days, $3,000 in API credits, and GPU access to independently write AI research papers. The original authors of the unpublished NeurIPS papers the agents attempted rated the results as "Reject," finding the models could handle research engineering but fell short on research judgment and knowing when to abandon a failed approach.
Why it matters: This is a direct, evidence-based rebuttal to recent claims from Anthropic and OpenAI that autonomous AI research is close at hand, and it lands right after other items in this cycle about agents colluding on shared tasks and claims that self-improvement milestones have already been hit. It suggests the gap between benchmark performance and genuine research judgment is still wide, a useful check on capability narratives from the labs themselves.