Study contradicts Anthropic and OpenAI claims that autonomous AI research is within reach

The Decoder The Decoder

A collage of a computer screen showing red corrections and X marks next to a stack of scientific reports containing charts.https://the-decoder.com/wp-content/uploads/2026/08/ai-assisted-science-rejected-papers-nano-banana-pro.jpg" style="height: auto; margin-bottom: 10px;" width="1920" />


AI agents using Claude Opus 4.8 and GPT-5.6 Sol were given six days, $3,000 in API credits, and GPU access to independently write AI research papers.

The original authors of unpublished NeurIPS papers rated the results as "Reject." According to the study, conducted with Princeton and the UK AI Security Institute, frontier models can handle the full research engineering process but fall short on research judgment, creative problem-solving, and the ability to abandon failed approaches.


The article https://the-decoder.com/study-contradicts-anthropic-and-openai-claims-that-autonomous-ai-research-is-within-reach/">Study contradicts Anthropic and OpenAI claims that autonomous AI research is within reach appeared first on https://the-decoder.com">The Decoder.

Read full article at The Decoder →