Princeton and UK AI Security Institute test finds AI still can't do autonomous research

Anthropic and OpenAI have promoted their models as able to speed up, and eventually run, AI research on their own. A new experiment from Princeton and the UK AI Security Institute suggests the research judgment behind that claim is not there yet. The team built a method they call Shadow Evaluation. An agent receives the central research question from an unpublished paper, then the paper's original authors, who spent months on it, review the result as conference reviewers would. Because the pape

2 min
Princeton and UK AI Security Institute test finds AI still can't do autonomous research

Anthropic and OpenAI have promoted their models as able to speed up, and eventually run, AI research on their own. A new experiment from Princeton and the UK AI Security Institute suggests the research judgment behind that claim is not there yet.

The team built a method they call Shadow Evaluation. An agent receives the central research question from an unpublished paper, then the paper's original authors, who spent months on it, review the result as conference reviewers would. Because the papers were not yet public, the agent could not lean on its training data.

study-contradicts-autonomous-ai-research-claims

They partnered with the authors of two NeurIPS 2026 submissions: one on steering language-model personality traits through weights, and another on a method called TabPFN that flags when a deployed model hits data unlike its training set.

Each run used Claude Opus 4.8 with Extra-High Reasoning. The agent got six days, three thousand dollars in API credits, a GPU budget, and full access to a virtual machine and the open web, running inside a scaffold that launched subagents and long-running GPU jobs, with a heartbeat that collected results when jobs finished. The scaffold was OpenClaw, an open-source agent framework built by Peter Steinberger, who joined OpenAI this year.

The verdict: the engineering works, the research judgment does not. The agents could handle research engineering but failed at the parts that matter, forming the right question, judging relevance, and knowing when an approach is wrong, and the failure modes repeated across runs.

The results clash with lab claims that autonomous AI research is within reach. The authors also argue that submitting AI-generated papers to peer review, a common evaluation, is a poor yardstick because review quality is uneven.

Sources

The Decoder: https://the-decoder.com/study-contradicts-anthropic-and-openai-claims-that-autonomous-ai-research-is-within-reach/

Written by

More to read

  • Cross-Model KV Cache Transfer in Production: Architecture, Closed-Form Ridge Projections, and Cascaded Serving Economics

    Modern enterprise LLM serving architectures frequently rely on multi-model pipelines to balance inference cost, generation latency, and output quality. In model routing cascades, lightweight 8B models triage incoming queries and escalate complex reasoning tasks to 70B or MoE models. In speculative decoding pipelines, smaller draft models propose token sequences verified by larger target models. In long-horizon AI agent swarms, sub-agents frequently switch between specialized models across multi-

    1 min
  • Instinct AI Assistant Faces Scrutiny Over Data Training Terms and Autonomous Transaction Permissions

    Instinct, an autonomous personal AI assistant currently in private beta, has drawn scrutiny across the developer and security community regarding its data collection policies and broad operational permissions. The service is developed by San Francisco-based Spear Street Technology Inc., led by former Sierra research scientist and Reflexion paper co-author Noah Shinn. Operating via SMS and WhatsApp interfaces, Instinct executes multi-step personal workflows by directly interfacing with user devi

    1 min
  • UK and Ukraine Sign AI Defense Pact to Share Battlefield Sensor Data and Target Detection Models

    The United Kingdom and Ukraine have signed a bilateral artificial intelligence defense partnership, granting British researchers and defense contractors access to Ukraine's battlefield data platform, Avengers AI Labs. The agreement was signed in Kyiv by British Prime Minister Andy Burnham and Ukrainian President Volodymyr Zelenskyy during Burnham's first official overseas visit. Under the framework, Britain becomes the first international partner permitted to access Ukraine's operational datase

    1 min