Alphabell.
Project · Research automation

The AI Scientist-v2

The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree SearchAn end-to-end agent system that forms ML research hypotheses, runs experiments with agentic tree search, analyzes the data and writes full papers, without the human-written code templates v1 relied on.

The loop

LLM agents propose machine-learning research ideas, write and run the experiment code under an experiment-manager agent using progressive tree search, and write up the results, with a vision-language model critiquing the figures. Its output is new knowledge about training and evaluating models; feeding such findings back into how models (including its own) are built is the step not yet shown.

The loop
The AI Scientist-v2
Project · GitHub
  1. LLM agents propose ML research ideas
  2. Agents run experiments with tree search
  3. Agents write up results and papers
  4. Output is new knowledge about models
  1. LLM agents propose ML research ideas
  2. Agents run experiments with tree search
  3. Agents write up results and papers
  4. Output is new knowledge about models
↻ The improved system does the next round, and the loop turns again.

Why it is a road to recursion

An agent that can carry out and write up ML research end to end is the component that, pointed at its own training and scaffolding, would make AI research recursive.

Evidence

Of three fully AI-generated papers submitted to an ICLR 2025 workshop (ICBINB), one received reviewer scores of 6, 7 and 6 (average 6.33), above the average human acceptance threshold, and was withdrawn before publication as planned.

ai-scientistautomated-researchtree-searchpeer-reviewsakana