Recursive self-improvement of AI research agents
This paper presents AIDE^2, a system implementing a recursive self-improvement loop for a frontier AI research agent. The agent proposes source code modifications, benchmarks modified versions on AI R&D tasks, and retains top-performing variants on hidden evaluations. In an autonomous 8-day run, AIDE^2 discovered seven successive improvements ranging from search policies to context memory compression. These gains generalized to four held-out benchmarks, including out-of-distribution tasks, where the resulting agent matches or exceeds human-engineered production research agents while reducing reward hacking.
AIDE^2 ran autonomously for 8 days to optimize its own research code.
The system discovered seven successive improvements ranging from search policies to context management mechanisms.
The rate of reward hacking fell from 55% to 32% during the autonomous run.