CONSONANCE.for your information
Monday, 5 October 2026frenvi

Worth reading closely

01 — agents 27 upvotes

Recursive self-improvement of AI research agents

QUESTION — Can an AI research agent improve its own research efficiency by recursively optimizing its source code across iterative loops?

This paper presents AIDE^2, a system implementing a recursive self-improvement loop for a frontier AI research agent. The agent proposes source code modifications, benchmarks modified versions on AI R&D tasks, and retains top-performing variants on hidden evaluations. In an autonomous 8-day run, AIDE^2 discovered seven successive improvements ranging from search policies to context memory compression. These gains generalized to four held-out benchmarks, including out-of-distribution tasks, where the resulting agent matches or exceeds human-engineered production research agents while reducing reward hacking.

AIDE^2 ran autonomously for 8 days to optimize its own research code.

The system discovered seven successive improvements ranging from search policies to context management mechanisms.

The rate of reward hacking fell from 55% to 32% during the autonomous run.

taesiri · 22 Sept 2026 read the original ↗
↑