ScienceBuddy: Recursive-in-Recursive Self-Improvement for Interactive Scientific Agents
QUESTION — How can scientific AI agents continuously co-evolve their harness environments and model weights through a recursive self-improvement loop?
The authors introduce ScienceBuddy, an interactive scientific research workspace implementing recursive-in-recursive self-improvement to continuously adapt AI agents. The paradigm couples harness evolution with model reinforcement learning: an inner recursion optimizes the execution harness with fixed model weights, while an outer recursion trains the model using the improved harness. Case studies spanning four scientific task families demonstrate how this joint evolution drives discovery intelligence in sustained collaboration with human researchers.
ScienceBuddy couples harness evolution with model reinforcement learning via a recursive-in-recursive self-improvement paradigm.
Lingaaaaaaa · 15 Sept 2026
read the original ↗