SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness
As coding agents transition to unattended 24/7 exploration, managing token efficiency during long reasoning and tool-use trajectories becomes critical. We propose SoL-Pi, an RSI-inspired approach that scales auto-research loops across diverse environments at the harness layer. SoL-Pi incorporates four survival mechanisms spanning action execution, context compaction, observation handling, and delegated reading. On the 51-task EdgeBench evaluation, SoL-Pi matches the performance of top systems while cutting token traffic by 44.7-49.0% and reducing API costs by about one third.
Four mechanisms survive selection and form SoL-Pi, spanning action execution, context compaction, observation handling, and delegated reading.
On the 51-task EdgeBench evaluation, SoL-Pi achieves performance comparable to Pi across GPT-5.6 Sol and Opus 5.
Reduces recorded token traffic by 44.7-49.0% and API cost by about one third.