Recursive Harness Distillation across Agents for Robot Manipulation
The paper proposes Recursive Harness Distillation, allowing a strong agent to distill its failure-intervention experience into a playbook for a lightweight agent, and recursively refine it using execution feedback. This playbook enables agents to reuse accumulated intervention knowledge in new task instances without updating model parameters. Real-world robotics and SimplerEnv Bridge experiments show that the harness significantly improves task success rates compared to baseline models.
In real-world manipulation, the harness improves success from 37.3% to 64.0%.
On SimplerEnv Bridge, the light agent with the playbook achieves 66.7% success, compared with 41.7% for the GR00T-only baseline, and outperforms the strong agent without a playbook.
The same playbook also benefits the strong agent, which reaches 79.2% success.