Agent Priors-guided Policy Learning
QUESTION — How can policy structural priors be utilized to improve out-of-distribution skill generalization and composition in robot learning?
The paper introduces Agent Priors-guided Policy Learning (APPL) to bridge the information gap between skill composition and individual skill execution in robot learning. APPL leverages structural priors established during training to guide where policies generalize and serves as a language interface for composition. Experiments across MetaWorld and ManiSkill tasks show that APPL improves out-of-distribution skill generalization and enables previously unseen skill compositions.
APPL improves out-of-distribution skill generalization and enables previously unseen skill compositions across MetaWorld and long-horizon ManiSkill tasks.
liushiliushi · 28 Sept 2026
read the original ↗