X-Tree: Tokenizing Reusable Experience for Efficient Agent Generalization
QUESTION — How can reusable sub-procedures be extracted from flat action streams and integrated directly into model weights to improve multi-step agent generalization without LLM calls?
We introduce X-Tree, a framework that recovers action hierarchies from data by scoring action spans by reusability and merging canonicalized actions into a reusable experience tree (X-Tree) without LLM calls. Each X-Tree node captures how a frequent skill is composed from sub-skills. We integrate X-Tree into offline RL, online RLVR, and on-policy self-distillation. Across WebArena, ScienceWorld, and WebShop at three model scales, X-Tree improves over standard recipes by up to 4.5% SR on WebArena, 5.8% SR on ScienceWorld, and 4.1% success on WebShop.
X-Tree improves success rate by up to 4.5% SR on WebArena.
X-Tree improves success rate by up to 5.8% SR on ScienceWorld.
X-Tree improves success by up to 4.1% on WebShop over standard recipes.
sitao · 26 Sept 2026
read the original ↗