AdaTutoRank: Learning to Rerank Document Sets via Adaptive Tutoring Optimization for RAG and Deep Research
QUESTION — How can document set reranking for RAG and deep research be optimized to avoid sparse credit assignment and better capture complementary set composition?
The authors propose AdaTutoRank, a setwise reranker trained with Adaptive Tutoring Optimization (ATO) under a three-level hierarchy of nine rubric dimensions. The method supplies silver labels, reinforcement learning rewards, and hints derived from the policy's frozen snapshot to distill token-level advantages. Across ten benchmarks spanning RAG, deep research, and setwise evaluation, AdaTutoRank attains the best overall performance while issuing fewer retrieval calls.
Across ten benchmarks spanning RAG, deep research, and setwise evaluation, AdaTutoRank attains the best overall performance while issuing fewer retrieval calls.
kailinjiang · 26 Sept 2026
read the original ↗