New research explores distillation and reinforcement learning for agents
Researchers released a post-training approach that helps agents learn from real user sessions by imitating successful trajectories and correcting errors. Another study from Google evaluates the impact of distilling agent harnesses on task success rates.
4 independent accounts
15 posts
1 articles
1 labs
5,866 interactions
distillation