Back to papers
August 7, 2026cs.LGcs.CL

Trajectory-Relative Hindsight Distillation for Agentic Reinforcement Learning

Categories

cs.LG, cs.CL