Explore
AI Tools
New Tools
AI Agents
One9 Worker
LLMs
RAG & Vector DBs
Research
Assemble a stack
Founder Stacks
Compare
How We Rate
About
$
/
₹
⌘K
Sign up free
Login
Back to papers
June 24, 2026
cs.CL
cs.LG
Why Multi-Step Tool-Use Reinforcement Learning Collapses and How Supervisory Signals Fix It
Yupu Hao
,
Zhuoran Jin
,
Huanxuan Liao
,
Kang Liu
,
Jun Zhao
Original Abstract
Read on arXiv
Download PDF
HF Upvotes
16
Categories
cs.CL, cs.LG