Explore
Discover
Solutions
AI Services
Our Products
Submit a tool
Search the catalog
⌘K
$
/
₹
Sign up free
Login
Back to papers
September 28, 2026
cs.AI
Verifier Errors in RLVR: Reward Hacking, Limits of Feedback, and Selective Control
Christian Moya
,
Elliott Thornley
,
Guang Lin
Original Abstract
Read on arXiv
Download PDF
Categories
cs.AI
Verifier Errors in RLVR: Reward Hacking, Limits of Feedback, and Selective Control - AI | One9Founders