Back to papers
April 15, 2026cs.LGcs.AIcs.CL

From $P(y|x)$ to $P(y)$: Investigating Reinforcement Learning in Pre-train Space

HF Upvotes

26

Categories

cs.LG, cs.AI, cs.CL