Wave Pod
Discover
Library
Get Wave AI
Sign In
โ Daily Paper Cast
Daily Paper Cast
The Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning
July 7, 2026
ยท
00:21:25
Send to my inbox
Sign in to save
Share
Sign in to transcribe
Science
Technology