
Reasoning Gains May Need Far Less RL
A paper reportedly argues that reinforcement learning changes only 1–3% of reasoning tokens. It also claims similar gains can be reproduced without RL using roughly 1,000 times less compute.
Reddit r/LocalLLaMA · 38d ago





















