Fetching the paper…

Not All Rollouts are Useful: Down-Sampling Rollouts in LLM Reinforcement Learning · Around