Fetching the paper…

A Minimaximalist Approach to Reinforcement Learning from Human Feedback · Around