Fetching the paper…

Sample-Efficient Preference-based Reinforcement Learning with Dynamics Aware Rewards · Around