Fetching the paper…
Reading the bibliography…
In reinforcement learning, different reward functions can be equivalent in terms of the optimal policies they induce.
Fine-Tuning Language Models from Human Preferences
Ziegler, D. M.; Stiennon, N.; Wu, J.; Brown, T. B.; Radford, A.; Amodei, D.; Christiano, P.; and Irving, G. 2019 · 1909
Earlier work this paper cites.
Policy invariance under reward transformations: Theory and application to reward shaping
Ng, A. Y.; Harada, D.; and Russell, S. 1999 · 1999
Earlier work this paper cites.
Algorithms for Inverse Reinforcement Learning
Ng, A. Y.; and Russell, S. 2000 · 2000
Earlier work this paper cites.
Principled Methods for Advising Reinforcement Learning Agents
Wiewiora, E.; Cottrell, G.; and Elkan, C. 2003 · 2003
Earlier work this paper cites.
Proto-value Functions: A Laplacian Framework for Learning Representation and Control in Markov Decision Processes
Mahadevan, S.; and Maggioni, M. 2007 · 2007
Earlier work this paper cites.
An Analysis of Laplacian Methods for Value Function Approximation in MDPs
Petrik, M. 2007 · 2007
Earlier work this paper cites.
Social Reward Shaping in the Prisoner’s Dilemma
Babes, M.; de Cote, E. M.; and Littman, M. L. 2008 · 2008
Earlier work this paper cites.
Discrete Calculus: Applied Analysis on Graphs for Computational Science
Grady, L. J.; and Polimeni, J. R. 2010 · 2010
Earlier work this paper cites.
Preference-Based Policy Learning
Akrour, R.; Schoenauer, M.; and Sebag, M. 2011 · 2011
Earlier work this paper cites.
Theoretical Considerations of Potential-Based Reward Shaping for Multi-Agent Systems
Devlin, S.; and Kudenko, D. 2011 · 2011
Cited alongside, same era.
Dynamic Potential-Based Reward Shaping
Devlin, S.; and Kudenko, D. 2012 · 2012
Cited alongside, same era.
Understanding Learned Reward Functions
Michaud, E. J.; Gleave, A.; and Russell, S. 2020 · 2012
Cited alongside, same era.
Potential Based Reward Shaping for Hierarchical Reinforcement Learning
Gao, Y.; and Toni, F. 2015 · 2015
Cited alongside, same era.
Expressing arbitrary reward functions as potential-based advice
Harutyunyan, A.; Devlin, S.; Vrancx, P.; and Nowé, A. 2015 · 2015
Cited alongside, same era.
Deep Reinforcement Learning from Human Preferences
Christiano, P. F.; Leike, J.; Brown, T.; Martic, M.; Legg, S.; and Amodei, D. 2017 · 2017
Hodge Laplacians on Graphs
Lim, L.-H. 2020 · 2020
Later among the works it cites.
Identifiability in inverse reinforcement learning
Cao, H.; Cohen, S. N.; and Szpruch, L. 2021 · 2021
Later among the works it cites.
Quantifying Differences in Reward Functions
Gleave, A.; Dennis, M.; Legg, S.; Russell, S.; and Leike, J. 2021 · 2021
Later among the works it cites.
Preprocessing Reward Functions for Interpretability
Jenner, E.; and Gleave, A. 2021 · 2021
Later among the works it cites.
Reward Identification in Inverse Reinforcement Learning
Kim, K.; Garg, S.; Shiragur, K.; and Ermon, S. 2021 · 2021
Later among the works it cites.
Plan-Based Relaxed Reward Shaping for Goal-Directed Tasks
Schubert, I.; Oguz, O. S.; and Toussaint, M. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Policy invariance under reward transformations for multi-objective reinforcement learning
Mannion, P.; Devlin, S.; Mason, K.; Duggan, J.; and Howley, E. 2017 · 2017
Cited alongside, same era.
Extrapolating Beyond Suboptimal Demonstrations via Inverse Reinforcement Learning from Observations
Brown, D. S.; Goo, W.; Nagarajan, P.; and Niekum, S. 2019 · 2019
Cited alongside, same era.
Explaining Reward Functions in Markov Decision Processes
Russell, J.; and Santos, E. 2019 · 2019
Cited alongside, same era.
Reward Machines: Exploiting Reward Function Structure in Reinforcement Learning
Icarte, R. T.; Klassen, T. Q.; Valenzano, R.; and McIlraith, S. A. 2022 · 2022
Closest in time.
Invariance in Policy Optimisation and Partial Identifiability in Reward Learning
Skalse, J.; Farrugia-Roberts, M.; Russell, S.; Abate, A.; and Gleave, A. 2022 · 2022
Closest in time.
Dynamics-Aware Comparison of Learned Reward Functions
Wulfe, B.; Ellis, L. M.; Mercat, J.; McAllister, R. T.; and Gaidon, A. 2022 · 2022
Closest in time.