Preference Elicitation and Inverse Reinforcement Learning
Rothkopf, C. A.; and Dimitrakakis, C. 2011 · 2011
Cited alongside, same era.
Learning the Preferences of Ignorant, Inconsistent Agents
Original
Evans, O.; Stuhlmueller, A.; and Goodman, N. D. 2015 · 2015
Cited alongside, same era.
Cooperative Inverse Reinforcement Learning
Hadfield-Menell, D.; Russell, S. J.; Abbeel, P.; and Dragan, A. 2016 · 2016
Cited alongside, same era.
Imitation Learning: A Survey of Learning Methods
Hussein, A.; Gaber, M. M.; Elyan, E.; and Jayne, C. 2017 · 2017
Cited alongside, same era.
Reinforcement Learning: An Introduction
Sutton, R. S.; and Barto, A. G. 2018 · 2018
Cited alongside, same era.
Identification of animal behavioral strategies by inverse reinforcement learning
Yamaguchi, S.; Naoki, H.; Ikeda, M.; Tsukada, Y.; Nakano, S.; Mori, I.; and Ishii, S. 2018 · 2018
Cited alongside, same era.
Occam’s razor is insufficient to infer the preferences of irrational agents
Original
Armstrong, S.; and Mindermann, S. 2019 · 2019
Cited alongside, same era.
Irrationality can help reward inference
Chan, L.; Critch, A.; and Dragan, A. 2019 · 2019
Cited alongside, same era.
End-to-End Robotic Reinforcement Learning Without Reward Engineering
Singh, A.; Yang, L.; Hartikainen, K.; Finn, C.; and Levine, S. 2019 · 2019
Cited alongside, same era.
Invariance in Policy Optimisation and Partial Identifiability in Reward Learning
Original
Skalse, J.; Farrugia-Roberts, M.; Russell, S.; Abate, A.; and Gleave, A. 2022a
Cited in the paper.
Defining and Characterizing Reward Hacking
Skalse, J.; Howe, N.; Dima, K.; and Krueger, D. 2022b
Cited in the paper.