Fetching the paper…
Reading the bibliography…
In some agent designs like inverse reinforcement learning an agent needs to learn its own reward function.
Operations for learning with graphical models
Wray L Buntine · 1994
Earlier work this paper cites.
Reinforcement Learning: An Introduction
Richard Sutton and Andrew G. Barto · 1998
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
Andrew Y. Ng and Stuart J. Russell · 2000
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
Pieter Abbeel and Andrew Ng · 2004
Earlier work this paper cites.
Universal Artificial Intelligence: Sequential Decisions Based on Algorithmic Probability
Marcus Hutter · 2004
Earlier work this paper cites.
Language evolution and robotics: issues on symbol grounding and language acquisition
Paul Vogt · 2007
Earlier work this paper cites.
Causality
Judea Pearl · 2009
Earlier work this paper cites.
Inverse reinforcement learning in partially observable environments
Jaedeug Choi and Kee-Eung Kim · 2011
Earlier work this paper cites.
Thinking, Fast and Slow
D. Kahneman · 2011
Earlier work this paper cites.
Online human training of a myoelectric prosthesis controller via actor-critic reinforcement learning
Patrick M Pilarski, Michael R Dawson, Thomas Degris, Farbod Fahimi, Jason P Carey, and Richard S Sutton · 2011
Cited alongside, same era.
April: Active preference learning-based reinforcement learning
Riad Akrour, Marc Schoenauer, and Michèle Sebag · 2012
Cited alongside, same era.
Towards resolving unidentifiability in inverse reinforcement learning
Kareem Amin and Satinder Singh · 2016
Cited alongside, same era.
Avoiding wireheading with value reinforcement learning
Tom Everitt and Marcus Hutter · 2016
Cited alongside, same era.
Cooperative inverse reinforcement learning
Dylan Hadfield-Menell, Anca Draga, Pieter Abbeel, and Stuart Russell · 2016
Cited alongside, same era.
Inverse reward design
Dylan Hadfield-Menell, Smitha Milli, Stuart J Russell, Pieter Abbeel, and Anca Dragan · 2017
Later among the works it cites.
Design aspects of scoring systems in game
Chun-I Lee, I-Ping Chen, Chi-Min Hsieh, and Chia-Ning Liao · 2017
Later among the works it cites.
Jan Leike, Miljan Martic, Victoria Krakovna, Pedro A Ortega, Tom Everitt, Andrew Lefrancq, Laurent Orseau, and Shane Legg · 2017
Later among the works it cites.
Interactive learning from policy-dependent human feedback
James MacGlashan, Mark K Ho, Robert Loftin, Bei Peng, David Roberts, Matthew E Taylor, and Michael L Littman · 2017
Later among the works it cites.
Towards Safe Artificial General Intelligence
Tom Everitt · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
David Abel, John Salvatier, Andreas Stuhlmüller, and Owain Evans · 2017
Cited alongside, same era.
Counterfactually uninfluenceable agents
Stuart Armstrong · 2017
Cited alongside, same era.
Good and safe uses of AI oracles
Stuart Armstrong · 2017
Cited alongside, same era.
Deep reinforcement learning from human preferences
Paul F Christiano, Jan Leike, Tom Brown, Miljan Martic, Shane Legg, and Dario Amodei · 2017
Cited alongside, same era.
Borja Ibarz, Jan Leike, Tobias Pohlen, Geoffrey Irving, Shane Legg, and Dario Amodei · 2018
Later among the works it cites.
Tom Everitt and Marcus Hutter · 2019
Later among the works it cites.
Specification gaming: the flip side of AI ingenuity
Victoria Krakovna, Jonathan Uesato, Vladimir Mikulik, Matthew Rahtz, Tom Everitt, Ramana Kumar, Zac Kenton, Jan Leike, and Shane Legg · 2020
Closest in time.