Cooperative inverse reinforcement learning
Dylan Hadfield-Menell, Stuart J Russell, Pieter Abbeel, and Anca Dragan · 2016
Later among the works it cites.
Planning for autonomous cars that leverage effects on human actions
Dorsa Sadigh, Shankar Sastry, Sanjit A Seshia, and Anca D Dragan · 2016
Later among the works it cites.
An alternative softmax operator for reinforcement learning
Kavosh Asadi and Michael L Littman · 2017
Later among the works it cites.
Uncertain reward-transition mdps for negotiable reinforcement learning
Nishant Desai · 2017
Later among the works it cites.
Learning modular neural network policies for multi-task and multi-robot transfer
Coline Devin, Abhishek Gupta, Trevor Darrell, Pieter Abbeel, and Sergey Levine · 2017
Later among the works it cites.
Pragmatic-pedagogic value alignment
Original
Jaime F Fisac, Monica A Gates, Jessica B Hamrick, Chang Liu, Dylan Hadfield-Menell, Malayandi Palaniappan, Dhruv Malik, S Shankar Sastry, Thomas L Griffiths, and Anca D Dragan · 2017
Later among the works it cites.
Active preference-based learning of reward functions
Dorsa Sadigh, Anca D. Dragan, Shankar Sastry, and Sanjit A Seshia · 2017
Later among the works it cites.
Risk-sensitive inverse reinforcement learning via semi-and non-parametric methods
Original
Sumeet Singh, Jonathan Lacotte, Anirudha Majumdar, and Marco Pavone · 2017
Later among the works it cites.
Learning robust rewards with adverserial inverse reinforcement learning
Justin Fu, Katie Luo, and Sergey Levine · 2018
Later among the works it cites.
The surprising creativity of digital evolution: A collection of anecdotes from the evolutionary computation and artificial life research communities
Original
Joel Lehman, Jeff Clune, Dusan Misevic, Christoph Adami, Lee Altenberg, Julie Beaulieu, Peter J Bentley, Samuel Bernard, Guillaume Beslon, David M Bryson, et al · 2018
Later among the works it cites.
Where do you think you’re going?: Inferring beliefs about dynamics from behavior
Sid Reddy, Anca Dragan, and Sergey Levine · 2018
Later among the works it cites.
Task transfer by preference-based cost learning
Mingxuan Jing, Xiaojian Ma, Wenbing Huang, Fuchun Sun, and Huaping Liu · 2019
Later among the works it cites.