Fetching the paper…
Reading the bibliography…
Designing reward functions is a challenging problem in AI and robotics.
Batch active learning using determinantal point processes
Erdem Biyik, Kenneth Wang, Nima Anari, and Dorsa Sadigh · 1906
Earlier work this paper cites.
Policy invariance under reward transformations: Theory and application to reward shaping
Andrew Y Ng, Daishi Harada, and Stuart Russell · 1999
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
Andrew Y Ng and Stuart J Russell · 2000
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
Pieter Abbeel and Andrew Y Ng · 2004
Earlier work this paper cites.
Gaussian processes for machine learning
Carl Edward Rasmussen and Christopher KI Williams · 2005
Earlier work this paper cites.
Exploration and apprenticeship learning in reinforcement learning
Pieter Abbeel and Andrew Y Ng · 2005
Earlier work this paper cites.
Preference learning with gaussian processes
Wei Chu and Zoubin Ghahramani · 2005
Earlier work this paper cites.
Bayesian inverse reinforcement learning
Deepak Ramachandran and Eyal Amir · 2007
Earlier work this paper cites.
Active learning with gaussian processes for object categorization
Ashish Kapoor, Kristen Grauman, Raquel Urtasun, and Trevor Darrell · 2007
Earlier work this paper cites.
Fast poisson disk sampling in arbitrary dimensions
Robert Bridson · 2007
Earlier work this paper cites.
Maximum entropy inverse reinforcement learning
Brian D Ziebart, Andrew L Maas, J Andrew Bagnell, and Anind K Dey · 2008
Earlier work this paper cites.
Approximations for binary gaussian process classification
Hannes Nickisch and Carl Edward Rasmussen · 2008
Earlier work this paper cites.
Human preferences for robot-human hand-over configurations
Maya Cakmak, Siddhartha S Srinivasa, Min Kyung Lee, Jodi Forlizzi, and Sara Kiesler · 2011
Earlier work this paper cites.
Pairwise judgements and absolute ratings with gaussian process priors
Bjørn Sand Jensen and Jens Brehm Nielsen · 2011
Earlier work this paper cites.
Bayesian active learning for classification and preference learning
Neil Houlsby, Ferenc Huszár, Zoubin Ghahramani, and Máté Lengyel · 2011
Earlier work this paper cites.
April: Active preference learning-based reinforcement learning
Riad Akrour, Marc Schoenauer, and Michèle Sebag · 2012
Earlier work this paper cites.
Keyframe-based learning from demonstration
Baris Akgun, Maya Cakmak, Karl Jiang, and Andrea L Thomaz · 2012
Cited alongside, same era.
Formalizing assistive teleoperation
Anca D Dragan and Siddhartha S Srinivasa · 2012
Cited alongside, same era.
Collaborative gaussian processes for preference learning
Neil Houlsby, Ferenc Huszar, Zoubin Ghahramani, and Jose M Hernández-Lobato · 2012
Cited alongside, same era.
Mujoco: A physics engine for model-based control
Emanuel Todorov, Tom Erez, and Yuval Tassa · 2012
Cited alongside, same era.
Learning monocular reactive uav control in cluttered natural environments
Stéphane Ross, Narek Melik-Barkhudarov, Kumar Shaurya Shankar, Andreas Wendel, Debadeepta Dey, J Andrew Bagnell, and Martial Hebert · 2013
Cited alongside, same era.
Bayesian preference elicitation for multiobjective engineering design optimization
John R Lepird, Michael P Owen, and Mykel J Kochenderfer · 2015
Do you want your autonomous car to drive like you?
Chandrayee Basu, Qian Yang, David Hungerman, Mukesh Sinahal, and Anca D Dragan · 2017
Later among the works it cites.
Learning robot objectives from physical human interaction
Andrea Bajcsy, Dylan P Losey, Marcia K O’Malley, and Anca D Dragan · 2017
Later among the works it cites.
Preferential bayesian optimization
Javier González, Zhenwen Dai, Andreas Damianou, and Neil D Lawrence · 2017
Later among the works it cites.
Multi-agent generative adversarial imitation learning
Jiaming Song, Hongyu Ren, Dorsa Sadigh, and Stefano Ermon · 2018
Later among the works it cites.
Reward learning from human preferences and demonstrations in atari
Borja Ibarz, Jan Leike, Tobias Pohlen, Geoffrey Irving, Shane Legg, and Dario Amodei · 2018
Later among the works it cites.
Batch active preference-based learning of reward functions
Erdem Biyik and Dorsa Sadigh · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Shared autonomy via hindsight optimization
Shervin Javdani, Siddhartha S Srinivasa, and J Andrew Bagnell · 2015
Cited alongside, same era.
Data-driven motion mappings improve transparency in teleoperation
Rebecca P Khurshid and Katherine J Kuchenbecker · 2015
Cited alongside, same era.
Active reward learning with a novel acquisition function
Christian Daniel, Oliver Kroemer, Malte Viering, Jan Metz, and Jan Peters · 2015
Cited alongside, same era.
Generative adversarial imitation learning
Jonathan Ho and Stefano Ermon · 2016
Cited alongside, same era.
Faulty reward functions in the wild, Dec 2016
Jack Clark and Dario Amodei · 2016
Cited alongside, same era.
Planning for autonomous cars that leverage effects on human actions
Dorsa Sadigh, Shankar Sastry, Sanjit A Seshia, and Anca D Dragan · 2016
Cited alongside, same era.
Later among the works it cites.
Learning from physical human corrections, one feature at a time
Andrea Bajcsy, Dylan P Losey, Marcia K O’Malley, and Anca D Dragan · 2018
Later among the works it cites.
Deep bayesian reward learning from preferences
Daniel S Brown and Scott Niekum · 2019
Later among the works it cites.
Bayesian active learning for collaborative task specification using equivalence regions
Nils Wilde, Dana Kulić, and Stephen L Smith · 2019
Later among the works it cites.
Active learning of reward dynamics from hierarchical queries
Chandrayee Basu, Erdem Biyik, Zhixun He, Mukesh Singhal, and Dorsa Sadigh · 2019
Later among the works it cites.
Learning reward functions by integrating human demonstrations and preferences
Malayandi Palan, Nicholas C. Landolfi, Gleb Shevchuk, and Dorsa Sadigh · 2019
Later among the works it cites.
Learning an urban air mobility encounter model from expert preferences
Sydney M Katz, Anne-Claire Le Bihan, and Mykel J Kochenderfer · 2019
Later among the works it cites.
Teacher-aware active robot learning
Mattia Racca, Antti Oulasvirta, and Ville Kyrki · 2019
Later among the works it cites.
Preference-based learning for exoskeleton gait optimization
Maegan Tucker, Ellen Novoseller, Claudia Kann, Yanan Sui, Yisong Yue, Joel Burdick, and Aaron D Ames · 2020
Closest in time.
Controlling assistive robots with learned latent actions
Dylan P. Losey, Krishnan Srinivasan, Ajay Mandlekar, Animesh Garg, and Dorsa Sadigh · 2020
Closest in time.