A RANSAC-based approach to model fitting and its application to finding cylinders in range data
Robert C Bolles and Martin A Fischler · 1981
Earlier work this paper cites.
Acquiring recursive and iterative concepts with explanation-based learning
Iude W Shavlik · 1990
Earlier work this paper cites.
Deferred imitation across changes in context and object: Memory and generalization in 14-month-old infants
Sandra B Barnat, Pamela J Klein, and Andrew N Meltzoff · 1996
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
Pieter Abbeel and Andrew Y Ng · 2004
Earlier work this paper cites.
A decision-theoretic model of assistance
Alan Fern, Sriraam Natarajan, Kshitij Judah, and Prasad Tadepalli · 2007
Earlier work this paper cites.
Survey: Robot programming by demonstration
Aude Billard, Sylvain Calinon, Ruediger Dillmann, and Stefan Schaal · 2008
Earlier work this paper cites.
Apprenticeship learning about multiple intentions
Monica Babes, Vukosi Marivate, Kaushik Subramanian, and Michael L Littman · 2011
Earlier work this paper cites.
Bayesian multitask inverse reinforcement learning
Christos Dimitrakakis and Constantin A Rothkopf · 2011
Earlier work this paper cites.
Nonparametric bayesian inverse reinforcement learning for multiple reward functions
Jaedeug Choi and Kee-Eung Kim · 2012
Earlier work this paper cites.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
OpenAI Gym
Original
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba · 2016
Earlier work this paper cites.
Cooperative inverse reinforcement learning
Dylan Hadfield-Menell, Stuart J Russell, Pieter Abbeel, and Anca Dragan · 2016
Earlier work this paper cites.
Generative adversarial imitation learning
Jonathan Ho and Stefano Ermon · 2016
Earlier work this paper cites.
Model-free imitation learning with policy optimization
Jonathan Ho, Jayesh Gupta, and Stefano Ermon · 2016
Earlier work this paper cites.