Fetching the paper…
Reading the bibliography…
Observational learning is a type of learning that occurs as a function of observing, retaining and possibly replicating or imitating the behaviour of another agent.
Social learning and personality development
Albert Bandura and Richard H Walters · 1963
Earlier work this paper cites.
Social learning theory
Albert Bandura and Richard H Walters · 1977
Earlier work this paper cites.
Alvinn: An autonomous land vehicle in a neural network
Dean A. Pomerleau · 1989
Earlier work this paper cites.
Imitation, culture and cognition
Cecilia M Heyes · 1993
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Robot learning from demonstration
C.G. Atkeson and S. Schaal · 1997
Earlier work this paper cites.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 1998
Earlier work this paper cites.
Learning agents for uncertain environments
S. Russell · 1998
Earlier work this paper cites.
Is imitation learning the route to humanoid robots?
Stefan Schaal · 1999
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
Andrew Ng and Stuart Russell · 2000
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
Peter Abbeel and Andrew Y. Ng · 2004
Cited alongside, same era.
Imitation learning for locomotion and manipulation
Nathan Ratliff, J.Andrew Bagnell, and Siddhartha S. Srinivasa · 2007
Cited alongside, same era.
Maximum entropy inverse reinforcement learning
Brian D Ziebart, Andrew L Maas, J Andrew Bagnell, and Anind K Dey · 2008
Cited alongside, same era.
Hierarchical apprenticeship learning with application to quadruped locomotion
J Zico Kolter, Pieter Abbeel, and Andrew Y Ng · 2008
Cited alongside, same era.
A survey of robot learning from demonstration
Brenna. Argall, Sonnia Chernova, Manuella Veloso, and Brett Browning · 2009
Cited alongside, same era.
Training parsers by inverse reinforcement learning
Gergely Neu and Csaba Szepesvári · 2009
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Later among the works it cites.
Deep learning
Yann LeCun, Yoshua Bengio, and Geoffrey Hinton · 2015
Later among the works it cites.
Gated feedback recurrent neural networks
Junyoung Chung, Caglar Gulcehre, Kyunghyun Cho, and Yoshua Bengio · 2015
Later among the works it cites.
Asynchronous methods for deep reinforcement learning
Volodymyr Mnih, Adria Puigdomenech Badia, Mehdi Mirza, Alex Graves, Timothy Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu · 2016
Later among the works it cites.
End-to-end training of deep visuomotor policies
Sergey Levine, Chelsea Finn, Trevor Darrell, and Pieter Abbeel · 2016
Later among the works it cites.
Showing versus doing: Teaching by demonstration
Mark K Ho, Michael Littman, James MacGlashan, Fiery Cushman, and Joseph L Austerweil · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A reduction of imitation learning and structured prediction to no-regret online learning
Stéphane Ross, Geoffrey J. Gordon, and J. Andrew Bagnell · 2011
Cited alongside, same era.
Learning from demonstrations: is it worth estimating a reward function?
Bilal Piot, Matthieu Geist, and Olivier Pietquin · 2013
Cited alongside, same era.
Predicting when to laugh with structured classification
Bilal Piot, Olivier Pietquin, and Matthieu Geist · 2014
Cited alongside, same era.
Boosted and reward-regularized classification for apprenticeship learning
Bilal Piot, Matthieu Geist, and Olivier Pietquin · 2014
Cited alongside, same era.
Later among the works it cites.
Generative adversarial imitation learning
Jonathan Ho and Stefano Ermon · 2016
Later among the works it cites.
Learning to navigate in complex environments
Piotr Mirowski, Razvan Pascanu, Fabio Viola, Hubert Soyer, Andy Ballard, Andrea Banino, Misha Denil, Ross Goroshin, Laurent Sifre, Koray Kavukcuoglu, et al · 2017
Closest in time.
Third-person imitation learning
Bradly C Stadie, Pieter Abbeel, and Ilya Sutskever · 2017
Closest in time.