Fetching the paper…
Reading the bibliography…
Imitation learning enables robots to learn from demonstrations.
M. Bain and C. Sammut, “A framework for behavioural cloning.” in
1995
Earlier work this paper cites.
H. A. Simon,
1997
Earlier work this paper cites.
A. Y. Ng, S. J. Russell,
2000
Earlier work this paper cites.
C. L. Nehaniv, K. Dautenhahn,
2002
Earlier work this paper cites.
P. Abbeel and A. Y. Ng, “Apprenticeship learning via inverse reinforcement learning,” in
2004
Earlier work this paper cites.
S. Calinon, F. Guenter, and A. Billard, “On learning, representing, and generalizing a task in a humanoid robot,”
2007
Earlier work this paper cites.
B. D. Ziebart, A. L. Maas, J. A. Bagnell, and A. K. Dey, “Maximum entropy inverse reinforcement learning.” in
2008
Earlier work this paper cites.
H. Daumé, J. Langford, and D. Marcu, “Search-based structured prediction,”
2009
Earlier work this paper cites.
B. D. Argall, S. Chernova, M. Veloso, and B. Browning, “A survey of robot learning from demonstration,”
2009
Earlier work this paper cites.
C. Eppner, J. Sturm, M. Bennewitz, C. Stachniss, and W. Burgard, “Imitation learning with generalized task descriptions,” in
2009
Earlier work this paper cites.
S. Ross and D. Bagnell, “Efficient reductions for imitation learning,” in
2010
Earlier work this paper cites.
D. H. Grollman and A. Billard, “Donut as i do: Learning from failed demonstrations,” in
2011
Earlier work this paper cites.
S. Ross, G. Gordon, and D. Bagnell, “A reduction of imitation learning and structured prediction to no-regret online learning,” in
2011
Earlier work this paper cites.
B. Akgun, M. Cakmak, K. Jiang, and A. L. Thomaz, “Keyframe-based learning from demonstration,”
2012
Earlier work this paper cites.
E. Todorov, T. Erez, and Y. Tassa, “Mujoco: A physics engine for model-based control,” in
2012
Cited alongside, same era.
P. Englert, A. Paraschos, J. Peters, and M. P. Deisenroth, “Addressing the correspondence problem by model-based imitation learning,” in
2013
Cited alongside, same era.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,”
2014
Cited alongside, same era.
J. Schulman, S. Levine, P. Abbeel, M. Jordan, and P. Moritz, “Trust region policy optimization,” in
2015
Cited alongside, same era.
J. Ho and S. Ermon, “Generative adversarial imitation learning,” in
2016
Cited alongside, same era.
Y.-H. Wu, N. Charoenphakdee, H. Bao, V. Tangkaratt, and M. Sugiyama, “Imitation learning from imperfect demonstration,” in
2019
Later among the works it cites.
2019
Later among the works it cites.
F. Torabi, G. Warnell, and P. Stone, “Generative adversarial imitation from observation,”
2019
Later among the works it cites.
W. Sun, A. Vemula, B. Boots, and D. Bagnell, “Provably efficient imitation learning from observation alone,” in
2019
Later among the works it cites.
F. Liu, Z. Ling, T. Mu, and H. Su, “State alignment-based imitation learning,” in
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2016
Cited alongside, same era.
Y. Schroecker and C. L. Isbell, “State aware imitation learning,” in
2017
Cited alongside, same era.
C. Basu, Q. Yang, D. Hungerman, M. Singhal, and A. D. Dragan, “Do you want your autonomous car to drive like you?” in
2017
Cited alongside, same era.
J. Fu, K. Luo, and S. Levine, “Learning robust rewards with adverserial inverse reinforcement learning,” in
2018
Cited alongside, same era.
F. Torabi, G. Warnell, and P. Stone, “Behavioral cloning from observation,” in
2018
Cited alongside, same era.
F. Codevilla, M. Miiller, A. López, V. Koltun, and A. Dosovitskiy, “End-to-end driving via conditional imitation learning,” in
2018
Cited alongside, same era.
J. Zhang, Z. Ding, W. Li, and P. Ogunbona, “Importance weighted adversarial nets for partial domain adaptation,” in
2018
Cited alongside, same era.
2019
Later among the works it cites.
A. H. Qureshi, B. Boots, and M. C. Yip, “Adversarial imitation via variational inverse reinforcement learning,” in
2019
Later among the works it cites.
2020
Later among the works it cites.
D. P. Losey, K. Srinivasan, A. Mandlekar, A. Garg, and D. Sadigh, “Controlling assistive robots with learned latent actions,” in
2020
Later among the works it cites.
2020
Later among the works it cites.
V. Tangkaratt, B. Han, M. E. Khan, and M. Sugiyama, “Variational imitation learning with diverse-quality demonstrations,” in
2020
Later among the works it cites.
L. Ke, M. Barnes, W. Sun, G. Lee, S. Choudhury, and S. Srinivasa, “Imitation learning as
2020
Later among the works it cites.
M. Kwon, E. Biyik, A. Talati, K. Bhasin, D. P. Losey, and D. Sadigh, “When humans aren’t optimal: Robots that collaborate with risk-aware humans,” in
2020
Later among the works it cites.