Fetching the paper…
Reading the bibliography…
Interactions with either environments or expert policies during training are needed for most of the current imitation learning (IL) algorithms.
L. S. Pontryagin, E. Mishchenko, V. Boltyanskii, and R. Gamkrelidze, “The mathematical theory of optimal processes,” 1962
1962
Earlier work this paper cites.
J.-J. Slotine and S. S. Sastry, “Tracking control of non-linear systems using sliding surfaces, with application to robot manipulators,” International journal of control , vol. 38, no. 2, pp. 465–492, 1983
1983
Earlier work this paper cites.
D. A. Pomerleau, “Alvinn: An autonomous land vehicle in a neural network,” in Advances in neural information processing systems , 1989, pp. 305–313
1989
Earlier work this paper cites.
J.-J. E. Slotine et al. , Applied nonlinear control , 1991, vol. 199, no. 1
1991
Earlier work this paper cites.
S. A. Snell, D. F. Nns, and W. L. Arrard, “Nonlinear inversion flight control for a supermaneuverable aircraft,” Journal of guidance, control, and dynamics , vol. 15, no. 4, pp. 976–984, 1992
1992
Earlier work this paper cites.
D. Enns, D. Bugajski, R. Hendrick, and G. Stein, “Dynamic inversion: an evolving methodology for flight control design,” International Journal of control , vol. 59, no. 1, pp. 71–91, 1994
1994
Earlier work this paper cites.
K. Zhou, Essentials of robust control , 1998, vol. 104
1998
Earlier work this paper cites.
S. Schaal, “Is imitation learning the route to humanoid robots?” Trends in cognitive sciences , vol. 3, no. 6, pp. 233–242, 1999
1999
Earlier work this paper cites.
A. Y. Ng, S. J. Russell et al. , “Algorithms for inverse reinforcement learning.” 2000
2000
Earlier work this paper cites.
P. Abbeel and A. Y. Ng, “Apprenticeship learning via inverse reinforcement learning,” in Proceedings of the twenty-first international conference on Machine learning . ACM, 2004, p. 1
2004
Earlier work this paper cites.
S. Sieberling, Q. Chu, and J. Mulder, “Robust flight control using incremental nonlinear dynamic inversion and angular acceleration prediction,” Journal of guidance, control, and dynamics , vol. 33, no. 6, pp. 1732–1742, 2010
2010
Earlier work this paper cites.
S. Ross, G. Gordon, and D. Bagnell, “A reduction of imitation learning and structured prediction to no-regret online learning,” in Proceedings of the fourteenth international conference on artificial intelligence and statistics , 2011, pp. 627–635
2011
Earlier work this paper cites.
E. Todorov, T. Erez, and Y. Tassa, “Mujoco: A physics engine for model-based control,” in 2012 IEEE/RSJ International Conference on Intelligent Robots and Systems . IEEE, 2012, pp. 5026–5033
2012
Earlier work this paper cites.
2013
Cited alongside, same era.
K. Sohn, H. Lee, and X. Yan, “Learning structured output representation using deep conditional generative models,” in Advances in neural information processing systems , 2015, pp. 3483–3491
2015
Cited alongside, same era.
J. Schulman, S. Levine, P. Abbeel, M. Jordan, and P. Moritz, “Trust region policy optimization,” in International conference on machine learning , 2015, pp. 1889–1897
2015
Cited alongside, same era.
J. Ho and S. Ermon, “Generative adversarial imitation learning,” in Advances in neural information processing systems , 2016, pp. 4565–4573
2016
Cited alongside, same era.
2018
Later among the works it cites.
T. Osa, J. Pajarinen, G. Neumann, J. A. Bagnell, P. Abbeel, J. Peters et al. , “An algorithmic perspective on imitation learning,” Foundations and Trends® in Robotics , vol. 7, no. 1-2, pp. 1–179, 2018
2018
Later among the works it cites.
T. Q. Chen, Y. Rubanova, J. Bettencourt, and D. K. Duvenaud, “Neural ordinary differential equations,” in Advances in neural information processing systems , 2018, pp. 6571–6583
2018
Later among the works it cites.
P. Felsen, P. Lucey, and S. Ganguly, “Where will they go? predicting fine-grained adversarial multi-agent motion using conditional variational autoencoders,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2018, pp. 732–747
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2016
Cited alongside, same era.
J. Walker, C. Doersch, A. Gupta, and M. Hebert, “An uncertain future: Forecasting from static images using variational autoencoders,” in European Conference on Computer Vision . Springer, 2016, pp. 835–851
2016
Cited alongside, same era.
2016
Cited alongside, same era.
2017
Cited alongside, same era.
M. Laskey, J. Lee, R. Fox, A. Dragan, and K. Goldberg, “Dart: Noise injection for robust imitation learning,” in Conference on robot learning . PMLR, 2017, pp. 143–156
2017
Cited alongside, same era.
2017
Cited alongside, same era.
Z. Wang, J. S. Merel, S. E. Reed, N. de Freitas, G. Wayne, and N. Heess, “Robust imitation of diverse behaviors,” in Advances in Neural Information Processing Systems , 2017, pp. 5320–5329
2017
Cited alongside, same era.
2018
Cited alongside, same era.
A. Nagabandi, G. Kahn, R. S. Fearing, and S. Levine, “Neural network dynamics for model-based deep reinforcement learning with model-free fine-tuning,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2018, pp. 7559–7566
2018
Later among the works it cites.
2018
Later among the works it cites.
B. Amos, I. Jimenez, J. Sacks, B. Boots, and J. Z. Kolter, “Differentiable mpc for end-to-end planning and control,” in Advances in Neural Information Processing Systems , 2018, pp. 8289–8300
2018
Later among the works it cites.
A. Hill, A. Raffin, M. Ernestus, A. Gleave, A. Kanervisto, R. Traore, P. Dhariwal, C. Hesse, O. Klimov, A. Nichol, M. Plappert, A. Radford, J. Schulman, S. Sidor, and Y. Wu, “Stable baselines,” https://github.com/hill-a/stable-baselines , 2018
2018
Later among the works it cites.
Y. Ding, C. Florensa, P. Abbeel, and M. Phielipp, “Goal-conditioned imitation learning,” in Advances in Neural Information Processing Systems , 2019, pp. 15 298–15 309
2019
Later among the works it cites.
2019
Later among the works it cites.
Y. Rubanova, T. Q. Chen, and D. K. Duvenaud, “Latent ordinary differential equations for irregularly-sampled time series,” in Advances in Neural Information Processing Systems , 2019, pp. 5321–5331
2019
Later among the works it cites.