Fetching the paper…
Reading the bibliography…
Simulation is a crucial tool for accelerating the development of autonomous vehicles.
1910
Earlier work this paper cites.
D. A. Pomerleau, “ALVINN: an autonomous land vehicle in a neural network,” in Advances in neural information processing systems 1 . San Francisco, CA, USA: Morgan Kaufmann Publishers Inc., Dec. 1989, pp. 305–313
1989
Earlier work this paper cites.
C. E. Garcia, D. M. Prett, and M. Morari, “Model predictive control: Theory and practice—a survey,” Automatica , vol. 25, no. 3, pp. 335–348, 1989
1989
Earlier work this paper cites.
D. Michie, M. Bain, and J. Hayes-Miches, “Cognitive models from subcognitive skills,” IEE control engineering series , vol. 44, pp. 71–99, 1990
1990
Earlier work this paper cites.
R. J. Williams, “Simple statistical gradient-following algorithms for connectionist reinforcement learning,” Machine learning , vol. 8, no. 3, pp. 229–256, 1992
1992
Earlier work this paper cites.
M. L. Littman, “Markov games as a framework for multi-agent reinforcement learning,” in Machine learning proceedings 1994 . Elsevier, 1994, pp. 157–163
1994
Earlier work this paper cites.
G. Tesauro and G. R. Galperin, “On-line policy improvement using monte-carlo search,” in Proceedings of the 9th International Conference on Neural Information Processing Systems , 1996, pp. 1068–1074
1996
Earlier work this paper cites.
A. Y. Ng and S. J. Russell, “Algorithms for inverse reinforcement learning,” in Proceedings of the Seventeenth International Conference on Machine Learning , ser. ICML ’00. San Francisco, CA, USA: Morgan Kaufmann Publishers Inc., June 2000, pp. 663–670
2000
Earlier work this paper cites.
P. Abbeel and A. Y. Ng, “Apprenticeship learning via inverse reinforcement learning,” in Twenty-first international conference on Machine learning - ICML ’04 . New York, New York, USA: ACM Press, 2004
2004
Earlier work this paper cites.
D. Ramachandran and E. Amir, “Bayesian inverse reinforcement learning.” in IJCAI , vol. 7, 2007, pp. 2586–2591
2007
Earlier work this paper cites.
B. D. Ziebart, A. L. Maas, J. A. Bagnell, and A. K. Dey, “Maximum entropy inverse reinforcement learning,” in AAAI , vol. 8, 2008, pp. 1433–1438
2008
Earlier work this paper cites.
B. D. Argall, S. Chernova, M. Veloso, and B. Browning, “A survey of robot learning from demonstration,” Rob. Auton. Syst. , vol. 57, no. 5, pp. 469–483, May 2009
2009
Earlier work this paper cites.
S. Ross, G. Gordon, and D. Bagnell, “A reduction of imitation learning and structured prediction to no-regret online learning,” in Proceedings of the fourteenth international conference on artificial intelligence and statistics , 2011, pp. 627–635
2011
Earlier work this paper cites.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial nets,” in Advances in neural information processing systems , 2014, pp. 2672–2680
2014
Earlier work this paper cites.
J. Ho and S. Ermon, “Generative adversarial imitation learning,” June 2016
2016
Earlier work this paper cites.
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, et al. , “Mastering the game of go with deep neural networks and tree search,” Nature , vol. 529, no. 7587, pp. 484–489, 2016
2016
Cited alongside, same era.
N. Baram, O. Anschel, and S. Mannor, “Model-based adversarial imitation learning,” arXiv preprint arXiv:1612. 02179 , 2016
2016
Cited alongside, same era.
A. Tamar, Y. Wu, G. Thomas, S. Levine, and P. Abbeel, “Value iteration networks,” Feb. 2016
2016
Cited alongside, same era.
2016
Cited alongside, same era.
2018
Later among the works it cites.
F. Behbahani, K. Shiarlis, X. Chen, V. Kurin, S. Kasewa, C. Stirbu, J. Gomes, S. Paul, F. A. Oliehoek, J. Messias, and S. Whiteson, “Learning from demonstration in the wild,” Nov. 2018
2018
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A. Hussein, M. M. Gaber, E. Elyan, and C. Jayne, “Imitation learning: A survey of learning methods,” ACM Comput. Surv. , vol. 50, no. 2, pp. 1–35, Apr. 2017
2017
Cited alongside, same era.
N. Baram, O. Anschel, I. Caspi, and S. Mannor, “End-to-end differentiable adversarial imitation learning,” in International Conference on Machine Learning . PMLR, 2017, pp. 390–399
2017
Cited alongside, same era.
N. Lee, W. Choi, P. Vernaza, C. B. Choy, P. H. Torr, and M. Chandraker, “Desire: Distant future prediction in dynamic scenes with interacting agents,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2017, pp. 336–345
2017
Cited alongside, same era.
M. Laskey, J. Lee, R. Fox, A. Dragan, and K. Goldberg, “Dart: Noise injection for robust imitation learning,” in Conference on robot learning . PMLR, 2017, pp. 143–156
2017
Cited alongside, same era.
2017
Cited alongside, same era.
D. Silver, J. Schrittwieser, K. Simonyan, I. Antonoglou, A. Huang, A. Guez, T. Hubert, L. Baker, M. Lai, A. Bolton, et al. , “Mastering the game of go without human knowledge,” Nature , vol. 550, no. 7676, pp. 354–359, 2017
2017
Cited alongside, same era.
S. Casas, W. Luo, and R. Urtasun, “Intentnet: Learning to predict intention from raw sensor data,” in Conference on Robot Learning . PMLR, 2018, pp. 947–956
2018
Cited alongside, same era.
L. Lee, E. Parisotto, D. S. Chaplot, E. Xing, and R. Salakhutdinov, “Gated path planning networks,” 2018
2018
Cited alongside, same era.
J. B. Hamrick, A. L. Friesen, F. Behbahani, A. Guez, F. Viola, S. Witherspoon, T. Anthony, L. Buesing, P. Veličković, and T. Weber, “On the role of planning in model-based deep reinforcement learning,” 2020
2020
Later among the works it cites.
J. Schrittwieser, I. Antonoglou, T. Hubert, K. Simonyan, L. Sifre, S. Schmitt, A. Guez, E. Lockhart, D. Hassabis, T. Graepel, and et al., “Mastering atari, go, chess and shogi by planning with a learned model,” Nature , vol. 588, no. 7839, p. 604–609, Dec 2020. [Online]. Available: http://dx.doi.org/10.1038/s41586-020-03051-4
2020
Later among the works it cites.
2020
Later among the works it cites.
2021
Later among the works it cites.
L. Chen, K. Lu, A. Rajeswaran, K. Lee, A. Grover, M. Laskin, P. Abbeel, A. Srinivas, and I. Mordatch, “Decision transformer: Reinforcement learning via sequence modeling,” 2021
2021
Later among the works it cites.
M. Janner, Q. Li, and S. Levine, “Reinforcement learning as one big sequence modeling problem,” 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
S. Suo, S. Regalado, S. Casas, and R. Urtasun, “Trafficsim: Learning to simulate realistic multi-agent behaviors,” 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.