Fetching the paper…
Reading the bibliography…
We demonstrate the first large-scale application of model-based generative adversarial imitation learning (MGAIL) to the task of dense urban self-driving.
D. A. Pomerleau, “ALVINN: an autonomous land vehicle in a neural network,” in Advances in neural information processing systems 1 , 1989
1989
Earlier work this paper cites.
D. Michie, M. Bain, and J. Hayes-Miches, “Cognitive models from subcognitive skills,” IEE control engineering series , vol. 44, pp. 71–99, 1990
1990
Earlier work this paper cites.
A. Y. Ng and S. Russell, “Algorithms for inverse reinforcement learning,” in in Proc. 17th International Conf. on Machine Learning . Morgan Kaufmann, 2000, pp. 663–670
2000
Earlier work this paper cites.
A. Goldberg and C. Harrelson, “Computing the shortest path: A* search meets graph theory,” Symposium on Discrete Algorithms , 2003
2003
Earlier work this paper cites.
P. Abbeel and A. Y. Ng, “Apprenticeship learning via inverse reinforcement learning,” in ICML , 2004
2004
Earlier work this paper cites.
N. D. Ratliff, J. A. Bagnell, and M. A. Zinkevich, “Maximum margin planning,” in Proceedings of the 23rd international conference on Machine learning , 2006, pp. 729–736
2006
Earlier work this paper cites.
B. D. Ziebart, A. L. Maas, J. A. Bagnell, and A. K. Dey, “Maximum entropy inverse reinforcement learning,” in AAAI , vol. 8, 2008, pp. 1433–1438
2008
Earlier work this paper cites.
B. D. Argall, S. Chernova, M. Veloso, and B. Browning, “A survey of robot learning from demonstration,” Rob. Auton. Syst. , 2009
2009
Earlier work this paper cites.
S. Ross et al. , “A reduction of imitation learning and structured prediction to no-regret online learning,” in AI Stats , 2011
2011
Earlier work this paper cites.
M. P. Deisenroth et al. , “A survey on policy search for robotics,” Foundations and trends in Robotics , 2013
2013
Earlier work this paper cites.
2014
Earlier work this paper cites.
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
C. Finn, S. Levine, and P. Abbeel, “Guided cost learning: Deep inverse optimal control via policy optimization,” in International Conference on Machine Learning , 2016, pp. 49–58
2016
Earlier work this paper cites.
J. Ho and S. Ermon, “Generative adversarial imitation learning,” in Advances in Neural Information Processing Systems , vol. 29, 2016
2016
Earlier work this paper cites.
B. Paden, M. Cap, S. Z. Yong, D. Yershov, and E. Frazzoli, “A survey of motion planning and control techniques for self-driving urban vehicles,” Apr. 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
D. Amodei, C. Olah, J. Steinhardt, P. Christiano, J. Schulman, and D. Mané, “Concrete problems in ai safety,” 2016
2016
Earlier work this paper cites.
N. Baram, O. Anschel, I. Caspi, and S. Mannor, “End-to-end differentiable adversarial imitation learning,” in International Conference on Machine Learning . PMLR, 2017, pp. 390–399
2017
Earlier work this paper cites.
A. Hussein, M. M. Gaber, E. Elyan, and C. Jayne, “Imitation learning: A survey of learning methods,” ACM Comput. Surv. , vol. 50, no. 2, apr 2017. [Online]. Available: https://doi.org/10.1145/3054912
2017
Earlier work this paper cites.
J. Fu, K. Luo, and S. Levine, “Learning robust rewards with adversarial inverse reinforcement learning,” Oct. 2017
2017
Cited alongside, same era.
Y. Li, J. Song, and S. Ermon, “Infogail: Interpretable imitation learning from visual demonstrations,” in Advances in Neural Information Processing Systems , 2017, pp. 3812–3822
2017
Cited alongside, same era.
A. Dosovitskiy, G. Ros, F. Codevilla, A. Lopez, and V. Koltun, “CARLA: An open urban driving simulator,” arXiv preprint arXiv:1711. 03938 , 2017
2017
Cited alongside, same era.
X. Pan, Y. You, Z. Wang, and C. Lu, “Virtual to real reinforcement learning for autonomous driving,” 2017
2017
Cited alongside, same era.
A. Vaswani et al. , “Attention is all you need,” Advances in neural information processing systems , 2017
2017
Cited alongside, same era.
C. R. Garrett, R. Chitnis, R. Holladay, B. Kim, T. Silver, L. P. Kaelbling, and T. Lozano-Pérez, “Integrated task and motion planning,” 2020
2020
Later among the works it cites.
A. Mandlekar et al. , “Iris: Implicit reinforcement without interaction at scale for learning control from offline robot manipulation data,” in 2020 IEEE International Conference on Robotics and Automation (ICRA) , 2020, pp. 4414–4420
2020
Later among the works it cites.
2020
Later among the works it cites.
J. Mercat et al. , “Multi-head attention for multi-modal joint vehicle motion forecasting,” in ICRA , 2020
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
F. Behbahani, K. Shiarlis, X. Chen, V. Kurin, S. Kasewa, C. Stirbu, J. Gomes, S. Paul, F. A. Oliehoek, J. Messias, and S. Whiteson, “Learning from demonstration in the wild,” Nov. 2018
2018
Cited alongside, same era.
F. Codevilla, M. Müller, A. López, V. Koltun, and A. Dosovitskiy, “End-to-end driving via conditional imitation learning,” 2018
2018
Cited alongside, same era.
A. Kendall, J. Hawke, D. Janz, P. Mazur, D. Reda, J.-M. Allen, V.-D. Lam, A. Bewley, and A. Shah, “Learning to drive in a day,” July 2018
2018
Cited alongside, same era.
M. Müller, A. Dosovitskiy, B. Ghanem, and V. Koltun, “Driving policy transfer via modularity and abstraction,” 2018
2018
Cited alongside, same era.
2018
Cited alongside, same era.
Y. Ding, C. Florensa, M. Phielipp, and P. Abbeel, “Goal-conditioned imitation learning,” June 2019
2019
Cited alongside, same era.
2021
Later among the works it cites.
G. Swamy, S. Choudhury, J. A. Bagnell, and Z. S. Wu, “Of moments and matching: A game-theoretic framework for closing the imitation gap,” 2021
2021
Later among the works it cites.
A. Jain, L. D. Pero, H. Grimmett, and P. Ondruska, “Autonomy 2.0: Why is self-driving always 5 years away?” 2021
2021
Later among the works it cites.
P. A. Ortega et al. , “Shaking the foundations: delusions in sequence models for interaction and control,” 2021
2021
Later among the works it cites.
M. Vitelli et al. , “Safetynet: Safe planning for real-world self-driving vehicles using machine-learned policies,” 2021
2021
Later among the works it cites.
B. Varadarajan et al. , “Multipath++: Efficient information fusion and trajectory aggregation for behavior prediction,” 2021
2021
Later among the works it cites.
J. Liu, W. Zeng, R. Urtasun, and E. Yumer, “Deep structured reactive planning,” 2021
2021
Later among the works it cites.
S. Casas, A. Sadat, and R. Urtasun, “Mp3: A unified model to map, perceive, predict and plan,” 2021
2021
Later among the works it cites.
S. Ettinger et al. , “Large scale interactive motion forecasting for autonomous driving : The waymo open motion dataset,” 2021
2021
Later among the works it cites.
W. Zeng, W. Luo, S. Suo, A. Sadat, B. Yang, S. Casas, and R. Urtasun, “End-to-end interpretable neural motion planner,” 2021
2021
Later among the works it cites.
J. Ngiam et al. , “Scene transformer: A unified architecture for predicting multiple agent trajectories,” 2021
2021
Later among the works it cites.
S. Suo, S. Regalado, S. Casas, and R. Urtasun, “Trafficsim: Learning to simulate realistic multi-agent behaviors,” 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
Y. Liu, J. Zhang, L. Fang, Q. Jiang, and B. Zhou, “Multimodal motion prediction with stacked transformers,” 2021
2021
Later among the works it cites.
A. Jaegle et al. , “Perceiver: General perception with iterative attention,” in ICML , 2021
2021
Later among the works it cites.
M. Igl, D. Kim, A. Kuefler, P. Mougin, P. Shah, K. Shiarlis, D. Anguelov, M. Palatucci, B. White, and S. Whiteson, “Symphony: Learning realistic and diverse agents for autonomous driving simulation,” in Robotics and Automation (ICRA), 2022 IEEE International Conference on , 2022
2022
Closest in time.