Fetching the paper…
Reading the bibliography…
Imitation learning (IL) is a simple and powerful way to use high-quality human driving data, which can be collected at scale, to produce human-like behavior.
D. A. Pomerleau, “Alvinn: An autonomous land vehicle in a neural network,” Advances in neural information processing systems , vol. 1, 1988
1988
Earlier work this paper cites.
A. Y. Ng and S. J. Russell, “Algorithms for inverse reinforcement learning,” in Proceedings of 17th International Conference on Machine Learning, 2000 , 2000, pp. 663–670
2000
Earlier work this paper cites.
J. Frank, S. Mannor, and D. Precup, “Reinforcement learning in the presence of rare events,” in Proceedings of the 25th international conference on Machine learning , 2008, pp. 336–343
2008
Earlier work this paper cites.
S. Ross, G. Gordon, and D. Bagnell, “A reduction of imitation learning and structured prediction to no-regret online learning,” in Proceedings of the fourteenth international conference on artificial intelligence and statistics . JMLR Workshop and Conference Proceedings, 2011, pp. 627–635
2011
Earlier work this paper cites.
R. Rajamani, Vehicle dynamics and control . Springer Science & Business Media, 2011
2011
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski et al. , “Human-level control through deep reinforcement learning,” nature , vol. 518, no. 7540, pp. 529–533, 2015
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
J. Ho and S. Ermon, “Generative adversarial imitation learning,” Advances in neural information processing systems , vol. 29, 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
K. Ramamohanarao, H. Xie, L. Kulik, S. Karunasekera, E. Tanin, R. Zhang, and E. B. Khunayn, “Smarts: Scalable microscopic adaptive road traffic simulator,” ACM Transactions on Intelligent Systems and Technology (TIST) , vol. 8, no. 2, pp. 1–22, 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
N. Kalra and S. M. Paddock, Driving to Safety: How Many Miles of Driving Would It Take to Demonstrate Autonomous Vehicle Reliability? RAND Corporation, 2016
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
A. Dosovitskiy, G. Ros, F. Codevilla, A. Lopez, and V. Koltun, “Carla: An open urban driving simulator,” in Conference on robot learning . PMLR, 2017, pp. 1–16
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
D. Gandhi, L. Pinto, and A. Gupta, “Learning to fly by crashing,” in 2017 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2017, pp. 3948–3955
2017
Earlier work this paper cites.
F. Codevilla, M. Miiller, A. López, V. Koltun, and A. Dosovitskiy, “End-to-end driving via conditional imitation learning,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2018, pp. 1–9
2018
Earlier work this paper cites.
X. Liang, T. Wang, L. Yang, and E. Xing, “Cirl: Controllable imitative reinforcement learning for vision-based self-driving,” in Proceedings of the European conference on computer vision (ECCV) , 2018, pp. 584–599
2018
Earlier work this paper cites.
2018
Cited alongside, same era.
T. Hester, M. Vecerik, O. Pietquin, M. Lanctot, T. Schaul, B. Piot, D. Horgan, J. Quan, A. Sendonaris, I. Osband et al. , “Deep q-learning from demonstrations,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 32, no. 1, 2018
2018
Cited alongside, same era.
A. Rajeswaran, V. Kumar, A. Gupta, G. Vezzani, J. Schulman, E. Todorov, and S. Levine, “Learning complex dexterous manipulation with deep reinforcement learning and demonstrations,” in Robotics: Science and Systems , 2018
2018
Cited alongside, same era.
2018
Cited alongside, same era.
Z. Zhang, A. Liniger, D. Dai, F. Yu, and L. Van Gool, “End-to-end urban driving by imitating a reinforcement learning coach,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , October 2021, pp. 15 222–15 232
2021
Later among the works it cites.
2021
Later among the works it cites.
S. Fujimoto and S. S. Gu, “A minimalist approach to offline reinforcement learning,” Advances in neural information processing systems , vol. 34, pp. 20 132–20 145, 2021
2021
Later among the works it cites.
Z. Zhang, A. Liniger, D. Dai, F. Yu, and L. Van Gool, “End-to-end urban driving by imitating a reinforcement learning coach,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2021, pp. 15 222–15 232
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
P. Wang, C.-Y. Chan, and A. de La Fortelle, “A reinforcement learning based approach for automated lane change maneuvers,” in 2018 IEEE Intelligent Vehicles Symposium (IV) . IEEE, 2018, pp. 1379–1384
2018
Cited alongside, same era.
E. Leurent, “An environment for autonomous driving decision-making,” https://github.com/eleurent/highway-env , 2018
2018
Cited alongside, same era.
S. Fujimoto, H. Hoof, and D. Meger, “Addressing function approximation error in actor-critic methods,” in International conference on machine learning . PMLR, 2018, pp. 1587–1596
2018
Cited alongside, same era.
S. Paul, K. Chatzilygeroudis, K. Ciosek, J.-B. Mouret, M. Osborne, and S. Whiteson, “Alternating optimisation and quadrature for robust control,” in AAAI Conference on Artificial Intelligence , 2018
2018
Cited alongside, same era.
L. Espeholt, H. Soyer, R. Munos, K. Simonyan, V. Mnih, T. Ward, Y. Doron, V. Firoiu, T. Harley, I. Dunning et al. , “Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures,” in International conference on machine learning . PMLR, 2018, pp. 1407–1416
2018
Cited alongside, same era.
P. De Haan, D. Jayaraman, and S. Levine, “Causal confusion in imitation learning,” Advances in Neural Information Processing Systems , vol. 32, 2019
2019
Cited alongside, same era.
2019
Cited alongside, same era.
2019
Cited alongside, same era.
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
A. Jaegle, F. Gimeno, A. Brock, O. Vinyals, A. Zisserman, and J. Carreira, “Perceiver: General perception with iterative attention,” in International conference on machine learning . PMLR, 2021, pp. 4651–4664
2021
Later among the works it cites.
2022
Closest in time.
E. Bronstein, S. Srinivasan, S. Paul, A. Sinha, M. O’Kelly, P. Nikdel, and S. Whiteson, “Embedding synthetic off-policy experience for autonomous driving via zero-shot curricula,” in 6th Annual Conference on Robot Learning , 2022
2022
Closest in time.
M. Vitelli, Y. Chang, Y. Ye, A. Ferreira, M. Wołczyk, B. Osiński, M. Niendorf, H. Grimmett, Q. Huang, and A. Jain, “Safetynet: Safe planning for real-world self-driving vehicles using machine-learned policies,” in 2022 International Conference on Robotics and Automation (ICRA) . IEEE, 2022, pp. 897–904
2022
Closest in time.
2022
Closest in time.
V. Lioutas, A. Scibior, and F. Wood, “Titrated: Learned human driving behavior without infractions via amortized inference,” Transactions on Machine Learning Research , 2022
2022
Closest in time.
2022
Closest in time.
Q. Li, Z. Peng, L. Feng, Q. Zhang, Z. Xue, and B. Zhou, “Metadrive: Composing diverse driving scenarios for generalizable reinforcement learning,” IEEE transactions on pattern analysis and machine intelligence , 2022
2022
Closest in time.
2022
Closest in time.
2022
Closest in time.
D. Isele, R. Rahimi, A. Cosgun, K. Subramanian, and K. Fujimura, “Navigating occluded intersections with autonomous vehicles using deep reinforcement learning,” in 2018 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2018, pp. 2034–2039
2039
Closest in time.