Fetching the paper…
Reading the bibliography…
Decision-making for urban autonomous driving is challenging due to the stochastic nature of interactive traffic participants and the complexity of road structures.
N. Tishby, F. C. Pereira, and W. Bialek, “The information bottleneck method,” arXiv preprint physics/0004057 , 2000
2000
Earlier work this paper cites.
P. Bender, J. Ziegler, and C. Stiller, “Lanelets: Efficient map representation for autonomous driving,” in 2014 IEEE Intelligent Vehicles Symposium Proceedings . IEEE, 2014, pp. 420–425
2014
Earlier work this paper cites.
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
P. Wolf, C. Hubschneider, M. Weber, A. Bauer, J. Härtl, F. Dürr, and J. M. Zöllner, “Learning how to drive in a real world simulation with deep q-networks,” in 2017 IEEE Intelligent Vehicles Symposium (IV) . IEEE, 2017, pp. 244–250
2017
Earlier work this paper cites.
A. Dosovitskiy, G. Ros, F. Codevilla, A. Lopez, and V. Koltun, “Carla: An open urban driving simulator,” in Conference on robot learning . PMLR, 2017, pp. 1–16
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
L. Zhu, F. R. Yu, Y. Wang, B. Ning, and T. Tang, “Big data analytics in intelligent transportation systems: A survey,” IEEE Transactions on Intelligent Transportation Systems , vol. 20, no. 1, pp. 383–398, 2018
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
D. Zhu, T. Li, D. Ho, C. Wang, and M. Q.-H. Meng, “Deep reinforcement learning supervised autonomous exploration in office environments,” in 2018 IEEE international conference on robotics and automation (ICRA) . IEEE, 2018, pp. 7548–7555
2018
Earlier work this paper cites.
J. Chen, B. Yuan, and M. Tomizuka, “Model-free deep reinforcement learning for urban autonomous driving,” in 2019 IEEE intelligent transportation systems conference (ITSC) . IEEE, 2019, pp. 2765–2771
2019
Earlier work this paper cites.
J. Chen, B. Yuan, and M. Tomizuka, “Deep imitation learning for autonomous driving in generic urban scenarios with enhanced safety,” in 2019 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2019, pp. 2884–2890
2019
Earlier work this paper cites.
Á. Fehér, S. Aradi, F. Hegedüs, T. Bécsi, and P. Gáspár, “Hybrid ddpg approach for vehicle motion planning,” 2019
2019
Earlier work this paper cites.
C. Gelada, S. Kumar, J. Buckman, O. Nachum, and M. G. Bellemare, “Deepmdp: Learning continuous latent space models for representation learning,” in International Conference on Machine Learning . PMLR, 2019, pp. 2170–2179
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
B. Poole, S. Ozair, A. Van Den Oord, A. Alemi, and G. Tucker, “On variational bounds of mutual information,” in International Conference on Machine Learning . PMLR, 2019, pp. 5171–5180
2019
Earlier work this paper cites.
E. Yurtsever, J. Lambert, A. Carballo, and K. Takeda, “A survey of autonomous driving: Common practices and emerging technologies,” IEEE access , vol. 8, pp. 58 443–58 469, 2020
2020
Earlier work this paper cites.
S. Aradi, “Survey of deep reinforcement learning for motion planning of autonomous vehicles,” IEEE Transactions on Intelligent Transportation Systems , 2020
2020
Cited alongside, same era.
Z. Huang, C. Lv, Y. Xing, and J. Wu, “Multi-modal sensor fusion-based deep neural network for end-to-end autonomous driving with scene understanding,” IEEE Sensors Journal , vol. 21, no. 10, pp. 11 781–11 790, 2020
2020
Cited alongside, same era.
J. Duan, S. E. Li, Y. Guan, Q. Sun, and B. Cheng, “Hierarchical reinforcement learning for self-driving decision-making without reliance on labelled driving data,” IET Intelligent Transport Systems , vol. 14, no. 5, pp. 297–305, 2020
2020
Cited alongside, same era.
J. Gao, C. Sun, H. Zhao, Y. Shen, D. Anguelov, C. Li, and C. Schmid, “Vectornet: Encoding hd maps and agent dynamics from vectorized representation,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 11 525–11 533
2020
Cited alongside, same era.
2021
Later among the works it cites.
2021
Later among the works it cites.
M. Janner, Q. Li, and S. Levine, “Reinforcement learning as one big sequence modeling problem,” in ICML 2021 Workshop on Unsupervised Reinforcement Learning , 2021
2021
Later among the works it cites.
L. Chen, K. Lu, A. Rajeswaran, K. Lee, A. Grover, M. Laskin, P. Abbeel, A. Srinivas, and I. Mordatch, “Decision transformer: Reinforcement learning via sequence modeling,” Advances in neural information processing systems , vol. 34, pp. 15 084–15 097, 2021
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
M. Hügle, G. Kalweit, M. Werling, and J. Boedecker, “Dynamic interaction-aware scene understanding for reinforcement learning in autonomous driving,” in 2020 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2020, pp. 4329–4335
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2020
Cited alongside, same era.
J.-B. Grill, F. Strub, F. Altché, C. Tallec, P. Richemond, E. Buchatskaya, C. Doersch, B. Avila Pires, Z. Guo, M. Gheshlaghi Azar et al. , “Bootstrap your own latent-a new approach to self-supervised learning,” Advances in Neural Information Processing Systems , vol. 33, pp. 21 271–21 284, 2020
2020
Cited alongside, same era.
D. Yarats, I. Kostrikov, and R. Fergus, “Image augmentation is all you need: Regularizing deep reinforcement learning from pixels,” in International Conference on Learning Representations , 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
T. Wang, Y. Luo, J. Liu, R. Chen, and K. Li, “End-to-end self-driving approach independent of irrelevant roadside objects with auto-encoder,” IEEE Transactions on Intelligent Transportation Systems , vol. 23, no. 1, pp. 641–650, 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
S. Casas, A. Sadat, and R. Urtasun, “Mp3: A unified model to map, perceive, predict and plan,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2021, pp. 14 403–14 412
2021
Later among the works it cites.
2021
Later among the works it cites.
J. Ngiam, B. Caine, V. Vasudevan, Z. Zhang, H.-T. L. Chiang, J. Ling, R. Roelofs, A. Bewley, C. Liu, A. Venugopal et al. , “Scene transformer: A unified multi-task model for behavior prediction and planning,” arXiv e-prints , pp. arXiv–2106, 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
Y. Liu, J. Zhang, L. Fang, Q. Jiang, and B. Zhou, “Multimodal motion prediction with stacked transformers,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2021, pp. 7577–7586
2021
Later among the works it cites.
K. Chen, L. Hong, H. Xu, Z. Li, and D.-Y. Yeung, “Multisiam: Self-supervised multi-instance siamese representation learning for autonomous driving,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2021, pp. 7546–7554
2021
Later among the works it cites.
X. Chen and K. He, “Exploring simple siamese representation learning,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2021, pp. 15 750–15 758
2021
Later among the works it cites.
2021
Later among the works it cites.
Z. Huang, J. Wu, and C. Lv, “Efficient deep reinforcement learning with imitative expert priors for autonomous driving,” IEEE Transactions on Neural Networks and Learning Systems , 2022
2022
Closest in time.
Ó. Pérez-Gil, R. Barea, E. López-Guillén, L. M. Bergasa, C. Gomez-Huelamo, R. Gutiérrez, and A. Diaz-Diaz, “Deep reinforcement learning based control for autonomous vehicles in carla,” Multimedia Tools and Applications , vol. 81, no. 3, pp. 3553–3576, 2022
2022
Closest in time.
T. Gilles, S. Sabatini, D. Tsishkou, B. Stanciulescu, and F. Moutarde, “Gohome: Graph-oriented heatmap output for future motion estimation,” in 2022 International Conference on Robotics and Automation (ICRA) . IEEE, 2022, pp. 9107–9114
2022
Closest in time.