Fetching the paper…
Reading the bibliography…
Building a multi-modality multi-task neural network toward accurate and robust performance is a de-facto standard in perception task of autonomous driving.
D. A. Pomerleau, “Alvinn: An autonomous land vehicle in a neural network,” in NeurIPS
1988
Earlier work this paper cites.
2016
Earlier work this paper cites.
Y. Zhou and O. Tuzel, “Voxelnet: End-to-end learning for point cloud based 3d object detection,” in Proceedings of the IEEE conference on computer vision and pattern recognition
2018
Earlier work this paper cites.
A. H. Lang, S. Vora, H. Caesar, L. Zhou, J. Yang, and O. Beijbom, “Pointpillars: Fast encoders for object detection from point clouds,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition
2019
Earlier work this paper cites.
A. Kendall, J. Hawke, D. Janz, P. Mazur, D. Reda, J.-M. Allen, V.-D. Lam, A. Bewley, and A. Shah, “Learning to drive in a day,” in ICRA
2019
Earlier work this paper cites.
H. Caesar, V. Bankiti, A. H. Lang, S. Vora, V. E. Liong, Q. Xu, A. Krishnan, Y. Pan, G. Baldan, and O. Beijbom, “nuscenes: A multimodal dataset for autonomous driving,” in CVPR
2020
Earlier work this paper cites.
J. Philion and S. Fidler, “Lift, splat, shoot: Encoding images from arbitrary camera rigs by implicitly unprojecting to 3d,” in ECCV
2020
Earlier work this paper cites.
J. Gao, C. Sun, H. Zhao, Y. Shen, D. Anguelov, C. Li, and C. Schmid, “Vectornet: Encoding hd maps and agent dynamics from vectorized representation,” in CVPR
2020
Earlier work this paper cites.
M. Liang, B. Yang, R. Hu, Y. Chen, R. Liao, S. Feng, and R. Urtasun, “Learning lane graph representations for motion forecasting,” in ECCV
2020
Earlier work this paper cites.
H. Zhao, J. Gao, T. Lan, C. Sun, B. Sapp, B. Varadarajan, Y. Shen, Y. Shen, Y. Chai, C. Schmid, C. Li, and D. Anguelov, “TNT: Target-driven trajectory prediction,” in CoRL
2020
Earlier work this paper cites.
M. Liang, B. Yang, W. Zeng, Y. Chen, R. Hu, S. Casas, and R. Urtasun, “Pnpnet: End-to-end perception and prediction with tracking in the loop,” in CVPR
2020
Earlier work this paper cites.
N. Carion, F. Massa, G. Synnaeve, N. Usunier, A. Kirillov, and S. Zagoruyko, “End-to-end object detection with transformers,” in ECCV
2020
Earlier work this paper cites.
J. Chen, S. E. Li, and M. Tomizuka, “Interpretable end-to-end urban autonomous driving with latent deep reinforcement learning,” IEEE Transactions on Intelligent Transportation Systems
2020
Earlier work this paper cites.
A. Sadat, S. Casas, M. Ren, X. Wu, P. Dhawan, and R. Urtasun, “Perceive, predict, and plan: Safe motion planning through interpretable semantic representations,” in ECCV
2020
Earlier work this paper cites.
S. Casas, A. Sadat, and R. Urtasun, “Mp3: A unified model to map, perceive, predict and plan,” in CVPR
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
T. Yin, X. Zhou, and P. Krähenbühl, “Multimodal virtual point 3d detection,” Advances in Neural Information Processing Systems
2021
Cited alongside, same era.
T. Yin, X. Zhou, and P. Krahenbuhl, “Center-based 3d object detection and tracking,” in CVPR
2021
Cited alongside, same era.
J. Gu, C. Sun, and H. Zhao, “Densetnt: End-to-end trajectory prediction from dense goal sets,” in ICCV
2021
Cited alongside, same era.
F. Zeng, B. Dong, T. Wang, X. Zhang, and Y. Wei, “Motr: End-to-end multiple-object tracking with transformer,” in ECCV
2021
Cited alongside, same era.
A. Prakash, K. Chitta, and A. Geiger, “Multi-modal fusion transformer for end-to-end autonomous driving,” in CVPR
2021
Cited alongside, same era.
2022
Later among the works it cites.
F. Da and Y. Zhang, “Path-aware graph attention for hd maps in motion prediction,” in 2022 International Conference on Robotics and Automation (ICRA)
2022
Later among the works it cites.
L. Gao, Z. Gu, C. Qiu, L. Lei, S. E. Li, S. Zheng, W. Jing, and J. Chen, “Cola-hrl: Continuous-lattice hierarchical reinforcement learning for autonomous driving,” in IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)
2022
Later among the works it cites.
O. Scheel, L. Bergamini, M. Wolczyk, B. Osiński, and P. Ondruska, “Urban Driver: Learning to drive from real-world demonstrations using policy gradients,” in Conference on Robot Learning (CoRL)
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. Suo, S. Regalado, S. Casas, and R. Urtasun, “Trafficsim: Learning to simulate realistic multi-agent behaviors,” 2021 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
2021
Cited alongside, same era.
A. Hu, Z. Murez, N. Mohan, S. Dudas, J. Hawke, V. Badrinarayanan, R. Cipolla, and A. Kendall, “FIERY: Future instance prediction in bird’s-eye view from surround monocular cameras,” in ICCV
2021
Cited alongside, same era.
2021
Cited alongside, same era.
P. Hu, A. Huang, J. Dolan, D. Held, and D. Ramanan, “Safe local motion planning with self-supervised freespace forecasting,” in CVPR
2021
Cited alongside, same era.
Y. Hu, J. Yang, L. Chen, K. Li, C. Sima, X. Zhu, S. Chai, S. Du, T. Lin, W. Wang, L. Lu, X. Jia, Q. Liu, J. Dai, Y. Qiao, and H. Li, “Planning-oriented autonomous driving,” 2022
2022
Cited alongside, same era.
X. Bai, Z. Hu, X. Zhu, Q. Huang, Y. Chen, H. Fu, and C.-L. Tai, “TransFusion: Robust lidar-camera fusion for 3d object detection with transformers,” in CVPR
2022
Cited alongside, same era.
T. Liang, H. Xie, K. Yu, Z. Xia, Z. Lin, Y. Wang, T. Tang, B. Wang, and Z. Tang, “BEVFusion: A simple and robust lidar-camera fusion framework,” in NeurIPS
2022
Cited alongside, same era.
A. K. Akan and F. Güney, “StretchBEV: Stretching future instance prediction spatially and temporally,” in ECCV
2022
Later among the works it cites.
S. Hu, L. Chen, P. Wu, H. Li, J. Yan, and D. Tao, “ST-P3: End-to-end vision-based autonomous driving via spatial-temporal feature learning,” in ECCV
2022
Later among the works it cites.
2022
Later among the works it cites.
T. Khurana, P. Hu, A. Dave, J. Ziglar, D. Held, and D. Ramanan, “Differentiable raycasting for self-supervised occupancy forecasting,” in ECCV
2022
Later among the works it cites.
Y. Hu, J. Yang, L. Chen, K. Li, C. Sima, X. Zhu, S. Chai, S. Du, T. Lin, W. Wang, L. Lu, X. Jia, Q. Liu, J. Dai, Y. Qiao, and H. Li, “Planning-oriented autonomous driving,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
2023
Closest in time.
2023
Closest in time.
Z. Liu, H. Tang, A. Amini, X. Yang, H. Mao, D. Rus, and S. Han, “BEVFusion: Multi-task multi-sensor fusion with unified bird’s-eye view representation,” in ICRA
2023
Closest in time.
J. Gu, C. Hu, T. Zhang, X. Chen, Y. Wang, Y. Wang, and H. Zhao, “ViP3D: End-to-end visual trajectory prediction via 3d agent queries,” in CVPR
2023
Closest in time.
K. Guo, W. Jing, J. Chen, and J. Pan, “CCIL: Context-conditioned imitation learning for urban driving,” in Robotics: Science and Systems (RSS)
2023
Closest in time.
C. Hu, “Fusionformer: Unified multi-modal and temporal fusion with transformer for 3d detection in bird’s-eye-view,” 2023
2023
Closest in time.