Fetching the paper…
Reading the bibliography…
Most automated driving systems comprise a diverse sensor set, including several cameras, Radars, and LiDARs, ensuring a complete 360\deg coverage in near and far regions.
Noise and the reality gap: The use of simulation in evolutionary robotics
Nick Jakobi, Phil Husbands, and Inman Harvey · 1995
Earlier work this paper cites.
Are we ready for autonomous driving? The KITTI vision benchmark suite
A. Geiger, P. Lenz, and R. Urtasun · 2012
Earlier work this paper cites.
Densebox: Unifying landmark localization with end to end object detection
Lichao Huang, Yi Yang, Yafeng Deng, and Yinan Yu · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
High-resolution lidar-based depth mapping using bilateral filter
Cristiano Premebida, Luis Garrote, Alireza Asvadi, A Pedro Ribeiro, and Urbano Nunes · 2016
Earlier work this paper cites.
CARLA: An Open Urban Driving Simulator
Alexey Dosovitskiy, German Ros, Felipe Codevilla, Antonio Lopez, and Vladlen Koltun · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Pyramid scene parsing network
Hengshuang Zhao, Jianping Shi, Xiaojuan Qi, Xiaogang Wang, and Jiaya Jia · 2017
Earlier work this paper cites.
Near-field depth estimation using monocular fisheye camera: A semi-supervised learning approach using sparse LiDAR data
Varun Ravi Kumar, Stefan Milz, Christian Witt, Martin Simon, et al · 2018
Earlier work this paper cites.
Orthographic feature transform for monocular 3d object detection
Thomas Roddick, Alex Kendall, and Roberto Cipolla · 2018
Earlier work this paper cites.
Real-time joint object detection and semantic segmentation network for automated driving
Ganesh Sistu, Isabelle Leang, and Senthil Yogamani · 2018
Earlier work this paper cites.
Deep layer aggregation
Fisher Yu, Dequan Wang, Evan Shelhamer, and Trevor Darrell · 2018
Earlier work this paper cites.
Argoverse: 3d tracking and forecasting with rich maps
Ming-Fang Chang, John Lambert, Patsorn Sangkloy, Jagjeet Singh, Slawomir Bak, Andrew Hartnett, De Wang, Peter Carr, Simon Lucey, Deva Ramanan, et al · 2019
Earlier work this paper cites.
Pointpillars: Fast encoders for object detection from point clouds
Alex H Lang, Sourabh Vora, Holger Caesar, Lubing Zhou, Jiong Yang, and Oscar Beijbom · 2019
Earlier work this paper cites.
Monocular Semantic Occupancy Grid Mapping with Convolutional Variational Encoder-Decoder Networks
Chenyang Lu, Marinus Jacobus Gerardus van de Molengraft, and Gijs Dubbelman · 2019
Earlier work this paper cites.
Cross-view Semantic Segmentation for Sensing Surroundings
Bowen Pan, Jiankai Sun, Ho Yin Tiga Leung, Alex Andonian, and Bolei Zhou · 2019
Earlier work this paper cites.
Motion and depth augmented semantic segmentation for autonomous navigation
Hazem Rashed, Ahmad El Sallab, Senthil Yogamani, and Mohamed ElHelw · 2019
Earlier work this paper cites.
Efficientnet: Rethinking model scaling for convolutional neural networks
Mingxing Tan and Quoc Le · 2019
Earlier work this paper cites.
Challenges in designing datasets and validation for autonomous driving
Michal Uricár, David Hurych, Pavel Krizek, and Senthil Yogamani · 2019
Earlier work this paper cites.
Desoiling dataset: Restoring soiled areas on automotive fisheye cameras
Michal Uricár, Jan Ulicny, Ganesh Sistu, Hazem Rashed, et al · 2019
Cited alongside, same era.
nuscenes: A multimodal dataset for autonomous driving
Holger Caesar, Varun Bankiti, Alex H Lang, Sourabh Vora, Venice Erin Liong, Qiang Xu, Anush Krishnan, Yu Pan, Giancarlo Baldan, and Oscar Beijbom · 2020
Cited alongside, same era.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2020
Cited alongside, same era.
Unrectdepthnet: Self-supervised monocular depth estimation using a generic framework for handling common camera distortion models
Varun Ravi Kumar, Senthil Yogamani, Markus Bach, Christian Witt, Stefan Milz, and Patrick Mäder · 2020
Cited alongside, same era.
Dynamic task weighting methods for multi-task networks in autonomous driving systems
Isabelle Leang, Ganesh Sistu, Fabian Bürger, Andrei Bursuc, et al · 2020
Enabling spatio-temporal aggregation in birds-eye-view vehicle estimation
Avishkar Saha, Oscar Mendez, Chris Russell, and Richard Bowden · 2021
Later among the works it cites.
Fcos3d: Fully convolutional one-stage monocular 3d object detection
Tai Wang, Xinge Zhu, Jiangmiao Pang, and Dahua Lin · 2021
Later among the works it cites.
Center-based 3d object detection and tracking
Tianwei Yin, Xingyi Zhou, and Philipp Krahenbuhl · 2021
Later among the works it cites.
ViT-BEVSeg: A hierarchical transformer network for monocular birds-eye-view segmentation
Pramit Dutta, Ganesh Sistu, Senthil Yogamani, Edgar Galván, et al · 2022
Later among the works it cites.
GitNet: Geometric Prior-based Transformation for Birds-Eye-View Segmentation
Shi Gong, Xiaoqing Ye, Xiao Tan, Jingdong Wang, Errui Ding, Yu Zhou, and Xiang Bai · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Lift, splat, shoot: Encoding images from arbitrary camera rigs by implicitly unprojecting to 3d
Jonah Philion and Sanja Fidler · 2020
Cited alongside, same era.
A sim2real deep learning approach for the transformation of images from multiple vehicle-mounted cameras to a semantically segmented image in bird’s eye view
Lennart Reiher, Bastian Lampe, and Lutz Eckstein · 2020
Cited alongside, same era.
Predicting Semantic Map Representations From Images Using Pyramid Occupancy Networks
Thomas Roddick and Roberto Cipolla · 2020
Cited alongside, same era.
Scalability in perception for autonomous driving: Waymo open dataset
Pei Sun, Henrik Kretzschmar, Xerxes Dotiwalla, Aurelien Chouard, Vijaysai Patnaik, Paul Tsui, James Guo, Yin Zhou, Yuning Chai, Benjamin Caine, et al · 2020
Cited alongside, same era.
Efficientdet: Scalable and efficient object detection
Mingxing Tan, Ruoming Pang, and Quoc V Le · 2020
Cited alongside, same era.
Weather and light level classification for autonomous driving: Dataset, baseline and active learning
Mahesh M Dhananjaya, Varun Ravi Kumar, and Senthil Yogamani · 2021
Cited alongside, same era.
FIERY: Future Instance Prediction in Bird’s-Eye View From Surround Monocular Cameras
Anthony Hu, Zak Murez, Nikhil Mohan, Sofía Dudas, Jeffrey Hawke, Vijay Badrinarayanan, Roberto Cipolla, and Alex Kendall · 2021
Cited alongside, same era.
Bird’s-Eye-View Panoptic Segmentation Using Monocular Frontal View Images
Nikhil Gosala and Abhinav Valada · 2022
Later among the works it cites.
A simple baseline for BEV perception without lidar
Adam W Harley, Zhaoyuan Fang, Jie Li, Rares Ambrus, and Katerina Fragkiadaki · 2022
Later among the works it cites.
Bevdet4d: Exploit temporal cues in multi-camera 3d object detection
Junjie Huang and Guan Huang · 2022
Later among the works it cites.
Detecting adversarial perturbations in multi-task perception
Marvin Klingner, Varun Ravi Kumar, Senthil Yogamani, Andreas Bär, et al · 2022
Later among the works it cites.
Bevdepth: Acquisition of reliable depth for multi-view 3d object detection
Yinhao Li, Zheng Ge, Guanyi Yu, Jinrong Yang, Zengran Wang, Yukang Shi, Jianjian Sun, and Zeming Li · 2022
Later among the works it cites.
Zhiqi Li, Wenhai Wang, Hongyang Li, Enze Xie, Chonghao Sima, Tong Lu, Qiao Yu, and Jifeng Dai · 2022
Later among the works it cites.
BEVFusion: Multi-Task Multi-Sensor Fusion with Unified Bird’s-Eye View Representation
Zhijian Liu, Haotian Tang, Alexander Amini, Xinyu Yang, Huizi Mao, Daniela Rus, and Song Han · 2022
Later among the works it cites.
Vision-centric bev perception: A survey
Yuexin Ma, Tai Wang, Xuyang Bai, Huitong Yang, Yuenan Hou, Yaming Wang, Yu Qiao, Ruigang Yang, Dinesh Manocha, and Xinge Zhu · 2022
Later among the works it cites.
LiMoSeg: Real-time Bird’s Eye View based LiDAR Motion Segmentation
Sambit Mohapatra, Mona Hodaei, Senthil Yogamani, Stefan Milz, et al · 2022
Later among the works it cites.
Translating images into maps
Avishkar Saha, Oscar Mendez, Chris Russell, and Richard Bowden · 2022
Later among the works it cites.
Detr3d: 3d object detection from multi-view images via 3d-to-2d queries
Yue Wang, Vitor Campagnolo Guizilini, Tianyuan Zhang, Yilun Wang, Hang Zhao, and Justin Solomon · 2022
Later among the works it cites.
M2̂BEV: Multi-Camera Joint 3D Detection and Segmentation with Unified Birds-Eye View Representation
Enze Xie, Zhiding Yu, Daquan Zhou, Jonah Philion, Anima Anandkumar, Sanja Fidler, Ping Luo, and Jose M Alvarez · 2022
Later among the works it cites.
Cross-view Transformers for real-time Map-view Semantic Segmentation
Brady Zhou and Philipp Krähenbühl · 2022
Later among the works it cites.
X-Align: Cross-Modal Cross-View Alignment for Bird’s-Eye-View Segmentation
Shubhankar Borse, Marvin Klingner, Varun Ravi Kumar, Hong Cai, et al · 2023
Closest in time.