Fetching the paper…
Reading the bibliography…
Autonomous driving has attracted tremendous attention especially in the past few years.
R. E. Kalman et al. , “A new approach to linear filtering and prediction problems,” Journal of basic Engineering , vol. 82, no. 1, pp. 35–45, 1960
1960
Earlier work this paper cites.
A. Kar, S. Tulsiani, J. Carreira, and J. Malik, “Category-specific object reconstruction from a single image,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2015, pp. 1966–1974
1974
Earlier work this paper cites.
B. M. Haralick, C.-N. Lee, K. Ottenberg, and M. Nölle, “Review and analysis of solutions of the three point perspective pose estimation problem,” IJCV , vol. 13, no. 3, pp. 331–356, 1994
1994
Earlier work this paper cites.
P. David, D. Dementhon, R. Duraiswami, and H. Samet, “Softposit: Simultaneous pose and correspondence determination,” IJCV , vol. 59, no. 3, pp. 259–284, 2004
2004
Earlier work this paper cites.
F. Moreno-Noguer, V. Lepetit, and P. Fua, “Pose priors for simultaneously solving alignment and correspondence,” European Conference on Computer Vision , pp. 405–418, 2008
2008
Earlier work this paper cites.
G. J. Brostow, J. Fauqueur, and R. Cipolla, “Semantic object classes in video: A high-definition ground truth database,” Pattern Recognition Letters , vol. 30, no. 2, pp. 88–97, 2009
2009
Earlier work this paper cites.
M. Everingham, L. Van Gool, C. K. Williams, J. Winn, and A. Zisserman, “The pascal visual object classes (voc) challenge,” International journal of computer vision , vol. 88, no. 2, pp. 303–338, 2010
2010
Earlier work this paper cites.
N. Silberman, D. Hoiem, P. Kohli, and R. Fergus, “Indoor segmentation and support inference from rgbd images,” in European Conference on Computer Vision . Springer, 2012, pp. 746–760
2012
Earlier work this paper cites.
H. Su, J. Deng, and L. Fei-Fei, “Crowdsourcing annotations for visual object detection,” in Workshops at the Twenty-Sixth AAAI Conference on Artificial Intelligence , vol. 1, no. 2, 2012
2012
Earlier work this paper cites.
A. Geiger, P. Lenz, C. Stiller, and R. Urtasun, “Vision meets robotics: The kitti dataset,” International Journal of Robotics Research (IJRR) , 2013
2013
Earlier work this paper cites.
C. Hane, C. Zach, A. Cohen, R. Angst, and M. Pollefeys, “Joint 3d scene reconstruction and class segmentation,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2013, pp. 97–104
2013
Earlier work this paper cites.
Y. Xiang, R. Mottaghi, and S. Savarese, “Beyond pascal: A benchmark for 3d object detection in the wild,” in Applications of Computer Vision (WACV), 2014 IEEE Winter Conference on . IEEE, 2014, pp. 75–82
2014
Earlier work this paper cites.
A. Kundu, Y. Li, F. Dellaert, F. Li, and J. M. Rehg, “Joint semantic segmentation and 3d reconstruction from monocular video,” in European Conference on Computer Vision . Springer, 2014, pp. 703–718
2014
Earlier work this paper cites.
D. Scharstein, H. Hirschmüller, Y. Kitajima, G. Krathwohl, N. Nešić, X. Wang, and P. Westling, “High-resolution stereo datasets with subpixel-accurate ground truth,” in German Conference on Pattern Recognition . Springer, 2014, pp. 31–42
2014
Earlier work this paper cites.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft coco: Common objects in context,” in European conference on computer vision . Springer, 2014, pp. 740–755
2014
Earlier work this paper cites.
L. Kneip, H. Li, and Y. Seo, “Upnp: An optimal o (n) solution to the absolute pose problem with universal applicability,” in European Conference on Computer Vision . Springer, 2014, pp. 127–142
2014
Earlier work this paper cites.
J. Engel, T. Schöps, and D. Cremers, “Lsd-slam: Large-scale direct monocular slam,” in European Conference on Computer Vision . Springer, 2014, pp. 834–849
2014
Earlier work this paper cites.
S. Christoph Stein, M. Schoeler, J. Papon, and F. Worgotter, “Object partitioning using local convexity,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2014, pp. 304–311
2014
Earlier work this paper cites.
B. Hariharan, P. Arbeláez, R. Girshick, and J. Malik, “Simultaneous detection and segmentation,” in European Conference on Computer Vision . Springer, 2014, pp. 297–312
2014
Earlier work this paper cites.
A. Kendall, M. Grimes, and R. Cipolla, “Posenet: A convolutional network for real-time 6-dof camera relocalization,” in Proceedings of the IEEE international conference on computer vision , 2015, pp. 2938–2946
2015
Earlier work this paper cites.
J. Long, E. Shelhamer, and T. Darrell, “Fully convolutional networks for semantic segmentation,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2015, pp. 3431–3440
2015
Earlier work this paper cites.
F. Guney and A. Geiger, “Displets: Resolving stereo ambiguities using object knowledge,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2015, pp. 4165–4175
2015
Earlier work this paper cites.
K. Vishal, C. Jawahar, and V. Chari, “Accurate localization by fusing images and gps signals,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition Workshops , 2015, pp. 17–24
2015
Earlier work this paper cites.
W. Byeon, T. M. Breuel, F. Raue, and M. Liwicki, “Scene labeling with lstm recurrent neural networks,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2015, pp. 3547–3555
2015
Earlier work this paper cites.
A. Dosovitskiy, P. Fischer, E. Ilg, P. Hausser, C. Hazirbas, V. Golkov, P. Van Der Smagt, D. Cremers, and T. Brox, “Flownet: Learning optical flow with convolutional networks,” in Proceedings of the IEEE International Conference on Computer Vision , 2015, pp. 2758–2766
2015
Earlier work this paper cites.
S. Ren, K. He, R. Girshick, and J. Sun, “Faster r-cnn: Towards real-time object detection with region proposal networks,” in Advances in neural information processing systems , 2015, pp. 91–99
2015
Earlier work this paper cites.
P. Wang, X. Shen, Z. Lin, S. Cohen, B. Price, and A. L. Yuille, “Joint object and part segmentation using deep learned potentials,” in Proceedings of the IEEE International Conference on Computer Vision , 2015, pp. 1573–1581
2015
Earlier work this paper cites.
D. Eigen and R. Fergus, “Predicting depth, surface normals and semantic labels with a common multi-scale convolutional architecture,” in Proceedings of the IEEE International Conference on Computer Vision , 2015, pp. 2650–2658
2015
Earlier work this paper cites.
B.-H. Lee, J.-H. Song, J.-H. Im, S.-H. Im, M.-B. Heo, and G.-I. Jee, “Gps/dr error estimation for autonomous vehicle localization,” Sensors , vol. 15, no. 8, pp. 20 779–20 798, 2015
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
R. Mur-Artal, J. M. M. Montiel, and J. D. Tardos, “Orb-slam: a versatile and accurate monocular slam system,” IEEE transactions on robotics , vol. 31, no. 5, pp. 1147–1163, 2015
2015
Cited alongside, same era.
M. Cordts, M. Omran, S. Ramos, T. Rehfeld, M. Enzweiler, R. Benenson, U. Franke, S. Roth, and B. Schiele, “The cityscapes dataset for semantic urban scene understanding,” in Proc. of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2016
2016
Cited alongside, same era.
G. Ros, L. Sellart, J. Materzynska, D. Vazquez, and A. M. Lopez, “The synthia dataset: A large collection of synthetic images for semantic segmentation of urban scenes,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 3234–3243
2016
Cited alongside, same era.
2016
B. Ummenhofer, H. Zhou, J. Uhrig, N. Mayer, E. Ilg, A. Dosovitskiy, and T. Brox, “Demon: Depth and motion network for learning monocular stereo,” in IEEE Conference on computer vision and pattern recognition (CVPR) , vol. 5, 2017, p. 6
2017
Later among the works it cites.
H. Zhao, J. Shi, X. Qi, X. Wang, and J. Jia, “Pyramid scene parsing network,” CVPR , 2017
2017
Later among the works it cites.
2017
Later among the works it cites.
R. Gadde, V. Jampani, and P. V. Gehler, “Semantic video cnns through representation warping,” Proceedings of the International Conference on Computer Vision (ICCV) , 2017
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
A. Arnab, S. Jayasumana, S. Zheng, and P. H. Torr, “Higher order conditional random fields in deep neural networks,” in European Conference on Computer Vision . Springer, 2016, pp. 524–540
2016
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 770–778
2016
Cited alongside, same era.
2016
Cited alongside, same era.
A. Kundu, V. Vineet, and V. Koltun, “Feature space optimization for semantic video segmentation,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2016, pp. 3168–3175
2016
Cited alongside, same era.
J. Revaud, P. Weinzaepfel, Z. Harchaoui, and C. Schmid, “Deepmatching: Hierarchical deformable dense matching,” IJCV , vol. 120, no. 3, pp. 300–323, 2016
2016
Cited alongside, same era.
J. Xie, M. Kiefel, M.-T. Sun, and A. Geiger, “Semantic instance annotation of street scenes by 3d to 2d label transfer,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2016, pp. 3688–3697
2016
Cited alongside, same era.
2016
Cited alongside, same era.
A. Kovashka, O. Russakovsky, L. Fei-Fei, K. Grauman et al. , “Crowdsourcing in computer vision,” Foundations and Trends® in Computer Graphics and Vision , vol. 10, no. 3, pp. 177–243, 2016
2016
Cited alongside, same era.
2017
Later among the works it cites.
K. Tateno, F. Tombari, I. Laina, and N. Navab, “Cnn-slam: Real-time dense monocular slam with learned depth prediction,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , vol. 2, 2017
2017
Later among the works it cites.
C. R. Qi, L. Yi, H. Su, and L. J. Guibas, “Pointnet++: Deep hierarchical feature learning on point sets in a metric space,” in Advances in Neural Information Processing Systems , 2017, pp. 5105–5114
2017
Later among the works it cites.
T. Pohlen, A. Hermans, M. Mathias, and B. Leibe, “Full-resolution residual networks for semantic segmentation in street scenes,” Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2017
2017
Later among the works it cites.
Velodyne Lidar, “HDL-64E,” http://velodynelidar.com/ , 2018, [Online; accessed 01-March-2018]
2018
Closest in time.
P.-H. Huang, K. Matzen, J. Kopf, N. Ahuja, and J.-B. Huang, “Deepmvs: Learning multi-view stereopsis,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018, pp. 2821–2830
2018
Closest in time.
Y. Yao, Z. Luo, S. Li, T. Fang, and L. Quan, “Mvsnet: Depth inference for unstructured multi-view stereo,” in The European Conference on Computer Vision (ECCV) , September 2018
2018
Closest in time.
X. Cheng, P. Wang, and R. Yang, “Depth estimation via affinity learned with convolutional spatial propagation network,” European Conference on Computer Vision , 2018
2018
Closest in time.
J. L. Schönberger, M. Pollefeys, A. Geiger, and T. Sattler, “Semantic visual localization,” ISPRS Journal of Photogrammetry and Remote Sensing (JPRS) , 2018
2018
Closest in time.
K. He, G. Gkioxari, P. Dollár, and R. Girshick, “Mask r-cnn,” IEEE transactions on pattern analysis and machine intelligence , 2018
2018
Closest in time.
A. Kundu, Y. Li, and J. M. Rehg, “3d-rcnn: Instance-level 3d object reconstruction via render-and-compare,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018, pp. 3559–3568
2018
Closest in time.
——, “Instance segmentation,” https://www.kaggle.com/c/cvpr-2018-autonomous-driving
2018
Closest in time.
2018
Closest in time.
T. Sattler, W. Maddern, C. Toft, A. Torii, L. Hammarstrand, E. Stenborg, D. Safari, M. Okutomi, M. Pollefeys, J. Sivic et al. , “Benchmarking 6dof outdoor visual localization in changing conditions,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , vol. 1, 2018
2018
Closest in time.
Y. Chen, W. Li, and L. Van Gool, “Road: Reality oriented adaptation for semantic segmentation of urban scenes,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018, pp. 7892–7901
2018
Closest in time.
K.-N. Lianos, J. L. Schönberger, M. Pollefeys, and T. Sattler, “Vso: Visual semantic odometry,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2018, pp. 234–250
2018
Closest in time.
2018
Closest in time.
P. Wang, R. Yang, B. Cao, W. Xu, and Y. Lin, “Dels-3d: Deep localization and segmentation with a 3d semantic map,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018, pp. 5860–5869
2018
Closest in time.
S. Liu, L. Qi, H. Qin, J. Shi, and J. Jia, “Path aggregation network for instance segmentation,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2018, pp. 8759–8768
2018
Closest in time.
Z. Yueqing, L. Zeming, and G. Yu, “Find tiny instance segmentation,” http://www.skicyyu.org/WAD/wad_final.pdf , 2018
2018
Closest in time.
Smart_Vision_SG, “Wad instance segmentation 2nd place,” https://github.com/Computational-Camera/Kaggle-CVPR-2018-WAD-Video-Segmentation-Challenge-Solution , 2018
2018
Closest in time.
SZU_N606, “Wad instance segmentation 3rd place,” https://github.com/wwoody827/cvpr-2018-autonomous-driving-autopilot-solution , 2018
2018
Closest in time.
2018
Closest in time.
X. Song, P. Wang, D. Zhou, R. Zhu, C. Guan, Y. Dai, H. Su, H. Li, and R. Yang, “Apollocar3d: A large 3d car instance understanding benchmark for autonomous driving,” CVPR , 2019
2019
Closest in time.
Y. Ma, X. Zhu, S. Zhang, R. Yang, W. Wang, and D. Manocha, “Trafficpredict: Trajectory prediction for heterogeneous traffic-agents,” AAAI , 2019
2019
Closest in time.