Fetching the paper…
Reading the bibliography…
The success of monocular depth estimation relies on large and diverse training sets.
D. Hoiem, A. A. Efros, and M. Hebert, “Automatic photo pop-up,” ACM Transactions on Graphics , vol. 24, no. 3, 2005
2005
Earlier work this paper cites.
A. Saxena, M. Sun, and A. Y. Ng, “Make3D: Learning 3D scene structure from a single still image,” PAMI , vol. 31, no. 5, 2009
2009
Earlier work this paper cites.
R. Neuman, “Bolt 3D: a case study,” in Stereoscopic Displays and Applications XX , vol. 7237. SPIE, 2009
2009
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L. Li, K. Li, and F. Li, “ImageNet: A large-scale hierarchical image database,” in CVPR , 2009
2009
Earlier work this paper cites.
F. Devernay and P. A. Beardsley, “Stereoscopic cinema,” in Image and Geometry Processing for 3-D Cinematography . Springer, 2010
2010
Earlier work this paper cites.
A. Torralba and A. A. Efros, “Unbiased look at dataset bias,” in CVPR , 2011
2011
Earlier work this paper cites.
K. Khoshelham and S. O. Elberink, “Accuracy and resolution of Kinect depth data for indoor mapping applications,” Sensors , vol. 12, no. 2, 2012
2012
Earlier work this paper cites.
A. Geiger, P. Lenz, and R. Urtasun, “Are we ready for autonomous driving? The KITTI vision benchmark suite,” in CVPR , 2012
2012
Earlier work this paper cites.
N. Silberman, D. Hoiem, P. Kohli, and R. Fergus, “Indoor segmentation and support inference from RGBD images,” in ECCV , 2012
2012
Earlier work this paper cites.
D. J. Butler, J. Wulff, G. B. Stanley, and M. J. Black, “A naturalistic open source movie for optical flow evaluation,” in ECCV , 2012
2012
Earlier work this paper cites.
J. Sturm, N. Engelhard, F. Endres, W. Burgard, and D. Cremers, “A benchmark for the evaluation of RGB-D SLAM systems,” in IROS , 2012
2012
Earlier work this paper cites.
M. Hansard, S. Lee, O. Choi, and R. Horaud, Time-of-Flight Cameras: Principles, Methods and Applications . Springer, 2013
2013
Earlier work this paper cites.
T.-Y. Lin, M. Maire, S. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft COCO: Common objects in context,” in ECCV , 2014
2014
Earlier work this paper cites.
K. Karsch, C. Liu, and S. B. Kang, “Depth transfer: Depth extraction from video using non-parametric sampling,” PAMI , vol. 36, no. 11, 2014
2014
Earlier work this paper cites.
D. Eigen, C. Puhrsch, and R. Fergus, “Depth map prediction from a single image using a multi-scale deep network,” in NIPS , 2014
2014
Earlier work this paper cites.
P. Fankhauser, M. Blösch, D. Rodriguez, R. Kaestner, M. Hutter, and R. Siegwart, “Kinect v2 for mobile robot navigation: Evaluation and modeling,” in International Conference on Advanced Robotics , 2015
2015
Earlier work this paper cites.
F. Liu, C. Shen, and G. Lin, “Deep convolutional neural fields for depth estimation from a single image,” in CVPR , 2015
2015
Earlier work this paper cites.
M. Menze and A. Geiger, “Object scene flow for autonomous vehicles,” in CVPR , 2015
2015
Earlier work this paper cites.
S. Song, S. P. Lichtenberg, and J. Xiao, “SUN RGB-D: A RGB-D scene understanding benchmark suite,” in CVPR , 2015
2015
Earlier work this paper cites.
D. P. Kingma and J. L. Ba, “Adam: A method for stochastic optimization,” in ICLR , 2015
2015
Earlier work this paper cites.
D. Eigen and R. Fergus, “Predicting depth, surface normals and semantic labels with a common multi-scale convolutional architecture,” in ICCV , 2015
2015
Earlier work this paper cites.
A. Hertzmann, “Why do line drawings work? a realism hypothesis,” Perception , 2020
2015
Earlier work this paper cites.
R. Garg, B. V. Kumar, G. Carneiro, and I. Reid, “Unsupervised CNN for single view depth estimation: Geometry to the rescue,” in ECCV , 2016
2016
Cited alongside, same era.
I. Laina, C. Rupprecht, V. Belagiannis, F. Tombari, and N. Navab, “Deeper depth prediction with fully convolutional residual networks,” in 3DV , 2016
2016
Cited alongside, same era.
A. Roy and S. Todorovic, “Monocular depth estimation using neural regression forest,” in CVPR , 2016
2016
Cited alongside, same era.
W. Chen, Z. Fu, D. Yang, and J. Deng, “Single-image depth perception in the wild,” in NIPS , 2016
2016
Cited alongside, same era.
M. Cordts, M. Omran, S. Ramos, T. Rehfeld, M. Enzweiler, R. Benenson, U. Franke, S. Roth, and B. Schiele, “The Cityscapes dataset for semantic urban scene understanding,” in CVPR , 2016
2016
Cited alongside, same era.
R. Li, K. Xian, C. Shen, Z. Cao, H. Lu, and L. Hang, “Deep attention-based classification network for robust depth prediction,” in ACCV , 2018
2018
Later among the works it cites.
X. Guo, H. Li, S. Yi, J. Ren, and X. Wang, “Learning monocular depth by distilling cross-domain stereo networks,” in ECCV , 2018
2018
Later among the works it cites.
Y. Luo, J. Ren, M. Lin, J. Pang, W. Sun, H. Li, and L. Lin, “Single view stereo matching,” in CVPR , 2018
2018
Later among the works it cites.
H. Zhan, R. Garg, C. S. Weerasekera, K. Li, H. Agarwal, and I. D. Reid, “Unsupervised learning of monocular depth estimation and visual odometry with deep feature reconstruction,” in CVPR , 2018
2018
Later among the works it cites.
R. Mahjourian, M. Wicke, and A. Angelova, “Unsupervised learning of depth and ego-motion from monocular video using 3D geometric constraints,” in CVPR , 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Xie, R. B. Girshick, and A. Farhadi, “Deep3D: Fully automatic 2D-to-3D video conversion with deep convolutional neural networks,” in ECCV , 2016
2016
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in CVPR , 2016
2016
Cited alongside, same era.
F. Perazzi, J. Pont-Tuset, B. McWilliams, L. J. V. Gool, M. H. Gross, and A. Sorkine-Hornung, “A benchmark dataset and evaluation methodology for video object segmentation,” in CVPR , 2016
2016
Cited alongside, same era.
C. Godard, O. Mac Aodha, and G. J. Brostow, “Unsupervised monocular depth estimation with left-right consistency,” in CVPR , 2017
2017
Cited alongside, same era.
T. Zhou, M. Brown, N. Snavely, and D. G. Lowe, “Unsupervised learning of depth and ego-motion from video,” in CVPR , 2017
2017
Cited alongside, same era.
T. Schöps, J. L. Schönberger, S. Galliani, T. Sattler, K. Schindler, M. Pollefeys, and A. Geiger, “A multi-view stereo benchmark with high-resolution images and multi-camera videos,” in CVPR , 2017
2017
Cited alongside, same era.
B. Ummenhofer, H. Zhou, J. Uhrig, N. Mayer, E. Ilg, A. Dosovitskiy, and T. Brox, “DeMoN: Depth and motion network for learning monocular stereo,” in CVPR , 2017
2017
Cited alongside, same era.
2018
Later among the works it cites.
Y. Kim, H. Jung, D. Min, and K. Sohn, “Deep monocular depth estimation via integration of global and local predictions,” IEEE Transactions on Image Processing , vol. 27, no. 8, 2018
2018
Later among the works it cites.
K. Xian, C. Shen, Z. Cao, H. Lu, Y. Xiao, R. Li, and Z. Luo, “Monocular relative depth perception with web stereo data supervision,” in CVPR , 2018
2018
Later among the works it cites.
FFmpeg developers, “FFmpeg,” https://ffmpeg.org , 2018
2018
Later among the works it cites.
D. Sun, X. Yang, M.-Y. Liu, and J. Kautz, “PWC-Net: CNNs for optical flow using pyramid, warping, and cost volume,” in CVPR , 2018
2018
Later among the works it cites.
S. Rota Bulò, L. Porzi, and P. Kontschieder, “In-place activated batchnorm for memory-optimized training of DNNs,” in CVPR , 2018
2018
Later among the works it cites.
D. Mahajan, R. Girshick, V. Ramanathan, K. He, M. Paluri, Y. Li, A. Bharambe, and L. van der Maaten, “Exploring the limits of weakly supervised pretraining,” in ECCV , 2018
2018
Later among the works it cites.
B. Zhou, P. Krähenbühl, and V. Koltun, “Does computer vision matter for action?” Science Robotics , vol. 4, no. 30, 2019
2019
Closest in time.
C. Godard, O. Mac Aodha, M. Firman, and G. J. Brostow, “Digging into self-supervised monocular depth prediction,” in ICCV , 2019
2019
Closest in time.
V. Casser, S. Pirk, R. Mahjourian, and A. Angelova, “Unsupervised learning of depth and ego-motion: A structured approach,” in AAAI , 2019
2019
Closest in time.
C. Wang, O. Wang, F. Perazzi, and S. Lucey, “Web stereo video supervision for depth prediction from dynamic scenes,” in 3DV , 2019
2019
Closest in time.
Z. Li, T. Dekel, F. Cole, R. Tucker, N. Snavely, C. Liu, and W. T. Freeman, “Learning the depths of moving people by watching frozen people,” in CVPR , 2019
2019
Closest in time.
W. Chen, S. Qian, and J. Deng, “Learning single-image depth from videos using quality assessment networks,” in CVPR , 2019
2019
Closest in time.
2019
Closest in time.
A. Gordon, H. Li, R. Jonschkowski, and A. Angelova, “Depth from videos in the wild: Unsupervised monocular depth learning from unknown cameras,” in ICCV , 2019
2019
Closest in time.
J. M. Facil, B. Ummenhofer, H. Zhou, L. Montesano, T. Brox, and J. Civera, “CAM-Convs: Camera-aware multi-scale convolutions for single-view depth,” in CVPR , 2019
2019
Closest in time.