Fetching the paper…
Reading the bibliography…
Single-View depth estimation using the CNNs trained from unlabelled videos has shown significant promise.
M. A. Fischler and R. C. Bolles, “Random sample consensus: a paradigm for model fitting with applications to image analysis and automated cartography,” Communications of the ACM , 1981
1981
Earlier work this paper cites.
Z. Zhang, “Determining the epipolar geometry and its uncertainty: A review,” International Journal on Computer Vision (IJCV) , 1998
1998
Earlier work this paper cites.
E. Trucco and A. Verri, Introductory techniques for 3-D computer vision . Prentice Hall Englewood Cliffs, 1998, vol. 201
1998
Earlier work this paper cites.
A. Fusiello, E. Trucco, and A. Verri, “A compact algorithm for rectification of stereo pairs,” Machine Vision and Applications , 2000
2000
Earlier work this paper cites.
R. Hartley and A. Zisserman, Multiple view geometry in computer vision . Cambridge university press, 2003
2003
Earlier work this paper cites.
D. G. Lowe, “Distinctive image features from scale-invariant keypoints,” International Journal on Computer Vision (IJCV) , 2004
2004
Earlier work this paper cites.
D. Nistér, “An efficient solution to the five-point relative pose problem,” IEEE Transactions on Pattern Recognition and Machine Intelligence (TPAMI) , 2004
2004
Earlier work this paper cites.
A. Saxena, S. H. Chung, and A. Y. Ng, “Learning depth from single monocular images,” in Neural Information Processing Systems (NeurIPS) , 2006
2006
Earlier work this paper cites.
A. J. Davison, I. D. Reid, N. D. Molton, and O. Stasse, “Monoslam: Real-time single camera slam,” IEEE Transactions on Pattern Recognition and Machine Intelligence (TPAMI) , 2007
2007
Earlier work this paper cites.
H. Hirschmuller, “Stereo processing by semiglobal matching and mutual information,” IEEE Transactions on Pattern Recognition and Machine Intelligence (TPAMI) , 2007
2007
Earlier work this paper cites.
R. A. Newcombe, S. J. Lovegrove, and A. J. Davison, “Dtam: Dense tracking and mapping in real-time,” in IEEE International Conference on Computer Vision (ICCV) , 2011
2011
Earlier work this paper cites.
N. Silberman, D. Hoiem, P. Kohli, and R. Fergus, “Indoor segmentation and support inference from rgbd images,” in European Conference on Computer Vision (ECCV) , 2012
2012
Earlier work this paper cites.
A. Geiger, P. Lenz, C. Stiller, and R. Urtasun, “Vision meets Robotics: The kitti dataset,” IJRR , 2013
2013
Earlier work this paper cites.
J. Shotton, B. Glocker, C. Zach, S. Izadi, A. Criminisi, and A. Fitzgibbon, “Scene coordinate regression forests for camera relocalization in rgb-d images,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2013
2013
Earlier work this paper cites.
D. Eigen, C. Puhrsch, and R. Fergus, “Depth map prediction from a single image using a multi-scale deep network,” in Neural Information Processing Systems (NeurIPS) , 2014
2014
Earlier work this paper cites.
K. Karsch, C. Liu, and S. B. Kang, “Depth transfer: Depth extraction from video using non-parametric sampling,” IEEE Transactions on Pattern Recognition and Machine Intelligence (TPAMI) , 2014
2014
Earlier work this paper cites.
M. Liu, M. Salzmann, and X. He, “Discrete-continuous depth estimation from a single image,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2014
2014
Earlier work this paper cites.
L. Ladicky, J. Shi, and M. Pollefeys, “Pulling things out of perspective,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2014
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
R. Mur-Artal, J. M. M. Montiel, and J. D. Tardos, “ORB-SLAM: a versatile and accurate monocular slam system,” IEEE Transactions on Robotics (TRO) , 2015
2015
Earlier work this paper cites.
D. Eigen and R. Fergus, “Predicting depth, surface normals and semantic labels with a common multi-scale convolutional architecture,” in IEEE International Conference on Computer Vision (ICCV) , 2015
2015
Earlier work this paper cites.
M. Jaderberg, K. Simonyan, A. Zisserman et al. , “Spatial transformer networks,” in Neural Information Processing Systems (NeurIPS) , 2015
2015
Earlier work this paper cites.
M. Jaderberg, K. Simonyan, A. Zisserman et al. , “Spatial transformer networks,” in Neural Information Processing Systems (NeurIPS) , 2015
2015
Cited alongside, same era.
B. Li, C. Shen, Y. Dai, A. Van Den Hengel, and M. He, “Depth and surface normal estimation from monocular images using regression on deep features and hierarchical crfs,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2015
2015
Cited alongside, same era.
P. Wang, X. Shen, Z. Lin, S. Cohen, B. Price, and A. L. Yuille, “Towards unified depth and semantic prediction from a single image,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2015
2015
Cited alongside, same era.
J. L. Schonberger and J.-M. Frahm, “Structure-from-motion revisited,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2016
2016
Cited alongside, same era.
Y. Zou, Z. Luo, and J.-B. Huang, “DF-Net: Unsupervised joint learning of depth and flow using cross-task consistency,” in European Conference on Computer Vision (ECCV) , 2018
2018
Later among the works it cites.
Z. Teed and J. Deng, “Deepv2d: Video to depth with differentiable structure from motion,” in International Conference on Learning Representations (ICLR) , 2018
2018
Later among the works it cites.
J.-W. Bian, Y.-H. Wu, J. Zhao, Y. Liu, L. Zhang, M.-M. Cheng, and I. Reid, “An evaluation of feature matchers for fundamental matrix estimation,” in British Machine Vision Conference (BMVC) , 2019
2019
Later among the works it cites.
W. Yin, Y. Liu, C. Shen, and Y. Yan, “Enforcing geometric constraints of virtual normal for depth prediction,” in IEEE International Conference on Computer Vision (ICCV) , 2019
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
F. Liu, C. Shen, G. Lin, and I. Reid, “Learning depth from single monocular images using deep convolutional neural fields,” IEEE Transactions on Pattern Recognition and Machine Intelligence (TPAMI) , 2016
2016
Cited alongside, same era.
A. Chakrabarti, J. Shao, and G. Shakhnarovich, “Depth from a single image by harmonizing overcomplete local network predictions,” in Neural Information Processing Systems (NeurIPS) , 2016
2016
Cited alongside, same era.
I. Laina, C. Rupprecht, V. Belagiannis, F. Tombari, and N. Navab, “Deeper depth prediction with fully convolutional residual networks,” in 3DV , 2016
2016
Cited alongside, same era.
R. Garg, V. K. BG, G. Carneiro, and I. Reid, “Unsupervised cnn for single view depth estimation: Geometry to the rescue,” in European Conference on Computer Vision (ECCV) , 2016
2016
Cited alongside, same era.
M. Cordts, M. Omran, S. Ramos, T. Rehfeld, M. Enzweiler, R. Benenson, U. Franke, S. Roth, and B. Schiele, “The cityscapes dataset for semantic urban scene understanding,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2016
2016
Cited alongside, same era.
J. L. Schönberger, E. Zheng, M. Pollefeys, and J.-M. Frahm, “Pixelwise view selection for unstructured multi-view stereo,” in European Conference on Computer Vision (ECCV) , 2016
2016
Cited alongside, same era.
2016
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2016
2016
Cited alongside, same era.
H. Zhan, C. S. Weerasekera, R. Garg, and I. Reid, “Self-supervised learning for single view depth and surface normal estimation,” in IEEE International Conference on Robotics and Automation (ICRA) , 2019
2019
Later among the works it cites.
J. Watson, M. Firman, G. J. Brostow, and D. Turmukhambetov, “Self-supervised monocular depth hints,” in IEEE International Conference on Computer Vision (ICCV) , 2019
2019
Later among the works it cites.
A. Ranjan, V. Jampani, K. Kim, D. Sun, J. Wulff, and M. J. Black, “Competitive Collaboration: Joint unsupervised learning of depth, camera motion, optical flow and motion segmentation,” IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2019
2019
Later among the works it cites.
C. Godard, O. Mac Aodha, M. Firman, and G. J. Brostow, “Digging into self-supervised monocular depth prediction,” in IEEE International Conference on Computer Vision (ICCV) , 2019
2019
Later among the works it cites.
A. Gordon, H. Li, R. Jonschkowski, and A. Angelova, “Depth from videos in the wild: Unsupervised monocular depth learning from unknown cameras,” in IEEE International Conference on Computer Vision (ICCV) , 2019
2019
Later among the works it cites.
Y. Chen, C. Schmid, and C. Sminchisescu, “Self-supervised learning with geometric constraints in monocular video: Connecting flow, depth, and camera,” in IEEE International Conference on Computer Vision (ICCV) , 2019
2019
Later among the works it cites.
J. Zhou, Y. Wang, K. Qin, and W. Zeng, “Moving indoor: Unsupervised video depth learning in challenging environments,” in IEEE International Conference on Computer Vision (ICCV) , 2019
2019
Later among the works it cites.
V. Casser, S. Pirk, R. Mahjourian, and A. Angelova, “Depth prediction without the sensors: Leveraging structure for unsupervised learning from monocular videos,” in Association for the Advancement of Artificial Intelligence (AAAI) , 2019
2019
Later among the works it cites.
C. Luo, Z. Yang, P. Wang, Y. Wang, W. Xu, R. Nevatia, and A. Yuille, “Every pixel counts++: Joint learning of geometry and motion with 3d holistic understanding,” IEEE Transactions on Pattern Recognition and Machine Intelligence (TPAMI) , vol. 42, no. 10, pp. 2624–2641, 2019
2019
Later among the works it cites.
J.-W. Bian, W.-Y. Lin, Y. Liu, L. Zhang, S.-K. Yeung, M.-M. Cheng, and I. Reid, “GMS: Grid-based motion statistics for fast, ultra-robust feature correspondence,” International Journal on Computer Vision (IJCV) , 2020
2020
Closest in time.
W. Zhao, S. Liu, Y. Shu, and Y.-J. Liu, “Towards better generalization: Joint depth-pose learning without posenet,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2020
2020
Closest in time.
V. Guizilini, R. Ambrus, S. Pillai, A. Raventos, and A. Gaidon, “3d packing for self-supervised monocular depth estimation,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2020
2020
Closest in time.
Z. Yu, L. Jin, and S. Gao, “P 2
2020
Closest in time.
T. Do, K. Vuong, S. I. Roumeliotis, and H. S. Park, “Surface normal estimation of tilted images via spatial rectifier,” in European Conference on Computer Vision (ECCV) , 2020
2020
Closest in time.
Z. Li, T. Dekel, F. Cole, R. L. Tucker, N. Snavely, C. Liu, and W. T. Freeman, “MannequinChallenge: Learning the depths of moving people by watching frozen people.” IEEE Transactions on Pattern Recognition and Machine Intelligence (TPAMI) , 2020
2020
Closest in time.
X. Luo, J.-B. Huang, R. Szeliski, K. Matzen, and J. Kopf, “Consistent video depth estimation,” in ACM Transactions on Graphics (SIGGRAPH) , 2020
2020
Closest in time.
J.-W. Bian, H. Zhan, N. Wang, Z. Li, L. Zhang, C. Shen, M.-M. Cheng, and I. Reid, “Unsupervised scale-consistent depth learning from video,” International Journal on Computer Vision (IJCV) , 2021
2021
Closest in time.
S. Lee, S. Im, S. Lin, and I. S. Kweon, “Learning monocular depth in dynamic scenes via instance-aware projection consistency,” in Proceedings of the AAAI Conference on Artificial Intelligence (AAAI) , 2021
2021
Closest in time.