Fetching the paper…
Reading the bibliography…
Self-supervised monocular depth estimation presents a powerful method to obtain 3D scene information from single camera images, which is trainable on arbitrary image sequences without requiring depth labels, e.g., from a LiDAR sensor.
Wang, Z., Bovik, A.C., Sheikh, H.R., Simoncelli, E.P.: Image Quality Assessment: From Error Visibility to Structural Similarity. IEEE Trans. on Image Processing 13
2004
Earlier work this paper cites.
Sun, J., Li, Y., Kang, S.B., Shum, H.Y.: Symmetric Stereo Matching for Occlusion Handling. In: Proc. of CVPR. pp. 399–406. San Diego, CA, USA (Jun 2005)
2005
Earlier work this paper cites.
Hirschmüller, H.: Stereo Processing by Semi-Global Matching and Mutual Information. IEEE Trans. on Pattern Analysis and Machine Intelligence (TPAMI) 30
2008
Earlier work this paper cites.
Akhter, I., Sheikh, Y., Khan, S., Kanade, T.: Nonrigid Structure from Motion in Trajectory Space. In: Proc. of NIPS. pp. 41–48. Vancouver, BC, Canada (Dec 2009)
2009
Earlier work this paper cites.
Szeliski, R.: Computer Vision: Algorithms and Applications. Springer Science & Business Media (2010)
2010
Earlier work this paper cites.
Geiger, A., Lenz, P., Stiller, C., Urtasun, R.: Vision Meets Robotics: The KITTI Dataset. International Journal of Robotics Research (IJRR) 32
2013
Earlier work this paper cites.
Eigen, D., Puhrsch, C., Fergus, R.: Depth Map Prediction from a Single Image Using a Multi-Scale Deep Network. In: Proc. of NIPS. pp. 2366–2374. Montréal, QC, Canada (Dec 2014)
2014
Earlier work this paper cites.
Eigen, D., Fergus, R.: Predicting Depth, Surface Normals and Semantic Labels With a Common Multi-Scale Convolutional Architecture. In: Proc. of ICCV. pp. 2650–2658. Santiago, Chile (Dec 2015)
2015
Earlier work this paper cites.
Everingham, M., Van Gool, L., Williams, C.K.I., Winn, J., Zisserman, A.: The Pascal Visual Object Classes Challenge: A Retrospective. International Journal of Computer Vision (IJCV) 111
2015
Earlier work this paper cites.
Ganin, Y., Lempitsky, V.: Unsupervised Domain Adaptation by Backpropagation. In: Proc. of ICML. pp. 1180–1189. Lille, France (Jul 2015)
2015
Earlier work this paper cites.
Jaderberg, M., Simonyan, K., Zisserman, A., Kayukcuoglu, K.: Spatial Transformer Networks. In: Proc. of NIPS. pp. 2017–2025. Montréal, QC, Canada (Dec 2015)
2015
Earlier work this paper cites.
Kingma, D.P., Ba, J.: Adam: A Method for Stochastic Optimization. In: Proc. of ICLR. pp. 1–15. San Diego, CA, USA (May 2015)
2015
Earlier work this paper cites.
Menze, M., Geiger, A.: Object Scene Flow for Autonomous Vehicles. In: Proc. of CVPR. pp. 3061–3070. Boston, MA, USA (Jun 2015)
2015
Earlier work this paper cites.
Russakovsky, O., Deng, J., Su, H., Krause, J., Satheesh, S., Ma, S., Huang, Z., Karpathy, A., Khosla, A., Bernstein, M., Berg, A.C., Fei-Fei, L.: ImageNet Large Scale Visual Recognition Challenge. International Journal of Computer Vision (IJCV) 115
2015
Earlier work this paper cites.
Cordts, M., Omran, M., Ramos, S., Rehfeld, T., Enzweiler, M., Benenson, R., Franke, U., Roth, S., Schiele, B.: The Cityscapes Dataset for Semantic Urban Scene Understanding. In: Proc. of CVPR. pp. 3213–3223. Las Vegas, NV, USA (Jun 2016)
2016
Earlier work this paper cites.
Garg, R., BG, V.K., Carneiro, G., Reid, I.: Unsupervised CNN for Single View Depth Estimation: Geometry to the Rescue. In: Proc. of ECCV. pp. 740–756. Amsterdam, The Netherlands (Oct 2016)
2016
Earlier work this paper cites.
He, K., Zhang, X., Ren, S., Sun, J.: Deep Residual Learning for Image Recognition. In: Proc. of CVPR. pp. 770–778. Las Vegas, NV, USA (Jun 2016)
2016
Earlier work this paper cites.
Liu, F., Shen, C., Lin, G., Reid, I.: Learning Depth from Single Monocular Images Using Deep Convolutional Neural Fields. IEEE Trans. on Pattern Analysis and Machine Intelligence (TPAMI) 38
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
Ranftl, R., Vineet, V., Chen, Q., Koltun, V.: Dense Monocular Depth Estimation in Complex Dynamic Scenes. In: Proc. of CVPR. pp. 4058–4066. Las Vegas, NV, USA (Jun 2016)
2016
Earlier work this paper cites.
Godard, C., Mac Aodha, O., Brostow, G.J.: Unsupervised Monocular Depth Estimation With Left-Right Consistency. In: Proc. of CVPR. pp. 270–279. Honolulu, HI, USA (Jul 2017)
2017
Earlier work this paper cites.
Kuznietsov, Y., Stuckler, J., Leibe, B.: Semi-Supervised Deep Learning for Monocular Depth Map Prediction. In: Proc. of CVPR. pp. 6647–6655. Honolulu, HI, USA (Jul 2017)
2017
Earlier work this paper cites.
Laina, I., Rupprecht, C., Belagiannis, V., Tombari, F., Navab, N.: Deeper Depth Prediction With Fully Convolutional Residual Networks. In: Proc. of 3DV. pp. 239–248. Stanford, CA, USA (Oct 2017)
2017
Earlier work this paper cites.
Ren, Z., Yan, J., Ni, B., Liu, B., Yang, X., Zha, H.: Unsupervised Deep Learning for Optical Flow Estimation. In: Proc. of AAAI. pp. 1495–1501. San Francisco, CA, USA (Feb 2017)
2017
Earlier work this paper cites.
Uhrig, J., Schneider, N., Schneider, L., Franke, U., Brox, T., Geiger, A.: Sparsity Invariant CNNs. In: Proc. of 3DV. pp. 11–20. Verona, Italy (Oct 2017)
2017
Cited alongside, same era.
Vijayanarasimhan, S., Ricco, S., Schmid, C., Sukthankar, R., Fragkiadaki, K.: SfM-Net: Learning of Structure and Motion from Video. arXiv (1704.0780) (Apr 2017)
2017
Cited alongside, same era.
Zhou, T., Brown, M., Snavely, N., Lowe, D.G.: Unsupervised Learning of Depth and Ego-Motion from Video. In: Proc. of CVPR. pp. 1851–1860. Honolulu, HI, USA (Jul 2017)
2017
Cited alongside, same era.
Aleotti, F., Tosi, F., Poggi, M., Mattoccia, S.: Generative Adversarial Networks for Unsupervised Monocular Depth Prediction. In: Proc. of ECCV - Workshops. pp. 1–18. Munich, Germany (Sep 2018)
2018
Cited alongside, same era.
Gordon, A., Li, H., Jonschkowski, R., Angelova, A.: Depth from Videos in the Wild: Unsupervised Monocular Depth Learning from Unknown Cameras. In: Proc. of ICCV. pp. 8977–8986. Seoul, Korea (Oct 2019)
2019
Later among the works it cites.
Kirillov, A., Girshick, R., He, K., Dollár, P.: Panoptic Feature Pyramid Networks. In: Proc. of CVPR. pp. 6399–6408. Long Beach, CA, USA (Jun 2019)
2019
Later among the works it cites.
Li, Z., Dekel, T., Cole, F., Tucker, R., Snavely, N., Liu, C., Freeman, W.T.: Learning the Depths of Moving People by Watching Frozen People. In: Proc. of CVPR. pp. 4521–4530. Long Beach, CA, USA (Jun 2019)
2019
Later among the works it cites.
Liu, L., Zhai, G., Ye, W., Liu, Y.: Unsupervised Learning of Scene Flow Estimation Fusing With Local Rigidity. In: Proc. of IJCAI. pp. 876–882. Macao, China (Aug 2019)
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
Fu, H., Gong, M., Wang, C., Batmanghelich, K., Tao, D.: Deep Ordinal Regression Network for Monocular Depth Estimation. In: Proc. of CVPR. pp. 2002–2011. Salt Lake City, UT, USA (Jun 2018)
2018
Cited alongside, same era.
Godard, C., Mac Aodha, O., Firman, M., Brostow, G.J.: Digging Into Self-Supervised Monocular Depth Estimation. arXiv (1806.01260v4) (Jun 2018)
2018
Cited alongside, same era.
Kendall, A., Gal, Y., Cipolla, R.: Multi-Task Learning Using Uncertainty to Weigh Losses for Scene Geometry and Semantics. In: Proc. of CVPR. pp. 7482–7491. Salt Lake City, UT, USA (Jun 2018)
2018
Cited alongside, same era.
Mahjourian, R., Wicke, M., Angelova, A.: Unsupervised Learning of Depth and Ego-Motion from Monocular Video Using 3D Geometric Constraints. In: Proc. of CVPR. pp. 5667–5675. Salt Lake City, UT, USA (Jun 2018)
2018
Cited alongside, same era.
Pilzer, A., Xu, D., Puscas, M., Ricci, E., Sebe, N.: Unsupervised Adversarial Depth Estimation Using Cycled Generative Networks. In: Proc. of 3DV. pp. 587–595. Verona, Italy (Sep 2018)
2018
Cited alongside, same era.
Ramirez, P.Z., Poggi, M., Tosi, F., Mattoccia, S., Di Stefano, L.: Geometry Meets Semantics for Semi-Supervised Monocular Depth Estimation. In: Proc. of ACCV. pp. 298–313. Perth, Australia (Dec 2018)
2018
Cited alongside, same era.
Wang, C., Miguel Buenaposada, J., Zhu, R., Lucey, S.: Learning Depth From Monocular Videos Using Direct Methods. In: Proc. of CVPR. pp. 2022–2030. Salt Lake City, UT, USA (Jun 2018)
2018
Cited alongside, same era.
Liu, P., Lyu, M., King, I., Xu, J.: SelFlow: Self-Supervised Learning of Optical Flow. In: Proc. of CVPR. pp. 4571–4580. Long Beach, CA, USA (Jun 2019)
2019
Later among the works it cites.
Luo, C., Yang, Z., Wang, P., Wang, Y., Xu, W., Nevatia, R., Yuille, A.: Every Pixel Counts ++: Joint Learning of Geometry and Motion with 3D Holistic Understanding. arXiv (1810.06125) (Jul 2019)
2019
Later among the works it cites.
Meng, Y., Lu, Y., Raj, A., Sunarjo, S., Guo, R., Javidi, T., Bansal, G., Bharadia, D.: SIGNet: Semantic Instance Aided Unsupervised 3D Geometry Perception. In: Proc. of CVPR. pp. 9810–9820. Long Beach, CA, USA (Jun 2019)
2019
Later among the works it cites.
Ochs, M., Kretz, A., Mester, R.: SDNet: Semantically Guided Depth Estimation Network. In: Proc. of GCPR. pp. 288–302. Dortmund, Germany (Sep 2019)
2019
Later among the works it cites.
Ors̆ić, M., Kres̆o, I., Bevandić, P., S̆egvić, S.: In Defense of Pre-Trained ImageNet Architectures for Real-Time Semantic Segmentation of Road-Driving Images. In: Proc. of CVPR. pp. 12607–12616. Long Beach, CA, USA (Jun 2019)
2019
Later among the works it cites.
Paszke, A., Gross, S., Massa, F., Lerer, A., Bradbury, J., Chanan, G., Killeen, T., Lin, Z., Gimelshein, N., Antiga, L., Desmaison, A., Kopf, A., Yang, E., DeVito, Z., Raison, M., Tejani, A., Chilamkurthy, S., Steiner, B., Fang, L., Bai, J., Chintala, S.: PyTorch: An Imperative Style, High-Performance Deep Learning Library. In: Proc. of NeurIPS. pp. 8024–8035. Vancouver, BC, Canada (Dec 2019)
2019
Later among the works it cites.
Pilzer, A., Lathuiliere, S., Sebe, N., Ricci, E.: Refine and Distill: Exploiting Cycle-Inconsistency and Knowledge Distillation for Unsupervised Monocular Depth Estimation. In: Proc. of CVPR. pp. 9768–9777. Long Beach, CA, USA (Jun 2019)
2019
Later among the works it cites.
Ranjan, A., Jampani, V., Balles, L., Kim, K., Sun, D., Wulff, J., Black, M.J.: Competitive Collaboration: Joint Unsupervised Learning of Depth, Camera Motion, Optical Flow and Motion Segmentation. In: Proc. of CVPR. pp. 12240–12249. Long Beach, CA, USA (Jun 2019)
2019
Later among the works it cites.
Tosi, F., Aleotti, F., Poggi, M., Mattoccia, S.: Learning Monocular Depth Estimation Infusing Traditional Stereo Knowledge. In: Proc. of CVPR. pp. 9799–9809. Long Beach, CA, USA (Jun 2019)
2019
Later among the works it cites.
Wang, R., Pizer, S.M., Frahm, J.M.: Recurrent Neural Network for (Un-)Supervised Learning of Monocular Video Visual Odometry and Depth. In: Proc. of CVPR. pp. 5555–5564. Long Beach, CA, USA (Jun 2019)
2019
Later among the works it cites.
Wang, Y., Wang, P., Yang, Z., Luo, C., Yang, Y., Xu, W.: UnOS: Unified Unsupervised Optical-Flow and Stereo-Depth Estimation by Watching Videos. In: Proc. of CVPR. pp. 8071–8081. Long Beach, CA, USA (Jun 2019)
2019
Later among the works it cites.
Zhang, F., Prisacariu, V., Yang, R., Torr, P.H.S.: GA-Net: Guided Aggregation Net for End-to-End Stereo Matching. In: Proc. of CVPR. pp. 185–194. Long Beach, CA, USA (Jun 2019)
2019
Later among the works it cites.
Zhang, H., Shen, C., Li, Y., Cao, Y., Liu, Y., Yan, Y.: Exploiting Temporal Consistency for Real-Time Video Depth Estimation. In: Proc. of ICCV. pp. 1725–1734. Seoul, Korea (Oct 2019)
2019
Later among the works it cites.
Zhang, Z., Cui, Z., Xu, C., Yan, Y., Sebe, N., Yang, J.: Pattern-Affinitive Propagation Across Depth, Surface Normal and Semantic Segmentation. In: Proc. of CVPR. pp. 4106–4115. Long Beach, CA, USA (Jun 2019)
2019
Later among the works it cites.
Zhao, S., Fu, H., Gong, M., Tao, D.: Geometry-Aware Symmetric Domain Adaptation for Monocular Depth Estimation. In: Proc. of CVPR. Long Beach, CA, USA (Jun 2019)
2019
Later among the works it cites.
Zhou, J., Wang, Y., Qin, K., Zeng, W.: Unsupervised High-Resolution Depth Learning From Videos With Dual Networks. In: Proc. of ICCV. pp. 6872–6881. Seoul, Korea (Oct 2019)
2019
Later among the works it cites.
Guizilini, V., Ambrus, R., Pillai, S., Gaidon, A.: 3D Packing for Self-Supervised Monocular Depth Estimation. In: Proc. of CVPR. pp. 2485–2494. Seattle, WA, USA (Jun 2020)
2020
Closest in time.
Guizilini, V., Hou, R., Li, J., Ambrus, R., Gaidon, A.: Semantically-Guided Representation Learning for Self-Supervised Monocular Depth. In: Proc. of ICLR. pp. 1–14. Addis Ababa, Ethiopia (Apr 2020)
2020
Closest in time.
Tosi, F., Aleotti, F., Ramirez, P.Z., Poggi, M., Salti, S., Stefano, L.D., Mattoccia, S.: Distilled Semantics for Comprehensive Scene Understanding from Videos. In: Proc. of CVPR. pp. 4654–4665. Seattle, WA, USA (Jun 2020)
2020
Closest in time.