Fetching the paper…
Reading the bibliography…
Due to difficulties in acquiring ground truth depth of equirectangular (360) images, the quality and quantity of equirectangular depth data today is insufficient to represent the various scenes in the world.
Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer
Ranftl, R.; Lasinger, K.; Hafner, D.; Schindler, K.; and Koltun, V. 2019 · 1907
Earlier work this paper cites.
Image quality assessment: from error visibility to structural similarity
Wang, Z.; Bovik, A. C.; Sheikh, H. R.; and Simoncelli, E. P. 2004 · 2004
Earlier work this paper cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Dosovitskiy, A.; Beyer, L.; Kolesnikov, A.; Weissenborn, D.; Zhai, X.; Unterthiner, T.; Dehghani, M.; Minderer, M.; Heigold, G.; Gelly, S.; et al. 2020 · 2010
Earlier work this paper cites.
Vision meets robotics: The kitti dataset
Geiger, A.; Lenz, P.; Stiller, C.; and Urtasun, R. 2013 · 2013
Earlier work this paper cites.
Depth map prediction from a single image using a multi-scale deep network
Eigen, D.; Puhrsch, C.; and Fergus, R. 2014 · 2014
Earlier work this paper cites.
Convolutional LSTM network: A machine learning approach for precipitation nowcasting
Xingjian, S.; Chen, Z.; Wang, H.; Yeung, D.-Y.; Wong, W.-K.; and Woo, W.-c. 2015 · 2015
Earlier work this paper cites.
Single-image depth perception in the wild
Chen, W.; Fu, Z.; Yang, D.; and Deng, J. 2016 · 2016
Earlier work this paper cites.
Unsupervised cnn for single view depth estimation: Geometry to the rescue
Garg, R.; Bg, V. K.; Carneiro, G.; and Reid, I. 2016 · 2016
Earlier work this paper cites.
Joint 2d-3d-semantic data for indoor scene understanding
Armeni, I.; Sax, S.; Zamir, A. R.; and Savarese, S. 2017 · 2017
Earlier work this paper cites.
Matterport3d: Learning from rgb-d data in indoor environments
Chang, A.; Dai, A.; Funkhouser, T.; Halber, M.; Niessner, M.; Savva, M.; Song, S.; Zeng, A.; and Zhang, Y. 2017 · 2017
Earlier work this paper cites.
Unsupervised monocular depth estimation with left-right consistency
Godard, C.; Mac Aodha, O.; and Brostow, G. J. 2017 · 2017
Earlier work this paper cites.
Low-cost 360 stereo photography and video capture
Matzen, K.; Cohen, M. F.; Evans, B.; Kopf, J.; and Szeliski, R. 2017 · 2017
Cited alongside, same era.
Semantic scene completion from a single depth image
Song, S.; Yu, F.; Zeng, A.; Chang, A. X.; Savva, M.; and Funkhouser, T. 2017 · 2017
Cited alongside, same era.
Attention is all you need
Vaswani, A.; Shazeer, N.; Parmar, N.; Uszkoreit, J.; Jones, L.; Gomez, A. N.; Kaiser, Ł.; and Polosukhin, I. 2017 · 2017
Cited alongside, same era.
Cube padding for weakly-supervised saliency prediction in 360 videos
Cheng, H.-T.; Chao, C.-H.; Dong, J.-D.; Wen, H.-K.; Liu, T.-L.; and Sun, M. 2018 · 2018
Cited alongside, same era.
Eliminating the blind spot: Adapting 3d object detection and monocular depth estimation to 360 panoramic imagery
Payen de La Garanderie, G.; Atapour Abarghouei, A.; and Breckon, T. P. 2018 · 2018
Cited alongside, same era.
Omnidepth: Dense depth estimation for indoors spherical panoramas
360ˆ ∘ Camera Alignment via Segmentation
Davidson, B.; Alvi, M. S.; and Henriques, J. F. 2020 · 2020
Later among the works it cites.
Geometric Structure Based and Regularized Depth Estimation From 360 Indoor Imagery
Jin, L.; Xu, Y.; Zheng, J.; Zhang, J.; Tang, R.; Xu, S.; Yu, J.; and Gao, S. 2020 · 2020
Later among the works it cites.
360sd-net: 360 stereo depth estimation with learnable cost volume
Wang, N.-H.; Solarte, B.; Tsai, Y.-H.; Chiu, W.-C.; and Sun, M. 2020b · 2020
Later among the works it cites.
Joint 3d layout and depth prediction from a single indoor panorama image
Zeng, W.; Karaoglu, S.; and Gevers, T. 2020 · 2020
Later among the works it cites.
Structured3d: A large photo-realistic dataset for structured 3d modeling
Zheng, J.; Zhang, J.; Li, J.; Tang, R.; Gao, S.; and Zhou, Z. 2020 · 2020
Later among the works it cites.
SliceNet: deep dense depth estimation from a single indoor panorama using a slice-based representation
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zioulis, N.; Karakottas, A.; Zarpalas, D.; and Daras, P. 2018 · 2018
Cited alongside, same era.
Digging into self-supervised monocular depth estimation
Godard, C.; Mac Aodha, O.; Firman, M.; and Brostow, G. J. 2019 · 2019
Cited alongside, same era.
Depth from videos in the wild: Unsupervised monocular depth learning from unknown cameras
Gordon, A.; Li, H.; Jonschkowski, R.; and Angelova, A. 2019 · 2019
Cited alongside, same era.
Web stereo video supervision for depth prediction from dynamic scenes
Wang, C.; Lucey, S.; Perazzi, F.; and Wang, O. 2019 · 2019
Cited alongside, same era.
UprightNet: geometry-aware camera orientation estimation from single images
Xian, W.; Li, Z.; Fisher, M.; Eisenmann, J.; Shechtman, E.; and Snavely, N. 2019 · 2019
Cited alongside, same era.
Spherical view synthesis for self-supervised 360° depth estimation
Zioulis, N.; Karakottas, A.; Zarpalas, D.; Alvarez, F.; and Daras, P. 2019 · 2019
Cited alongside, same era.
Self-supervised Learning of Depth and Camera Motion from 360 Videos
Wang, F.-E.; Hu, H.-N.; Cheng, H.-T.; Lin, J.-T.; Yang, S.-T.; Shih, M.-L.; Chu, H.-K.; and Sun, M. 2018a
Cited in the paper.
Pintore, G.; Agus, M.; Almansa, E.; Schneider, J.; and Gobbetti, E. 2021 · 2021
Closest in time.
Vision transformers for dense prediction
Ranftl, R.; Bochkovskiy, A.; and Koltun, V. 2021 · 2021
Closest in time.
Hohonet: 360 indoor holistic understanding with latent horizontal features
Sun, C.; Sun, M.; and Chen, H.-T. 2021 · 2021
Closest in time.
Mirror3D: Depth Refinement for Mirror Surfaces
Tan, J.; Lin, W.; Chang, A. X.; and Savva, M. 2021 · 2021
Closest in time.
Training data-efficient image transformers & distillation through attention
Touvron, H.; Cord, M.; Douze, M.; Massa, F.; Sablayrolles, A.; and Jégou, H. 2021 · 2021
Closest in time.
Megadepth: Learning single-view depth prediction from internet photos
Li, Z.; and Snavely, N. 2018 · 2050
Closest in time.