Fetching the paper…
Reading the bibliography…
Occupancy prediction reconstructs 3D structures of surrounding environments.
Z. Yin and J. Shi, “Geonet: Unsupervised learning of dense depth, optical flow and camera pose,” in IEEE Conference on Computer Vision Pattern Recognision , 2018, pp. 1983–1992
1992
Earlier work this paper cites.
N. Max, “Optical models for direct volume rendering,” IEEE Transactions on Visualization and Computer Graphics , vol. 1, no. 2, pp. 99–108, 1995
1995
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in IEEE Conference on Computer Vision Pattern Recognision , 2009, pp. 248–255
2009
Earlier work this paper cites.
H. Fu, M. Gong, C. Wang, K. Batmanghelich, and D. Tao, “Deep ordinal regression network for monocular depth estimation,” in IEEE Conference on Computer Vision Pattern Recognision , 2018, pp. 2002–2011
2011
Earlier work this paper cites.
F. Liu, C. Shen, G. Lin, and I. Reid, “Learning depth from single monocular images using deep convolutional neural fields,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 38, no. 10, pp. 2024–2039, 2015
2015
Earlier work this paper cites.
A. Roy and S. Todorovic, “Monocular depth estimation using neural regression forest,” in IEEE Conference on Computer Vision Pattern Recognision , 2016, pp. 5506–5514
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in IEEE Conference on Computer Vision Pattern Recognision , 2016, pp. 770–778
2016
Earlier work this paper cites.
T. Zhou, M. Brown, N. Snavely, and D. G. Lowe, “Unsupervised learning of depth and ego-motion from video,” in IEEE Conference on Computer Vision Pattern Recognision , 2017, pp. 1851–1858
2017
Earlier work this paper cites.
C. Godard, O. Mac Aodha, and G. J. Brostow, “Unsupervised monocular depth estimation with left-right consistency,” in IEEE Conference on Computer Vision Pattern Recognision , 2017, pp. 270–279
2017
Earlier work this paper cites.
Z. Zhang, C. Xu, J. Yang, J. Gao, and Z. Cui, “Progressive hard-mining network for monocular depth estimation,” IEEE Transactions on Image Processing , vol. 27, no. 8, pp. 3691–3702, 2018
2018
Earlier work this paper cites.
R. Mahjourian, M. Wicke, and A. Angelova, “Unsupervised learning of depth and ego-motion from monocular video using 3d geometric constraints,” in IEEE Conference on Computer Vision Pattern Recognision , 2018, pp. 5667–5675
2018
Earlier work this paper cites.
C. Lu, M. J. G. van de Molengraft, and G. Dubbelman, “Monocular semantic occupancy grid mapping with convolutional variational encoder–decoder networks,” IEEE Robotics and Automation Letters , vol. 4, no. 2, pp. 445–452, 2019
2019
Earlier work this paper cites.
J. Behley, M. Garbade, A. Milioto, J. Quenzel, S. Behnke, C. Stachniss, and J. Gall, “Semantickitti: A dataset for semantic scene understanding of lidar sequences,” in International Conference on Computer Vision , 2019, pp. 9297–9307
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
J. Watson, M. Firman, G. J. Brostow, and D. Turmukhambetov, “Self-Supervised Monocular Depth Hints,” in International Conference on Computer Vision , 2019, pp. 2162–2171
2019
Earlier work this paper cites.
A. Ranjan, V. Jampani, L. Balles, K. Kim, D. Sun, J. Wulff, and M. J. Black, “Competitive Collaboration: Joint Unsupervised Learning of Depth, Camera Motion, Optical Flow and Motion Segmentation,” in IEEE Conference on Computer Vision Pattern Recognision , 2019, pp. 12 240–12 249
2019
Earlier work this paper cites.
F. Tosi, F. Aleotti, M. Poggi, and S. Mattoccia, “Learning monocular depth estimation infusing traditional stereo knowledge,” in IEEE Conference on Computer Vision Pattern Recognision , 2019, pp. 9799–9809
2019
Earlier work this paper cites.
J.-W. Bian, Z. Li, N. Wang, H. Zhan, C. Shen, M.-M. Cheng, and I. Reid, “Unsupervised Scale-consistent Depth and Ego-motion Learning from Monocular Video,” in Advances in Neural Information Processing Systems , 2019, pp. 35–45
2019
Earlier work this paper cites.
C. Godard, O. Mac Aodha, M. Firman, and G. J. Brostow, “Digging into self-supervised monocular depth estimation,” in International Conference on Computer Vision , 2019, pp. 3828–3838
2019
Earlier work this paper cites.
B. Mildenhall, P. P. Srinivasan, M. Tancik, J. T. Barron, R. Ramamoorthi, and R. Ng, “Nerf: Representing scenes as neural radiance fields for view synthesis,” in European Conference on Computer Vision . Springer, 2020, pp. 405–421
2020
Earlier work this paper cites.
H. Caesar, V. Bankiti, A. H. Lang, S. Vora, V. E. Liong, Q. Xu, A. Krishnan, Y. Pan, G. Baldan, and O. Beijbom, “nuscenes: A multimodal dataset for autonomous driving,” in IEEE Conference on Computer Vision Pattern Recognision , 2020, pp. 11 621–11 631
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
V. Guizilini, R. Ambrus, S. Pillai, A. Raventos, and A. Gaidon, “3d packing for self-supervised monocular depth estimation,” in IEEE Conference on Computer Vision Pattern Recognision , 2020, pp. 2485–2494
2020
Earlier work this paper cites.
L. Roldao, R. de Charette, and A. Verroust-Blondet, “Lmscnet: Lightweight multiscale 3d semantic completion,” in International Conference on 3D Vision . IEEE, 2020, pp. 111–119
2020
Earlier work this paper cites.
X. Chen, K.-Y. Lin, C. Qian, G. Zeng, and H. Li, “3d sketch-aware semantic scene completion via semi-supervised structure prior,” in IEEE Conference on Computer Vision Pattern Recognision , 2020, pp. 4193–4202
2020
Earlier work this paper cites.
J. Li, K. Han, P. Wang, Y. Liu, and X. Yuan, “Anisotropic convolutional networks for 3d semantic scene completion,” in IEEE Conference on Computer Vision Pattern Recognision , 2020, pp. 3351–3359
2020
Earlier work this paper cites.
A. Hu, Z. Murez, N. Mohan, S. Dudas, J. Hawke, V. Badrinarayanan, R. Cipolla, and A. Kendall, “Fiery: Future instance prediction in bird’s-eye view from surround monocular cameras,” in International Conference on Computer Vision , 2021, pp. 15 273–15 282
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
J. T. Barron, B. Mildenhall, M. Tancik, P. Hedman, R. Martin-Brualla, and P. P. Srinivasan, “Mip-nerf: A multiscale representation for anti-aliasing neural radiance fields,” in International Conference on Computer Vision , 2021, pp. 5855–5864
2021
Earlier work this paper cites.
K. Park, U. Sinha, J. T. Barron, S. Bouaziz, D. B. Goldman, S. M. Seitz, and R. Martin-Brualla, “Nerfies: Deformable neural radiance fields,” in International Conference on Computer Vision , 2021, pp. 5865–5874
2021
Earlier work this paper cites.
C. Gao, A. Saraf, J. Kopf, and J.-B. Huang, “Dynamic view synthesis from dynamic monocular video,” in International Conference on Computer Vision , 2021, pp. 5712–5721
2021
Earlier work this paper cites.
A. Pumarola, E. Corona, G. Pons-Moll, and F. Moreno-Noguer, “D-nerf: Neural radiance fields for dynamic scenes,” in IEEE Conference on Computer Vision Pattern Recognision , 2021, pp. 10 318–10 327
2021
Cited alongside, same era.
E. Tretschk, A. Tewari, V. Golyanik, M. Zollhöfer, C. Lassner, and C. Theobalt, “Non-rigid neural radiance fields: Reconstruction and novel view synthesis of a dynamic scene from monocular video,” in International Conference on Computer Vision , 2021, pp. 12 959–12 970
2021
Cited alongside, same era.
Z. Li, S. Niklaus, N. Snavely, and O. Wang, “Neural scene flow fields for space-time view synthesis of dynamic scenes,” in IEEE Conference on Computer Vision Pattern Recognision , 2021, pp. 6498–6508
2021
Cited alongside, same era.
Y. Du, Y. Zhang, H.-X. Yu, J. B. Tenenbaum, and J. Wu, “Neural radiance flow for 4d view synthesis and video processing,” in 2021 IEEE/CVF International Conference on Computer Vision (ICCV) . IEEE Computer Society, 2021, pp. 14 304–14 314
2021
J. Xu, X. Liu, Y. Bai, J. Jiang, K. Wang, X. Chen, and X. Ji, “Multi-camera collaborative depth prediction via consistent structure estimation,” in ACM International Conference on Multimedia , 2022, pp. 2730–2738
2022
Later among the works it cites.
Y. Wei, L. Zhao, W. Zheng, Z. Zhu, J. Zhou, and J. Lu, “Surroundocc: Multi-camera 3d occupancy prediction for autonomous driving,” in International Conference on Computer Vision , 2023, pp. 21 729–21 740
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
K. Park, U. Sinha, P. Hedman, J. T. Barron, S. Bouaziz, D. B. Goldman, R. Martin-Brualla, and S. M. Seitz, “Hypernerf: a higher-dimensional representation for topologically varying neural radiance fields,” ACM Transactions on Graphics (TOG) , vol. 40, no. 6, pp. 1–12, 2021
2021
Cited alongside, same era.
P. Wang, L. Liu, Y. Liu, C. Theobalt, T. Komura, and W. Wang, “Neus: Learning neural implicit surfaces by volume rendering for multi-view reconstruction,” Advances in Neural Information Processing Systems , vol. 34, pp. 27 171–27 183, 2021
2021
Cited alongside, same era.
L. Yariv, J. Gu, Y. Kasten, and Y. Lipman, “Volume rendering of neural implicit surfaces,” Advances in Neural Information Processing Systems , vol. 34, pp. 4805–4815, 2021
2021
Cited alongside, same era.
M. Oechsle, S. Peng, and A. Geiger, “Unisurf: Unifying neural implicit surfaces and radiance fields for multi-view reconstruction,” in International Conference on Computer Vision , 2021, pp. 5589–5599
2021
Cited alongside, same era.
Y. Wei, S. Liu, Y. Rao, W. Zhao, J. Lu, and J. Zhou, “Nerfingmvs: Guided optimization of neural radiance fields for indoor multi-view stereo,” in International Conference on Computer Vision , 2021, pp. 5610–5619
2021
Cited alongside, same era.
S. J. Garbin, M. Kowalski, M. Johnson, J. Shotton, and J. Valentin, “Fastnerf: High-fidelity neural rendering at 200fps,” in International Conference on Computer Vision , 2021, pp. 14 346–14 355
2021
Cited alongside, same era.
C. Reiser, S. Peng, Y. Liao, and A. Geiger, “Kilonerf: Speeding up neural radiance fields with thousands of tiny mlps,” in International Conference on Computer Vision , 2021, pp. 14 335–14 345
2021
Cited alongside, same era.
X. Xu, Z. Chen, and F. Yin, “Multi-scale spatial attention-guided monocular depth estimation with semantic enhancement,” IEEE Transactions on Image Processing , vol. 30, pp. 8811–8822, 2021
2021
Cited alongside, same era.
Y. Zhang, Z. Zhu, and D. Du, “Occformer: Dual-path transformer for vision-based 3d semantic occupancy prediction,” in International Conference on Computer Vision , 2023
2023
Closest in time.
X. Wang, Z. Zhu, W. Xu, Y. Zhang, Y. Wei, X. Chi, Y. Ye, D. Du, J. Lu, and X. Wang, “Openoccupancy: A large scale benchmark for surrounding semantic occupancy perception,” in International Conference on Computer Vision , 2023
2023
Closest in time.
A. Kirillov, E. Mintun, N. Ravi, H. Mao, C. Rolland, L. Gustafson, T. Xiao, S. Whitehead, A. C. Berg, W.-Y. Lo et al. , “Segment anything,” in International Conference on Computer Vision , 2023
2023
Closest in time.
2023
Closest in time.
Y. Huang, W. Zheng, Y. Zhang, J. Zhou, and J. Lu, “Tri-perspective view for vision-based 3d semantic occupancy prediction,” in IEEE Conference on Computer Vision Pattern Recognision , 2023, pp. 9223–9232
2023
Closest in time.
2023
Closest in time.
W. Tong, C. Sima, T. Wang, L. Chen, S. Wu, H. Deng, Y. Gu, L. Lu, P. Luo, D. Lin et al. , “Scene as occupancy,” in International Conference on Computer Vision , 2023, pp. 8406–8415
2023
Closest in time.
A.-Q. Cao and R. de Charette, “Scenerf: Self-supervised monocular 3d scene reconstruction with radiance fields,” in International Conference on Computer Vision , 2023, pp. 9387–9398
2023
Closest in time.
2023
Closest in time.
Y. Li, Z. Yu, C. Choy, C. Xiao, J. M. Alvarez, S. Fidler, C. Feng, and A. Anandkumar, “Voxformer: Sparse voxel transformer for camera-based 3d semantic scene completion,” in IEEE Conference on Computer Vision Pattern Recognision , 2023, pp. 9087–9098
2023
Closest in time.
Z. Li, Z. Yu, W. Wang, A. Anandkumar, T. Lu, and J. M. Alvarez, “Fb-bev: Bev representation from forward-backward view transformations,” in International Conference on Computer Vision , 2023, pp. 6919–6928
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
J. T. Barron, B. Mildenhall, D. Verbin, P. P. Srinivasan, and P. Hedman, “Zip-nerf: Anti-aliased grid-based neural radiance fields,” in International Conference on Computer Vision , 2023
2023
Closest in time.
W. Zhang, R. Xing, Y. Zeng, Y.-S. Liu, K. Shi, and Z. Han, “Fast learning radiance fields by shooting much fewer rays,” IEEE Transactions on Image Processing , vol. 32, pp. 2703–2718, 2023
2023
Closest in time.
G. Li, R. Huang, H. Li, Z. You, and W. Chen, “Sense: Self-evolving learning for self-supervised monocular depth estimation,” IEEE Transactions on Image Processing , 2023
2023
Closest in time.
Y. Wei, L. Zhao, W. Zheng, Z. Zhu, Y. Rao, G. Huang, J. Lu, and J. Zhou, “Surrounddepth: Entangling surrounding views for self-supervised multi-camera depth estimation,” in Conference on Robot Learning . PMLR, 2023, pp. 539–549
2023
Closest in time.
A. Schmied, T. Fischer, M. Danelljan, M. Pollefeys, and F. Yu, “R3d3: Dense 3d reconstruction of dynamic scenes from multiple cameras,” in International Conference on Computer Vision , 2023, pp. 3216–3226
2023
Closest in time.
Y. Shi, H. Cai, A. Ansari, and F. Porikli, “Ega-depth: Efficient guided attention for self-supervised multi-camera depth estimation,” in IEEE Conference on Computer Vision Pattern Recognision Workshops , 2023, pp. 119–129
2023
Closest in time.
A. W. Harley, Z. Fang, J. Li, R. Ambrus, and K. Fragkiadaki, “Simple-BEV: What really matters for multi-sensor bev perception?” in IEEE International Conference on Robotics and Automation , 2023
2023
Closest in time.
W. Gan, W. Wu, S. Chen, Y. Zhao, and P. K. Wong, “Rethinking 3d cost aggregation in stereo matching,” Pattern Recognition Letters , vol. 167, pp. 75–81, 2023
2023
Closest in time.
2023
Closest in time.
M. Pan, J. Liu, R. Zhang, P. Huang, X. Li, L. Liu, and S. Zhang, “Renderocc: Vision-centric 3d occupancy prediction with 2d rendering supervision,” in IEEE International Conference on Robotics and Automation , 2024
2024
Closest in time.
M. Chen, L. Wang, Y. Lei, Z. Dong, and Y. Guo, “Learning spherical radiance field for efficient 360° unbounded novel view synthesis,” IEEE Transactions on Image Processing , 2024
2024
Closest in time.
J. L. G. Bello, J. Moon, and M. Kim, “Self-supervised monocular depth estimation with positional shift depth variance and adaptive disparity quantization,” IEEE Transactions on Image Processing , 2024
2024
Closest in time.
C. Wang, J. Miguel Buenaposada, R. Zhu, and S. Lucey, “Learning depth from monocular videos using direct methods,” in IEEE Conference on Computer Vision Pattern Recognision , 2018, pp. 2022–2030
2030
Closest in time.