Fetching the paper…
Reading the bibliography…
This work addresses the challenge of high-quality surface normal estimation from monocular colored inputs (i.e., images and videos), a field which has recently been revolutionized by repurposing diffusion priors.
DIODE: A Dense Indoor and Outdoor DEpth Dataset
Igor Vasiljevic, Nick Kolkin, Shanyi Zhang, Ruotian Luo, Haochen Wang, Falcon Z. Dai, Andrea F. Daniele, Mohammadreza Mostajabi, Steven Basart, Matthew R. Walter, and Gregory Shakhnarovich. 2019 · 1908
Earlier work this paper cites.
Automatic photo pop-up
Derek Hoiem, Alexei A. Efros, and Martial Hebert. 2005 · 2005
Earlier work this paper cites.
Recovering Surface Layout from an Image
Derek Hoiem, Alexei A. Efros, and Martial Hebert. 2007 · 2007
Earlier work this paper cites.
Hypersim: A Photorealistic Synthetic Dataset for Holistic Indoor Scene Understanding
Mike Roberts, Jason Ramapuram, Anurag Ranjan, Atulit Kumar, Miguel Angel Bautista, Nathan Paczan, Russ Webb, and Joshua M. Susskind. 2021 · 2011
Earlier work this paper cites.
Indoor Segmentation and Support Inference from RGBD Images. In European Conference on Computer Vision
Nathan Silberman, Derek Hoiem, Pushmeet Kohli, and Rob Fergus. 2012 · 2012
Earlier work this paper cites.
Data-Driven 3D Primitives for Single Image Understanding. In 2013 IEEE International Conference on Computer Vision
David F. Fouhey, Abhinav Gupta, and Martial Hebert. 2013b · 2013
Earlier work this paper cites.
Unfolding an Indoor Origami World
David Ford Fouhey, Abhinav Gupta, and Martial Hebert. 2014 · 2014
Earlier work this paper cites.
Large Scale Multi-view Stereopsis Evaluation
Rasmus Ramsbøl Jensen, A. Dahl, George Vogiatzis, Engil Tola, and Henrik Aanæs. 2014 · 2014
Earlier work this paper cites.
Discriminatively Trained Dense Surface Normal Estimation
L’ubor Ladický, Bernhard Zeisl, and Marc Pollefeys. 2014 · 2014
Earlier work this paper cites.
Predicting Depth, Surface Normals and Semantic Labels with a Common Multi-Scale Convolutional Architecture. In 2015 IEEE International Conference on Computer Vision (ICCV)
David Eigen and Rob Fergus. 2015b · 2015
Earlier work this paper cites.
U-Net: Convolutional Networks for Biomedical Image Segmentation
Olaf Ronneberger, Philipp Fischer, and Thomas Brox. 2015 · 2015
Earlier work this paper cites.
Designing Deep Networks for Surface Normal Estimation. In 2015 IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
Xiaolong Wang, David F. Fouhey, and Abhinav Gupta. 2015b · 2015
Earlier work this paper cites.
Marr Revisited: 2D-3D Alignment via Surface Normal Prediction. In 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
Aayush Bansal, Bryan Russell, and Abhinav Gupta. 2016b · 2016
Earlier work this paper cites.
SURGE: surface regularized geometry estimation from a single image
Peng Wang, Xiaohui Shen, Bryan Russell, Scott Cohen, Brian Price, and AlanL. Yuille. 2016 · 2016
Earlier work this paper cites.
ScanNet: Richly-annotated 3D Reconstructions of Indoor Scenes
Angela Dai, Angel X. Chang, Manolis Savva, Maciej Halber, Thomas Funkhouser, and Matthias Nießner. 2017 · 2017
Earlier work this paper cites.
Evaluation of CNN-based Single-Image Depth Estimation Methods
Tobias Koch, Lukas Liebel, Friedrich Fraundorfer, and Marco Körner. 2018 · 2018
Earlier work this paper cites.
GeoNet: Geometric Neural Network for Joint Depth and Surface Normal Estimation. In 2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition
Xiaojuan Qi, Renjie Liao, Zhengzhe Liu, Raquel Urtasun, and Jiaya Jia. 2018 · 2018
Earlier work this paper cites.
FrameNet: Learning Local Canonical Frames of 3D Surfaces from a Single RGB Image
Jingwei Huang, Yichao Zhou, Thomas Funkhouser, and LeonidasJ. Guibas. 2019 · 2019
Earlier work this paper cites.
Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer
Katrin Lasinger, René Ranftl, Konrad Schindler, and Vladlen Koltun. 2019 · 2019
Earlier work this paper cites.
Spherical Regression: Learning Viewpoints, Surface Normals and 3D Rotations on n-Spheres
Shuai Liao, Efstratios Gavves, and CeesG.M. Snoek. 2019 · 2019
Earlier work this paper cites.
Decoupled Weight Decay Regularization
Ilya Loshchilov and Frank Hutter. 2019 · 2019
Cited alongside, same era.
3D Ken Burns Effect from a Single Image
Simon Niklaus, Long Mai, Jimei Yang, and Feng Liu. 2019 · 2019
Cited alongside, same era.
A Benchmark Dataset and Evaluation for Non-Lambertian and Uncalibrated Photometric Stereo
Boxin Shi, Zhipeng Mo, Zhe Wu, Dinglong Duan, Sai-Kit Yeung, and Ping Tan. 2019 · 2019
Cited alongside, same era.
The Replica dataset: A digital replica of indoor spaces
Julian Straub, Thomas Whelan, Lingni Ma, Yufan Chen, Erik Wijmans, Simon Green, Jakob J Engel, Raul Mur-Artal, Carl Ren, Shobhit Verma, et al · 2019
Cited alongside, same era.
Pattern-Affinitive Propagation across Depth, Surface Normal and Semantic Segmentation. In 2019 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
Zhenyu Zhang, Zhen Cui, Chunyan Xu, Yan Yan, Nicu Sebe, and Jian Yang. 2019 · 2019
Scalable Diffusion Models with Transformers
William Peebles and Saining Xie. 2022 · 2022
Later among the works it cites.
Laion-5b: An open large-scale dataset for training next generation image-text models
Christoph Schuhmann, Romain Beaumont, Richard Vencu, Cade Gordon, Ross Wightman, Mehdi Cherti, Theo Coombes, Aarush Katta, Clayton Mullis, Mitchell Wortsman, et al · 2022
Later among the works it cites.
Background Prompting for Improved Object Depth
Manel Baradad, Yuanzhen Li, Forrester Cole, Michael Rubinstein, Antonio Torralba, William T. Freeman, and Varun Jampani. 2023 · 2023
Later among the works it cites.
Ddp: Diffusion model for dense visual prediction. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 21741–21752
Yuanfeng Ji, Zhe Chen, Enze Xie, Lanqing Hong, Xihui Liu, Zhaoqiang Liu, Tong Lu, Zhenguo Li, and Ping Luo. 2023 · 2023
Later among the works it cites.
Hyperhuman: Hyper-realistic human generation with latent structural diffusion
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
MVD 2 : Efficient Multiview 3D Reconstruction for Multiview Diffusion
Xin-Yang Zheng, Hao Pan, Yu-Xiao Guo, Xin Tong, and Yang Liu. 2024 · 2019
Cited alongside, same era.
Surface Normal Estimation of Tilted Images via Spatial Rectifier
TienVan Do, Khiem Vuong, StergiosI. Roumeliotis, and HyunSoo Park. 2020 · 2020
Cited alongside, same era.
Denoising diffusion probabilistic models
Jonathan Ho, Ajay Jain, and Pieter Abbeel. 2020 · 2020
Cited alongside, same era.
GeoNet++: Iterative Geometric Neural Network with Edge-Aware Refinement for Joint Depth and Surface Normal Estimation
Xiaojuan Qi, Zhengzhe Liu, Renjie Liao, Philip H. S. Torr, Raquel Urtasun, and Jiaya Jia. 2022 · 2020
Cited alongside, same era.
Denoising diffusion implicit models
Jiaming Song, Chenlin Meng, and Stefano Ermon. 2020 · 2020
Cited alongside, same era.
VPLNet: Deep Single View Normal Estimation With Vanishing Points and Lines. In 2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
Rui Wang, David Geraghty, Kevin Matzen, Richard Szeliski, and Jan-Michael Frahm. 2020 · 2020
Cited alongside, same era.
Estimating and Exploiting the Aleatoric Uncertainty in Surface Normal Estimation. In 2021 IEEE/CVF International Conference on Computer Vision (ICCV)
Gwangbin Bae, Ignas Budvytis, and Roberto Cipolla. 2021 · 2021
Cited alongside, same era.
Xian Liu, Jian Ren, Aliaksandr Siarohin, Ivan Skorokhodov, Yanyu Li, Dahua Lin, Xihui Liu, Ziwei Liu, and Sergey Tulyakov. 2023 · 2023
Later among the works it cites.
Wonder3d: Single image to 3d using cross-domain diffusion
Xiaoxiao Long, Yuan-Chen Guo, Cheng Lin, Yuan Liu, Zhiyang Dou, Lingjie Liu, Yuexin Ma, Song-Hai Zhang, Marc Habermann, Christian Theobalt, et al · 2023
Later among the works it cites.
Dreamfusion: Text-to-3d using 2d diffusion
Ben Poole, Ajay Jain, Jonathan T Barron, and Ben Mildenhall. 2023 · 2023
Later among the works it cites.
In-context learning unlocked for diffusion models
Zhendong Wang, Yifan Jiang, Yadong Lu, Pengcheng He, Weizhu Chen, Zhangyang Wang, Mingyuan Zhou, et al · 2023
Later among the works it cites.
I2vgen-xl: High-quality image-to-video synthesis via cascaded diffusion models
Shiwei Zhang, Jiayu Wang, Yingya Zhang, Kang Zhao, Hangjie Yuan, Zhiwu Qin, Xiang Wang, Deli Zhao, and Jingren Zhou. 2023b · 2023
Later among the works it cites.
Unleashing text-to-image diffusion models for visual perception. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 5729–5739
Wenliang Zhao, Yongming Rao, Zuyan Liu, Benlin Liu, Jie Zhou, and Jiwen Lu. 2023 · 2023
Later among the works it cites.
Rethinking Inductive Biases for Surface Normal Estimation. In IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
Gwangbin Bae and Andrew J. Davison. 2024 · 2024
Closest in time.
Exploiting the signal-leak bias in diffusion models. In Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision . 4025–4034
Martin Nicolas Everaert, Athanasios Fitsios, Marco Bocchio, Sami Arpa, Sabine Süsstrunk, and Radhakrishna Achanta. 2024 · 2024
Closest in time.
GeoWizard: Unleashing the Diffusion Priors for 3D Geometry Estimation from a Single Image
Xiao Fu, Wei Yin, Mu Hu, Kaixuan Wang, Yuexin Ma, Ping Tan, Shaojie Shen, Dahua Lin, and Xiaoxiao Long. 2024b · 2024
Closest in time.
2D Gaussian Splatting for Geometrically Accurate Radiance Fields. In SIGGRAPH 2024 Conference Papers . Association for Computing Machinery
Binbin Huang, Zehao Yu, Anpei Chen, Andreas Geiger, and Shenghua Gao. 2024 · 2024
Closest in time.
Intrinsic Image Diffusion for Single-view Material Estimation. In Computer Vision and Pattern Recognition (CVPR)
Peter Kocsis, Vincent Sitzmann, and Matthias Nießner. 2024 · 2024
Closest in time.
Direct2.5: Diverse Text-to-3D Generation via Multi-view 2.5D Diffusion
Yuanxun Lu, Jingyang Zhang, Shiwei Li, Tian Fang, David McKinnon, Yanghai Tsin, Long Quan, Xun Cao, and Yao Yao. 2024 · 2024
Closest in time.
DINOv2: Learning Robust Visual Features without Supervision
Maxime Oquab, Timothée Darcet, Théo Moutakanni, Huy Vo, Marc Szafraniec, Vasil Khalidov, Pierre Fernandez, Daniel Haziza, Francisco Massa, Alaaeldin El-Nouby, Mahmoud Assran, Nicolas Ballas, Wojciech Galuba, Russell Howes, Po-Yao Huang, Shang-Wen Li, Ishan Misra, Michael Rabbat, Vasu Sharma, Gabriel Synnaeve, Hu Xu, Hervé Jegou, Julien Mairal, Patrick Labatut, Armand Joulin, and Piotr Bojanowski. 2024 · 2024
Closest in time.
Richdreamer: A generalizable normal-depth diffusion model for detail richness in text-to-3d. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 9914–9925
Lingteng Qiu, Guanying Chen, Xiaodong Gu, Qi Zuo, Mutian Xu, Yushuang Wu, Weihao Yuan, Zilong Dong, Liefeng Bo, and Xiaoguang Han. 2024 · 2024
Closest in time.
Diffuse, Attend, and Segment: Unsupervised Zero-Shot Segmentation using Stable Diffusion
Junjiao Tian, Lavisha Aggarwal, Andrea Colaco, Zsolt Kira, and Mar Gonzalez-Franco. 2024 · 2024
Closest in time.
Diffusion Models Trained with Large Data Are Transferable Visual Models
Guangkai Xu, Yongtao Ge, Mingyu Liu, Chengxiang Fan, Kangyang Xie, Zhiyue Zhao, Hao Chen, and Chunhua Shen. 2024 · 2024
Closest in time.