Fetching the paper…
Reading the bibliography…
In this paper, we present DM-Calib, a diffusion-based approach for estimating pinhole camera intrinsic parameters from a single input image.
Euclidean reconstruction from uncalibrated views
Richard I Hartley · 1993
Earlier work this paper cites.
Camera Self-Calibration from Video Sequences: the Kruppa Equations Revisited
Cyril Zeller and Olivier Faugeras · 1996
Earlier work this paper cites.
Kruppa’s equations derived from the fundamental matrix
Richard I. Hartley · 1997
Earlier work this paper cites.
Self-calibration of a moving camera from point correspondences and fundamental matrices
Quang-Tuan Luong and Olivier D. Faugeras · 1997
Earlier work this paper cites.
A stratified approach to metric self-calibration
Marc Pollefeys and Luc Van Gool · 1997
Earlier work this paper cites.
Autocalibration and the absolute quadric
Bill Triggs · 1997
Earlier work this paper cites.
Manhattan world: Compass direction from a single image by bayesian inference
James M Coughlan and Alan L Yuille · 1999
Earlier work this paper cites.
Procrustes alignment with the em algorithm
Bin Luo and Edwin R Hancock · 1999
Earlier work this paper cites.
A flexible new technique for camera calibration
Zhengyou Zhang · 2000
Earlier work this paper cites.
A general imaging model and a method for finding its parameters
Michael D Grossberg and Shree K Nayar · 2001
Earlier work this paper cites.
An expectation maximization framework for simultaneous low-level edge grouping and camera calibration in complex man-made environments
Grant Schindler et al · 2004
Earlier work this paper cites.
A comparison and evaluation of multi-view stereo reconstruction algorithms
Steven M Seitz, Brian Curless, James Diebel, Daniel Scharstein, and Richard Szeliski · 2006
Earlier work this paper cites.
Camera calibration from images of spheres
Hui Zhang, K Wong Kwan-yee, and Guoqiang Zhang · 2007
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Symmetric architecture modeling with a single image
Nianjuan Jiang, Ping Tan, and Loong-Fah Cheong · 2009
Earlier work this paper cites.
Google street view: Capturing the world at street level
Dragomir Anguelov, Carole Dulong, Daniel Filip, Christian Frueh, Stéphane Lafon, Richard Lyon, Abhijit Ogale, Luc Vincent, and Josh Weaver · 2010
Earlier work this paper cites.
Are we ready for autonomous driving? The KITTI vision benchmark suite
Andreas Geiger, Philip Lenz, and Raquel Urtasun · 2012
Earlier work this paper cites.
Indoor segmentation and support inference from rgbd images
Pushmeet Kohli Nathan Silberman, Derek Hoiem and Rob Fergus · 2012
Earlier work this paper cites.
A benchmark for the evaluation of rgb-d slam systems
Jürgen Sturm, Nikolas Engelhard, Felix Endres, Wolfram Burgard, and Daniel Cremers · 2012
Earlier work this paper cites.
Robust camera self-calibration from monocular images of manhattan worlds
Horst Wildenauer and Allan Hanbury · 2012
Earlier work this paper cites.
Auto-encoding variational bayes
Diederik P Kingma · 2013
Earlier work this paper cites.
Automatic upright adjustment of photographs with robust camera calibration
Hyunjoon Lee, Eli Shechtman, Jue Wang, and Seungyong Lee · 2013
Earlier work this paper cites.
Sun3d: A database of big spaces reconstructed using sfm and object labels
Jianxiong Xiao, Andrew Owens, and Antonio Torralba · 2013
Earlier work this paper cites.
Mve-a multi-view reconstruction environment
Simon Fuhrmann, Fabian Langguth, and Michael Goesele · 2014
Earlier work this paper cites.
Shapenet: An information-rich 3d model repository
Angel X Chang, Thomas Funkhouser, Leonidas Guibas, Pat Hanrahan, Qixing Huang, Zimo Li, Silvio Savarese, Manolis Savva, Shuran Song, Hao Su, et al · 2015
Earlier work this paper cites.
The cityscapes dataset for semantic urban scene understanding
Marius Cordts, Mohamed Omran, Sebastian Ramos, Timo Rehfeld, Markus Enzweiler, Rodrigo Benenson, Uwe Franke, Stefan Roth, and Bernt Schiele · 2016
Earlier work this paper cites.
Structure-from-motion revisited
Johannes Lutz Schönberger and Jan-Michael Frahm · 2016
Earlier work this paper cites.
Detecting vanishing points using global image context in a non-manhattan world
Menghua Zhai, Scott Workman, and Nathan Jacobs · 2016
Earlier work this paper cites.
A flexible online camera calibration using line segments
Yueqiang Zhang, Langming Zhou, Haibo Liu, and Yang Shang · 2016
Earlier work this paper cites.
Scannet: Richly-annotated 3d reconstructions of indoor scenes
Angela Dai, Angel X. Chang, Manolis Savva, Maciej Halber, Thomas Funkhouser, and Matthias Nießner · 2017
Earlier work this paper cites.
A multi-view stereo benchmark with high-resolution images and multi-camera videos
Thomas Schöps, Johannes L. Schönberger, Silvano Galliani, Torsten Sattler, Konrad Schindler, Marc Pollefeys, and Andreas Geiger · 2017
Cited alongside, same era.
Guillermo Gallego, Elias Mueggler, and Peter F. Sturm · 2018
Cited alongside, same era.
A-contrario horizon-first vanishing point detection using second-order grouping laws
Gilles Simon, Antoine Fond, and Marie-Odile Berger · 2018
Cited alongside, same era.
Taskonomy: Disentangling task transfer learning
Amir R Zamir, Alexander Sax, William B Shen, Leonidas Guibas, Jitendra Malik, and Silvio Savarese · 2018
Cited alongside, same era.
DIODE: A dense indoor and outdoor depth dataset
Igor Vasiljevic, Nicholas I. Kolkin, Shanyi Zhang, Ruotian Luo, Haochen Wang, Falcon Z. Dai, Andrea F. Daniele, Mohammadreza Mostajabi, Steven Basart, Matthew R. Walter, and Gregory Shakhnarovich · 2019
Laion-5b: An open large-scale dataset for training next generation image-text models
Christoph Schuhmann, Romain Beaumont, Richard Vencu, Cade Gordon, Ross Wightman, Mehdi Cherti, Theo Coombes, Aarush Katta, Clayton Mullis, Mitchell Wortsman, et al · 2022
Later among the works it cites.
Towards accurate reconstruction of 3d scene shape from a single monocular image
Wei Yin, Jianming Zhang, Oliver Wang, Simon Niklaus, Simon Chen, Yifan Liu, and Chunhua Shen · 2022
Later among the works it cites.
Hierarchical normalization for robust monocular depth estimation
Chi Zhang, Wei Yin, Billzb Wang, Gang Yu, Bin Fu, and Chunhua Shen · 2022
Later among the works it cites.
Zoedepth: Zero-shot transfer by combining relative and metric depth
Shariq Farooq Bhat, Reiner Birkl, Diana Wofk, Peter Wonka, and Matthias Müller · 2023
Later among the works it cites.
Camera self-calibration using human faces
Masa Hu, Garrick Brazil, Nanxiang Li, Liu Ren, and Xiaoming Liu · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Yohann Cabon, Naila Murray, and Martin Humenberger · 2020
Cited alongside, same era.
nuscenes: A multimodal dataset for autonomous driving
Holger Caesar, Varun Bankiti, Alex H. Lang, Sourabh Vora, Venice Erin Liong, Qiang Xu, Anush Krishnan, Yu Pan, Giancarlo Baldan, and Oscar Beijbom · 2020
Cited alongside, same era.
Denoising diffusion probabilistic models
Jonathan Ho, Ajay Jain, and Pieter Abbeel · 2020
Cited alongside, same era.
Comparison of monocular depth estimation methods using geometrically relevant metrics on the IBims-1 dataset
Tobias Koch, Lukas Liebel, Marco Körner, and Friedrich Fraundorfer · 2020
Cited alongside, same era.
Towards robust monocular depth estimation: Mixing datasets for zero-shot cross-dataset transfer
René Ranftl, Katrin Lasinger, David Hafner, Konrad Schindler, and Vladlen Koltun · 2020
Cited alongside, same era.
Why having 10,000 parameters in your camera model is better than twelve
Thomas Schops, Viktor Larsson, Marc Pollefeys, and Torsten Sattler · 2020
Cited alongside, same era.
Denoising diffusion implicit models
Jiaming Song, Chenlin Meng, and Stefano Ermon · 2020
Cited alongside, same era.
Perspective fields for single image camera calibration
Linyi Jin, Jianming Zhang, Yannick Hold-Geoffroy, Oliver Wang, Kevin Blackburn-Matzen, Matthew Sticha, and David F Fouhey · 2023
Later among the works it cites.
Multi-resolution noise for diffusion model training
Kasiopy · 2023
Later among the works it cites.
Camera self-calibration network based on face shape estimation
Yunxiang Liu and Ye Cui · 2023
Later among the works it cites.
Latent consistency models: Synthesizing high-resolution images with few-step inference, 2023
Simian Luo, Yiqin Tan, Longbo Huang, Jian Li, and Hang Zhao · 2023
Later among the works it cites.
iDisc: Internal discretization for monocular depth estimation
Luigi Piccinelli, Christos Sakaridis, and Fisher Yu · 2023
Later among the works it cites.
Metric3d: Towards zero-shot metric 3d prediction from a single image
Wei Yin, Chi Zhang, Hao Chen, Zhipeng Cai, Gang Yu, Kaixuan Wang, Xiaozhi Chen, and Chunhua Shen · 2023
Later among the works it cites.
Mvimgnet: A large-scale dataset of multi-view images
Xianggang Yu, Mutian Xu, Yidan Zhang, Haolin Liu, Chongjie Ye, Yushuang Wu, Zizheng Yan, Chenming Zhu, Zhangyang Xiong, Tianyou Liang, et al · 2023
Later among the works it cites.
Adding conditional control to text-to-image diffusion models
Lvmin Zhang, Anyi Rao, and Maneesh Agrawala · 2023
Later among the works it cites.
Tame a wild camera: In-the-wild monocular camera calibration
Shengjie Zhu, Abhinav Kumar, Masa Hu, and Xiaoming Liu · 2023
Later among the works it cites.
Stylegan knows normal, depth, albedo, and more
Anand Bhattad, Daniel McKee, Derek Hoiem, and David Forsyth · 2024
Closest in time.
Geowizard: Unleashing the diffusion priors for 3d geometry estimation from a single image
Xiao Fu, Wei Yin, Mu Hu, Kaixuan Wang, Yuexin Ma, Ping Tan, Shaojie Shen, Dahua Lin, and Xiaoxiao Long · 2024
Closest in time.
Fine-tuning image-conditional diffusion models is easier than you think
Gonzalo Martin Garcia, Karim Abou Zeid, Christian Schmidt, Daan de Geus, Alexander Hermans, and Bastian Leibe · 2024
Closest in time.
Diffcalib: Reformulating monocular camera calibration as diffusion-based dense incident map generation
Xiankang He, Guangkai Xu, Bo Zhang, Hao Chen, Ying Cui, and Dongyan Guo · 2024
Closest in time.
Repurposing diffusion-based image generators for monocular depth estimation
Bingxin Ke, Anton Obukhov, Shengyu Huang, Nando Metzger, Rodrigo Caye Daudt, and Konrad Schindler · 2024
Closest in time.
Wonder3d: Single image to 3d using cross-domain diffusion
Xiaoxiao Long, Yuan-Chen Guo, Cheng Lin, Yuan Liu, Zhiyang Dou, Lingjie Liu, Yuexin Ma, Song-Hai Zhang, Marc Habermann, Christian Theobalt, et al · 2024
Closest in time.
Unidepth: Universal monocular metric depth estimation
Luigi Piccinelli, Yung-Hsu Yang, Christos Sakaridis, Mattia Segu, Siyuan Li, Luc Van Gool, and Fisher Yu · 2024
Closest in time.
Geocalib: Learning single-image calibration with geometric optimization
Alexander Veicht, Paul-Edouard Sarlin, Philipp Lindenberger, and Marc Pollefeys · 2024
Closest in time.
Dust3r: Geometric 3d vision made easy
Shuzhe Wang, Vincent Leroy, Yohann Cabon, Boris Chidlovskii, and Jerome Revaud · 2024
Closest in time.
Diffusion models trained with large data are transferable visual models
Guangkai Xu, Yongtao Ge, Mingyu Liu, Chengxiang Fan, Kangyang Xie, Zhiyue Zhao, Hao Chen, and Chunhua Shen · 2024
Closest in time.
Depth anything: Unleashing the power of large-scale unlabeled data
Lihe Yang, Bingyi Kang, Zilong Huang, Xiaogang Xu, Jiashi Feng, and Hengshuang Zhao · 2024
Closest in time.
Stablenormal: Reducing diffusion variance for stable and sharp normal
Chongjie Ye, Lingteng Qiu, Xiaodong Gu, Qi Zuo, Yushuang Wu, Zilong Dong, Liefeng Bo, Yuliang Xiu, and Xiaoguang Han · 2024
Closest in time.
Cameras as rays: Pose estimation via ray diffusion
Jason Y Zhang, Amy Lin, Moneish Kumar, Tzu-Hsuan Yang, Deva Ramanan, and Shubham Tulsiani · 2024
Closest in time.
Depth anything at any condition
Boyuan Sun, Modi Jin, Bowen Yin, and Qibin Hou · 2025
Closest in time.
Geocalib: Learning single-image calibration with geometric optimization
Alexander Veicht, Paul-Edouard Sarlin, Philipp Lindenberger, and Marc Pollefeys · 2025
Closest in time.