Fetching the paper…
Reading the bibliography…
Estimating geometry from dynamic scenes, where objects move and deform over time, remains a core challenge in computer vision.
Random sample consensus: a paradigm for model fitting with applications to image analysis and automated cartography
Martin A Fischler and Robert C. Bolles · 1981
Earlier work this paper cites.
Procrustes alignment with the EM algorithm
Bin Luo and Edwin R. Hancock · 1999
Earlier work this paper cites.
Multiple view geometry in computer vision
Richard Hartley and Andrew Zisserman · 2003
Earlier work this paper cites.
EPnP: An accurate O(n) solution to the PnP problem
Vincent Lepetit, Francesc Moreno-Noguer, and Pascal Fua · 2009
Earlier work this paper cites.
DTAM: Dense tracking and mapping in real-time
Richard A Newcombe, Steven J. Lovegrove, and Andrew J. Davison · 2011
Earlier work this paper cites.
A naturalistic open source movie for optical flow evaluation
Daniel J. Butler, Jonas Wulff, Garrett B. Stanley, and Michael J. Black · 2012
Earlier work this paper cites.
Indoor segmentation and support inference from RGBD images
Nathan Silberman, Derek Hoiem, Pushmeet Kohli, and Rob Fergus · 2012
Earlier work this paper cites.
A benchmark for the evaluation of RGB-D SLAM systems
Jürgen Sturm, Nikolas Engelhard, Felix Endres, Wolfram Burgard, and Daniel Cremers · 2012
Earlier work this paper cites.
Vision meets robotics: The KITTI dataset
Andreas Geiger, Philip Lenz, Christoph Stiller, and Raquel Urtasun · 2013
Earlier work this paper cites.
LSD-SLAM: Large-scale direct monocular SLAM
Jakob Engel, Thomas Schöps, and Daniel Cremers · 2014
Earlier work this paper cites.
Fast R-CNN
Ross Girshick · 2015
Earlier work this paper cites.
ORB-SLAM: a versatile and accurate monocular SLAM system
Raul Mur-Artal, Jose Maria Martinez Montiel, and Juan D Tardos · 2015
Earlier work this paper cites.
3D-R2N2: A unified approach for single and multi-view 3D object reconstruction
Christopher B. Choy, Danfei Xu, JunYoung Gwak, Kevin Chen, and Silvio Savarese · 2016
Earlier work this paper cites.
Temporally coherent 4D reconstruction of complex dynamic scenes
Armin Mustafa, Hansung Kim, Jean-Yves Guillemaut, and Adrian Hilton · 2016
Earlier work this paper cites.
A benchmark dataset and evaluation methodology for video object segmentation
Federico Perazzi, Jordi Pont-Tuset, Brian McWilliams, Luc Van Gool, Markus Gross, and Alexander Sorkine-Hornung · 2016
Earlier work this paper cites.
Structure-from-motion revisited
Johannes Lutz Schönberger and Jan-Michael Frahm · 2016
Earlier work this paper cites.
Pixelwise view selection for unstructured multi-view stereo
Johannes Lutz Schönberger, Enliang Zheng, Marc Pollefeys, and Jan-Michael Frahm · 2016
Earlier work this paper cites.
ScanNet: Richly-annotated 3D reconstructions of indoor scenes
Angela Dai, Angel X. Chang, Manolis Savva, Maciej Halber, Thomas Funkhouser, and Matthias Nießner · 2017
Earlier work this paper cites.
Direct sparse odometry
Jakob Engel, Vladlen Koltun, and Daniel Cremers · 2017
Earlier work this paper cites.
Monocular dense 3D reconstruction of a complex dynamic scene from two perspective frames
Suryansh Kumar, Yuchao Dai, and Hongdong Li · 2017
Earlier work this paper cites.
ORB-SLAM2: An open-source SLAM system for monocular, stereo, and RGB-D cameras
Raul Mur-Artal and Juan D. Tardós · 2017
Earlier work this paper cites.
Multi-view supervision for single-view reconstruction via differentiable ray consistency
Shubham Tulsiani, Tinghui Zhou, Alexei A. Efros, and Jitendra Malik · 2017
Earlier work this paper cites.
Robust dense mapping for large-scale dynamic environments
Ioan Andrei Bârsan, Peidong Liu, Marc Pollefeys, and Andreas Geiger · 2018
Earlier work this paper cites.
Learning efficient point cloud generation for dense 3D object reconstruction
Chen-Hsuan Lin, Chen Kong, and Simon Lucey · 2018
Cited alongside, same era.
Unsupervised learning of depth and ego-motion from monocular video using 3D geometric constraints
Reza Mahjourian, Martin Wicke, and Anelia Angelova · 2018
Cited alongside, same era.
BA-Net: Dense bundle adjustment network
Chengzhou Tang and Ping Tan · 2018
Cited alongside, same era.
DeepV2D: Video to depth with differentiable structure from motion
Zachary Teed and Jia Deng · 2018
Cited alongside, same era.
Pixel2Mesh: Generating 3D mesh models from single RGB images
Nanyang Wang, Yinda Zhang, Zhuwen Li, Yanwei Fu, Wei Liu, and Yu-Gang Jiang · 2018
Cited alongside, same era.
The temporal opportunist: Self-supervised multi-frame monocular depth
Jamie Watson, Oisin Mac Aodha, Victor Prisacariu, Gabriel Brostow, and Michael Firman · 2021
Later among the works it cites.
Learning to recover 3d scene shape from a single image
Wei Yin, Jianming Zhang, Oliver Wang, Simon Niklaus, Long Mai, Simon Chen, and Chunhua Shen · 2021
Later among the works it cites.
Consistent depth of moving objects in video
Zhoutong Zhang, Forrester Cole, Richard Tucker, William T Freeman, and Tali Dekel · 2021
Later among the works it cites.
CroCo: Self-supervised pre-training for 3D vision tasks by cross-view completion
Philippe Weinzaepfel, Vincent Leroy, Thomas Lucas, Romain Brégier, Yohann Cabon, Vaibhav Arora, Leonid Antsfeld, Boris Chidlovskii, Gabriela Csurka, and Jérôme Revaud · 2022
Later among the works it cites.
Structure and motion from casual videos
Zhoutong Zhang, Forrester Cole, Zhengqi Li, Michael Rubinstein, Noah Snavely, and William T. Freeman · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Every pixel counts: Unsupervised geometry learning with holistic 3D motion understanding
Zhenheng Yang, Peng Wang, Yang Wang, Wei Xu, and Ram Nevatia · 2018
Cited alongside, same era.
Learning implicit fields for generative shape modeling
Zhiqin Chen and Hao Zhang · 2019
Cited alongside, same era.
Mesh R-CNN
Georgia Gkioxari, Jitendra Malik, and Justin Johnson · 2019
Cited alongside, same era.
Digging into self-supervised monocular depth estimation
Clément Godard, Oisin Mac Aodha, Michael Firman, and Gabriel J Brostow · 2019
Cited alongside, same era.
Depth from videos in the wild: Unsupervised monocular depth learning from unknown cameras
Ariel Gordon, Hanhan Li, Rico Jonschkowski, and Anelia Angelova · 2019
Cited alongside, same era.
Refusion: 3d reconstruction in dynamic environments for RGB-D cameras exploiting residuals
Emanuele Palazzolo, Jens Behley, Philipp Lottes, Philippe Giguere, and Cyrill Stachniss · 2019
Cited alongside, same era.
DeepVoxels: Learning persistent 3D feature embeddings
Vincent Sitzmann, Justus Thies, Felix Heide, Matthias Nießner, Gordon Wetzstein, and Michael Zollhofer · 2019
Cited alongside, same era.
ParticleSfM: Exploiting dense point trajectories for localizing moving cameras in the wild
Wang Zhao, Shaohui Liu, Hengkai Guo, Wenping Wang, and Yong-Jin Liu · 2022
Later among the works it cites.
3d gaussian splatting for real-time radiance field rendering
Bernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, and George Drettakis · 2023
Later among the works it cites.
Spring: A high-resolution high-detail dataset and benchmark for scene flow, optical flow and stereo
Lukas Mehl, Jenny Schmalfuss, Azin Jahedi, Yaroslava Nalivayko, and Andrés Bruhn · 2023
Later among the works it cites.
DytanVO: Joint refinement of visual odometry and motion segmentation in dynamic environments
Shihao Shen, Yilin Cai, Wenshan Wang, and Sebastian Scherer · 2023
Later among the works it cites.
Sc-depthv3: Robust self-supervised monocular depth estimation for dynamic scenes
Libo Sun, Jia-Wang Bian, Huangying Zhan, Wei Yin, Ian Reid, and Chunhua Shen · 2023
Later among the works it cites.
Neural video depth stabilizer
Yiran Wang, Min Shi, Jiaqi Li, Zihao Huang, Zhiguo Cao, Jianming Zhang, Ke Xian, and Guosheng Lin · 2023
Later among the works it cites.
CroCo v2: Improved cross-view completion pre-training for stereo matching and optical flow
Philippe Weinzaepfel, Thomas Lucas, Vincent Leroy, Yohann Cabon, Vaibhav Arora, Romain Brégier, Gabriela Csurka, Leonid Antsfeld, Boris Chidlovskii, and Jérôme Revaud · 2023
Later among the works it cites.
PointOdyssey: A large-scale synthetic dataset for long-term point tracking
Yang Zheng, Adam W. Harley, Bokui Shen, Gordon Wetzstein, and Leonidas J. Guibas · 2023
Later among the works it cites.
LEAP-VO: Long-term effective any point tracking for visual odometry
Weirong Chen, Le Chen, Rui Wang, and Marc Pollefeys · 2024
Closest in time.
DreamScene4D: Dynamic multi-object scene generation from monocular videos
Wen-Hsuan Chu, Lei Ke, and Katerina Fragkiadaki · 2024
Closest in time.
DepthCrafter: Generating consistent long depth sequences for open-world videos
Wenbo Hu, Xiangjun Gao, Xiaoyu Li, Sijie Zhao, Xiaodong Cun, Yong Zhang, Long Quan, and Ying Shan · 2024
Closest in time.
Repurposing diffusion-based image generators for monocular depth estimation
Bingxin Ke, Anton Obukhov, Shengyu Huang, Nando Metzger, Rodrigo Caye Daudt, and Konrad Schindler · 2024
Closest in time.
MoSca: Dynamic gaussian fusion from casual videos via 4D motion scaffolds
Jiahui Lei, Yijia Weng, Adam Harley, Leonidas Guibas, and Kostas Daniilidis · 2024
Closest in time.
MoDGS: Dynamic gaussian splatting from causually-captured monocular videos
Qingming Liu, Yuan Liu, Jiepeng Wang, Xianqiang Lv, Peng Wang, Wenping Wang, and Junhui Hou · 2024
Closest in time.
The surprising effectiveness of diffusion models for optical flow and monocular depth estimation
Saurabh Saxena, Charles Herrmann, Junhwa Hur, Abhishek Kar, Mohammad Norouzi, Deqing Sun, and David J. Fleet · 2024
Closest in time.
Learning temporally consistent video depth from video diffusion priors
Jiahao Shao, Yuanbo Yang, Hongyu Zhou, Youmin Zhang, Yujun Shen, Matteo Poggi, and Yiyi Liao · 2024
Closest in time.
Deep patch visual odometry
Zachary Teed, Lahav Lipson, and Jia Deng · 2024
Closest in time.