Fetching the paper…
Reading the bibliography…
DUSt3R has recently shown that one can reduce many tasks in multi-view geometry, including estimating camera intrinsics and extrinsics, reconstructing the scene in 3D, and establishing image correspondences, to the prediction of a pair of viewpoint-invariant point maps, i.e., pixel-aligned point clouds defined in a common reference frame.
An iterative image registration technique with an application to stereo vision
Bruce D. Lucas and Takeo Kanade · 1981
Earlier work this paper cites.
Motion and structure from motion from point and line matches
O. D. Faugeras, F. Lustman, and G. Toscani · 1987
Earlier work this paper cites.
Least-squares estimation of transformation parameters between two point patterns
Shinji Umeyama · 1991
Earlier work this paper cites.
Determining optical flow: A retrospective
Berthold K. P. Horn and Brian G. Schunck · 1993
Earlier work this paper cites.
Recovering non-rigid 3D shape from image streams
Christoph Bregler, Aaron Hertzmann, and Henning Biermann · 2000
Earlier work this paper cites.
Multiple View Geometry in Computer Vision
Richard Hartley and Andrew Zisserman · 2000
Earlier work this paper cites.
Multi-view matching for unordered image sets, or ”How do I organize my holiday snaps?”
Frederik Schaffalitzky and Andrew Zisserman · 2002
Earlier work this paper cites.
High accuracy optical flow estimation based on a theory for warping
Thomas Brox, Andrés Bruhn, Nils Papenberg, and Joachim Weickert · 2004
Earlier work this paper cites.
Learning non-rigid 3D shape from 2D motion
L. Torresani, A. Hertzmann, and C. Bregler · 2004
Earlier work this paper cites.
Lucas/kanade meets horn/schunck: Combining local and global optic flow methods
Andrés Bruhn, Joachim Weickert, and Christoph Schnörr · 2005
Earlier work this paper cites.
SURF: speeded up robust features
Herbert Bay, Tinne Tuytelaars, and Luc Van Gool · 2006
Earlier work this paper cites.
Nonrigid structure from motion in trajectory space
Ijaz Akhter, Yaser Sheikh, Sohaib Khan, and T. Kanade · 2008
Earlier work this paper cites.
Building rome in a day
Sameer Agarwal, Noah Snavely, Ian Simon, Steven M. Seitz, and Richard Szeliski · 2009
Earlier work this paper cites.
Large displacement optical flow
Thomas Brox, Christoph Bregler, and Jitendra Malik · 2009
Earlier work this paper cites.
Towards internet-scale multi-view stereo
Yasutaka Furukawa, Brian Curless, Steven M. Seitz, and Richard Szeliski · 2010
Earlier work this paper cites.
Trajectory Space: A dual representation for nonrigid structure from motion
Ijaz Akhter, Yaser Sheikh, Sohaib Khan, and T. Kanade · 2011
Earlier work this paper cites.
Trajectory triangulation: 3D motion reconstruction with
Mingyu Chen, G. Al-Regib, and B. Juang · 2011
Earlier work this paper cites.
General trajectory prior for non-rigid reconstruction
Jack Valmadre and Simon Lucey · 2012
Earlier work this paper cites.
Video pop-up: Monocular 3D reconstruction of dynamic scenes
Chris Russell, Rui Yu, and Lourdes Agapito · 2014
Earlier work this paper cites.
Flow fields: Dense correspondence fields for highly accurate large displacement optical flow estimation
Christian Bailer, Bertram Taetz, and Didier Stricker · 2015
Earlier work this paper cites.
FlowNet: Learning optical flow with convolutional networks
Alexey Dosovitskiy, Philipp Fischer, Eddy Ilg, Philip Häusser, Caner Hazirbas, Vladimir Golkov, Patrick van der Smagt, Daniel Cremers, and Thomas Brox · 2015
Cited alongside, same era.
FlowNet 2.0: Evolution of Optical Flow Estimation with Deep Networks
Eddy Ilg, Nikolaus Mayer, Tonmoy Saikia, Margret Keuper, Alexey Dosovitskiy, and Thomas Brox · 2016
Cited alongside, same era.
Dense monocular depth estimation in complex dynamic scenes
Rene Ranftl, Vibhav Vineet, Qifeng Chen, and Vladlen Koltun · 2016
Cited alongside, same era.
Structure-from-motion revisited
Johannes Lutz Schönberger and Jan-Michael Frahm · 2016
Cited alongside, same era.
LIFT: learned invariant feature transform
Kwang Moo Yi, Eduard Trulls, Vincent Lepetit, and Pascal Fua · 2016
Cited alongside, same era.
SuperPoint: self-supervised interest point detection and description
Dynamic view synthesis from dynamic monocular video
Chen Gao, Ayush Saraf, Johannes Kopf, and Jia-Bin Huang · 2021
Later among the works it cites.
Neural scene flow fields for space-time view synthesis of dynamic scenes
Zhengqi Li, Simon Niklaus, Noah Snavely, and Oliver Wang · 2021
Later among the works it cites.
D-NeRF: Neural radiance fields for dynamic scenes
Albert Pumarola, Enric Corona, Gerard Pons-Moll, and Francesc Moreno-Noguer · 2021
Later among the works it cites.
Space-time neural irradiance fields for free-viewpoint video
Wenqi Xian, Jia-Bin Huang, Johannes Kopf, and Changil Kim · 2021
Later among the works it cites.
Kubric: a scalable dataset generator
Klaus Greff, Francois Belletti, Lucas Beyer, Carl Doersch, Yilun Du, Daniel Duckworth, David J Fleet, Dan Gnanapragasam, Florian Golemo, Charles Herrmann, Thomas Kipf, Abhijit Kundu, Dmitry Lagun, Issam Laradji, Hsueh-Ti (Derek) Liu, Henning Meyer, Yishu Miao, Derek Nowrouzezahrai, Cengiz Oztireli, Etienne Pot, Noha Radwan, Daniel Rebain, Sara Sabour, Mehdi S. M. Sajjadi, Matan Sela, Vincent Sitzmann, Austin Stone, Deqing Sun, Suhani Vora, Ziyu Wang, Tianhao Wu, Kwang Moo Yi, Fangcheng Zhong, and Andrea Tagliasacchi · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
DeTone Daniel, Malisiewicz Tomasz, and Rabinovich Andrew · 2017
Cited alongside, same era.
What uncertainties do we need in Bayesian deep learning for computer vision?
Alex Kendall and Yarin Gal · 2017
Cited alongside, same era.
Monocular dense 3D reconstruction of a complex dynamic scene from two perspective frames
Suryansh Kumar, Yuchao Dai, and Hongdong Li · 2017
Cited alongside, same era.
Learning 3D object categories by looking around them
David Novotný, Diane Larlus, and Andrea Vedaldi · 2017
Cited alongside, same era.
DeMoN: depth and motion network for learning monocular stereo
Benjamin Ummenhofer, Huizhong Zhou, Jonas Uhrig, Nikolaus Mayer, Eddy Ilg, Alexey Dosovitskiy, and Thomas Brox · 2017
Cited alongside, same era.
Unsupervised learning of multi-frame optical flow with occlusions
J. Janai, F. Güney, A. Ranjan, M. Black, and A. Geiger · 2018
Cited alongside, same era.
A simple and effective fusion approach for multi-frame optical flow estimation
Zhile Ren, Orazio Gallo, Deqing Sun, Ming-Hsuan Yang, Erik B. Sudderth, and Jan Kautz · 2018
Cited alongside, same era.
FlowFormer: a transformer architecture for optical flow
Zhaoyang Huang, Xiaoyu Shi, Chao Zhang, Qiang Wang, Ka Chun Cheung, Hongwei Qin, Jifeng Dai, and Hongsheng Li · 2022
Later among the works it cites.
ClusterGNN: cluster-based coarse-to-fine graph neural network for efficient feature matching
Yan Shi, Jun-Xiong Cai, Yoli Shavit, Tai-Jiang Mu, Wensen Feng, and Kai Zhang · 2022
Later among the works it cites.
GMFlow: learning optical flow via global matching
Haofei Xu, Jing Zhang, Jianfei Cai, Hamid Rezatofighi, and Dacheng Tao · 2022
Later among the works it cites.
DynIBaR: Neural dynamic image-based rendering
Zhengqi Li, Qianqian Wang, Forrester Cole, Richard Tucker, and Noah Snavely · 2023
Later among the works it cites.
LightGlue: local feature matching at light speed
Philipp Lindenberger, Paul-Edouard Sarlin, and Marc Pollefeys · 2023
Later among the works it cites.
4D gaussian splatting for real-time dynamic scene rendering
Guanjun Wu, Taoran Yi, Jiemin Fang, Lingxi Xie, Xiaopeng Zhang, Wei Wei, Wenyu Liu, Qi Tian, and Xinggang Wang · 2023
Later among the works it cites.
Stereo4d: Learning how things move in 3d from internet stereo videos
Linyi Jin, Richard Tucker, Zhengqi Li, David Fouhey, Noah Snavely, and Aleksander Holynski · 2024
Later among the works it cites.
CoTracker: It is better to track together
Nikita Karaev, Ignacio Rocco, Ben Graham, Natalia Neverova, Andrea Vedaldi, and Christian Rupprecht · 2024
Later among the works it cites.
TAPVid-3D: a benchmark for tracking any point in 3D
Skanda Koppula, Ignacio Rocco, Yi Yang, Joe Heyward, João Carreira, Andrew Zisserman, Gabriel Brostow, and Carl Doersch · 2024
Later among the works it cites.
MoSca: dynamic gaussian fusion from casual videos via 4d motion scaffolds
Jiahui Lei, Yijia Weng, Adam Harley, Leonidas Guibas, and Kostas Daniilidis · 2024
Later among the works it cites.
Spatial cognition from egocentric video: Out of sight, not out of mind
Chiara Plizzari, Shubham Goel, Toby Perrett, Jacob Chalk, Angjoo Kanazawa, and Dima Damen · 2024
Later among the works it cites.
Dynamic Gaussian marbles for novel view synthesis of casual monocular videos
Colton Stearns, Adam Harley, Mikaela Uy, Florian Dubost, Federico Tombari, Gordon Wetzstein, and Leonidas Guibas · 2024
Later among the works it cites.
SpatialTracker: tracking any 2d pixels in 3d space
Yuxi Xiao, Qianqian Wang, Shangzhan Zhang, Nan Xue, Sida Peng, Yujun Shen, and Xiaowei Zhou · 2024
Later among the works it cites.
MonST3R: a simple approach for estimating geometry in the presence of motion
Junyi Zhang, Charles Herrmann, Junhwa Hur, Varun Jampani, Trevor Darrell, Forrester Cole, Deqing Sun, and Ming-Hsuan Yang · 2024
Later among the works it cites.
Continuous 3d perception model with persistent state
Qianqian Wang, Yifei Zhang, Aleksander Holynski, Alexei A Efros, and Angjoo Kanazawa · 2025
Closest in time.