Fetching the paper…
Reading the bibliography…
We present an unsupervised learning framework for the task of monocular depth and camera motion estimation from unstructured video sequences.
Hierarchical model-based motion estimation
J. Bergen, P. Anandan, K. Hanna, and R. Hingorani · 1992
Earlier work this paper cites.
View interpolation for image synthesis
S. E. Chen and L. Williams · 1993
Earlier work this paper cites.
Modeling and rendering architecture from photographs: A hybrid geometry-and image-based approach
P. E. Debevec, C. J. Taylor, and J. Malik · 1996
Earlier work this paper cites.
View morphing
S. M. Seitz and C. R. Dyer · 1996
Earlier work this paper cites.
About direct methods
M. Irani and P. Anandan · 1999
Earlier work this paper cites.
Prediction error as a quality metric for motion and stereo
R. Szeliski · 1999
Earlier work this paper cites.
Depth-image-based rendering (dibr), compression, and transmission for a new approach on 3d-tv
C. Fehn · 2004
Earlier work this paper cites.
High-quality video view interpolation using a layered representation
C. L. Zitnick, S. B. Kang, M. Uyttendaele, S. Winder, and R. Szeliski · 2004
Earlier work this paper cites.
Image-based rendering using image-based priors
A. Fitzgibbon, Y. Wexler, and A. Zisserman · 2005
Earlier work this paper cites.
Automatic photo pop-up
D. Hoiem, A. A. Efros, and M. Hebert · 2005
Earlier work this paper cites.
Make3D: Learning 3D scene structure from a single still image
A. Saxena, M. Sun, and A. Y. Ng · 2009
Earlier work this paper cites.
Towards internet-scale multi-view stereo
Y. Furukawa, B. Curless, S. M. Seitz, and R. Szeliski · 2010
Earlier work this paper cites.
DTAM: Dense tracking and mapping in real-time
R. A. Newcombe, S. J. Lovegrove, and A. J. Davison · 2011
Earlier work this paper cites.
VisualSFM: A visual structure from motion system
C. Wu · 2011
Earlier work this paper cites.
Are we ready for autonomous driving? The KITTI vision benchmark suite
A. Geiger, P. Lenz, and R. Urtasun · 2012
Earlier work this paper cites.
Depth map prediction from a single image using a multi-scale deep network
D. Eigen, C. Puhrsch, and R. Fergus · 2014
Earlier work this paper cites.
Depth transfer: Depth extraction from video using non-parametric sampling
K. Karsch, C. Liu, and S. B. Kang · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. Kingma and J. Ba · 2014
Earlier work this paper cites.
Discrete-continuous depth estimation from a single image
M. Liu, M. Salzmann, and X. He · 2014
Earlier work this paper cites.
OpenDR: An approximate differentiable renderer
M. M. Loper and M. J. Black · 2014
Cited alongside, same era.
Learning to see by moving
P. Agrawal, J. Carreira, and J. Malik · 2015
Cited alongside, same era.
Single image 3D without a single 3D image
D. F. Fouhey, W. Hussain, A. Gupta, and M. Hebert · 2015
Cited alongside, same era.
Unsupervised learning of spatiotemporally coherent metrics
R. Goroshin, J. Bruna, J. Tompson, D. Eigen, and Y. LeCun · 2015
Cited alongside, same era.
MatchNet: Unifying feature and metric learning for patch-based matching
X. Han, T. Leung, Y. Jia, R. Sukthankar, and A. C. Berg · 2015
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
S. Ioffe and C. Szegedy · 2015
Deeper depth prediction with fully convolutional residual networks
I. Laina, C. Rupprecht, V. Belagiannis, F. Tombari, and N. Navab · 2016
Later among the works it cites.
Learning depth from single monocular images using deep convolutional neural fields
F. Liu, C. Shen, G. Lin, and I. Reid · 2016
Later among the works it cites.
A large dataset to train convolutional networks for disparity, optical flow, and scene flow estimation
N. Mayer, E. Ilg, P. Hausser, P. Fischer, D. Cremers, A. Dosovitskiy, and T. Brox · 2016
Later among the works it cites.
Shuffle and learn: unsupervised learning using temporal order verification
I. Misra, C. L. Zitnick, and M. Hebert · 2016
Later among the works it cites.
Dense monocular depth estimation in complex dynamic scenes
R. Ranftl, V. Vineet, Q. Chen, and V. Koltun · 2016
Later among the works it cites.
Unsupervised learning of 3d structure from images
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Spatial transformer networks
M. Jaderberg, K. Simonyan, A. Zisserman, et al · 2015
Cited alongside, same era.
Learning image representations tied to egomotion
D. Jayaraman and K. Grauman · 2015
Cited alongside, same era.
PoseNet: A convolutional network for real-time 6-DOF camera relocalization
A. Kendall, M. Grimes, and R. Cipolla · 2015
Cited alongside, same era.
Deep convolutional inverse graphics network
T. D. Kulkarni, W. F. Whitney, P. Kohli, and J. Tenenbaum · 2015
Cited alongside, same era.
ORB-SLAM: a versatile and accurate monocular SLAM system
R. Mur-Artal, J. M. M. Montiel, and J. D. Tardos · 2015
Cited alongside, same era.
Unsupervised learning of visual representations using videos
X. Wang and A. Gupta · 2015
Cited alongside, same era.
D. J. Rezende, S. A. Eslami, S. Mohamed, P. Battaglia, M. Jaderberg, and N. Heess · 2016
Later among the works it cites.
Multi-view 3d models from single images with a convolutional network
M. Tatarchenko, A. Dosovitskiy, and T. Brox · 2016
Later among the works it cites.
DeMoN: Depth and motion network for learning monocular stereo
B. Ummenhofer, H. Zhou, J. Uhrig, N. Mayer, E. Ilg, A. Dosovitskiy, and T. Brox · 2016
Later among the works it cites.
Deep3D: Fully automatic 2D-to-3D video conversion with deep convolutional neural networks
J. Xie, R. B. Girshick, and A. Farhadi · 2016
Later among the works it cites.
Perspective transformer nets: Learning single-view 3d object reconstruction without 3d supervision
X. Yan, J. Yang, E. Yumer, Y. Guo, and H. Lee · 2016
Later among the works it cites.
Stereo matching by training a convolutional neural network to compare image patches
J. Zbontar and Y. LeCun · 2016
Later among the works it cites.
View synthesis by appearance flow
T. Zhou, S. Tulsiani, W. Sun, J. Malik, and A. A. Efros · 2016
Later among the works it cites.
Unsupervised monocular depth estimation with left-right consistency
C. Godard, O. Mac Aodha, and G. J. Brostow · 2017
Closest in time.
End-to-end learning of geometry and context for deep stereo regression
A. Kendall, H. Martirosyan, S. Dasgupta, P. Henry, R. Kennedy, A. Bachrach, and A. Bry · 2017
Closest in time.
Semi-supervised deep learning for monocular depth map prediction
Y. Kuznietsov, J. Stückler, and B. Leibe · 2017
Closest in time.
Learning features by watching objects move
D. Pathak, R. Girshick, P. Dollár, T. Darrell, and B. Hariharan · 2017
Closest in time.
Multi-view supervision for single-view reconstruction via differentiable ray consistency
S. Tulsiani, T. Zhou, A. A. Efros, and J. Malik · 2017
Closest in time.
SfM-Net: Learning of structure and motion from video
S. Vijayanarasimhan, S. Ricco, C. Schmid, R. Sukthankar, and K. Fragkiadaki · 2017
Closest in time.