Fetching the paper…
Reading the bibliography…
Predicting the future to anticipate the outcome of events and actions is a critical attribute of autonomous agents; particularly for agents which must rely heavily on real time visual data for decision making.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Parsing video events with goal inference and intent prediction
M. Pei, Y. Jia, and S. C. Zhu · 2011
Earlier work this paper cites.
Max-margin early event detectors
M. Hoai and F. D. la Torre · 2012
Earlier work this paper cites.
Activity forecasting
K. M. Kitani, B. D. Ziebart, J. A. Bagnell, and M. Hebert · 2012
Earlier work this paper cites.
Predicting object dynamics in scenes
D. F. Fouhey and C. L. Zitnick · 2014
Earlier work this paper cites.
A hierarchical representation for future action prediction
T. Lan, T.-C. Chen, and S. Savarese · 2014
Earlier work this paper cites.
Video (language) modeling: a baseline for generative models of natural videos
M. Ranzato, A. Szlam, J. Bruna, M. Mathieu, R. Collobert, and S. Chopra · 2014
Earlier work this paper cites.
Unsupervised learning of video representations using lstms
N. Srivastava, E. Mansimov, and R. Salakhutdinov · 2015
Earlier work this paper cites.
Xception: Deep learning with depthwise separable convolutions
F. Chollet · 2016
Cited alongside, same era.
The cityscapes dataset for semantic urban scene understanding
M. Cordts, M. Omran, S. Ramos, T. Rehfeld, M. Enzweiler, R. Benenson, U. Franke, S. Roth, and B. Schiele · 2016
Cited alongside, same era.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Cited alongside, same era.
Video pixel networks
N. Kalchbrenner, A. van den Oord, K. Simonyan, I. Danihelka, O. Vinyals, A. Graves, and K. Kavukcuoglu · 2016
Cited alongside, same era.
Deep multi-scale video prediction beyond mean square error
M. Mathieu, C. Couprie, and Y. LeCun · 2016
Cited alongside, same era.
Video scene parsing with predictive feature learning
X. Jin, X. Li, H. Xiao, X. Shen, Z. Lin, J. Yang, Y. Chen, J. Dong, L. Liu, Z. Jie, J. Feng, and S. Yan · 2017
Later among the works it cites.
Predicting deeper into the future of semantic segmentation
P. Luc, N. Neverova, C. Couprie, J. Verbeek, and Y. LeCun · 2017
Later among the works it cites.
Geometry-based next frame prediction from monocular video
R. Mahjourian, M. Wicke, and A. Angelova · 2017
Later among the works it cites.
Decomposing motion and content for natural video sequence prediction
R. Villegas, J. Yang, S. Hong, X. Lin, and H. Lee · 2017
Later among the works it cites.
Unsupervised learning of depth and ego-motion from video
T. Zhou, M. Brown, N. Snavely, and D. G. Lowe · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. Shalev-Shwartz, N. Ben-Zrihem, A. Cohen, and A. Shashua · 2016
Cited alongside, same era.
On the sample complexity of end-to-end training vs. semantic abstraction training
S. Shalev-Shwartz and A. Shashua · 2016
Cited alongside, same era.
Se3-nets: Learning rigid body motion using deep neural networks
A. Byravan and D. Fox · 2017
Cited alongside, same era.
L. C. Chen, G. Papandreou, I. Kokkinos, K. Murphy, and A. L. Yuille · 2018
Closest in time.
Unsupervised learning of depth and ego-motion from monocular video using 3d geometric constraints
R. Mahjourian, M. Wicke, and A. Angelova · 2018
Closest in time.