Fetching the paper…
Reading the bibliography…
Action anticipation and forecasting in videos do not require a hat-trick, as far as there are signs in the context to foresee how actions are going to be deployed.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Recognizing realistic actions from videos ”in the wild”
J. Liu, J. Luo, and M. Shah · 2009
Earlier work this paper cites.
UT-Interaction Dataset, ICPR contest on Semantic Description of Human Activities (SDHA)
M. S. Ryoo and J. K. Aggarwal · 2010
Earlier work this paper cites.
Hmdb: a large video database for human motion recognition
H. Kuehne, H. Jhuang, E. Garrote, T. Poggio, and T. Serre · 2011
Earlier work this paper cites.
A large-scale benchmark dataset for event recognition in surveillance video
S. Oh, A. Hoogs, A. Perera, N. Cuntoor, C.-C. Chen, J. T. Lee, S. Mukherjee, J. Aggarwal, H. Lee, L. Davis, et al · 2011
Earlier work this paper cites.
Human activity prediction: Early recognition of ongoing activities from streaming videos
M. S. Ryoo · 2011
Earlier work this paper cites.
Activity forecasting
K. M. Kitani, B. D. Ziebart, J. A. Bagnell, and M. Hebert · 2012
Earlier work this paper cites.
A database for fine grained activity detection of cooking activities
M. Rohrbach, S. Amin, M. Andriluka, and B. Schiele · 2012
Earlier work this paper cites.
UCF101: A dataset of 101 human actions classes from videos in the wild
K. Soomro, A. Roshan Zamir, and M. Shah · 2012
Earlier work this paper cites.
Towards understanding action recognition
H. Jhuang, J. Gall, S. Zuffi, C. Schmid, and M. J. Black · 2013
Earlier work this paper cites.
Learning human activities and object affordances from rgb-d videos
H. S. Koppula, R. Gupta, and A. Saxena · 2013
Earlier work this paper cites.
Combining embedded accelerometers with computer vision for recognizing food preparation activities
S. Stein and S. J. McKenna · 2013
Earlier work this paper cites.
Context-aware activity forecasting
A. Chakraborty and A. K. Roy-Chowdhury · 2014
Earlier work this paper cites.
Anticipating human actions for collaboration in the presence of task and sensor uncertainty
K. P. Hawkins, S. Bansal, N. N. Vo, and A. F. Bobick · 2014
Earlier work this paper cites.
Large-scale video classification with convolutional neural networks
A. Karpathy, G. Toderici, S. Shetty, T. Leung, R. Sukthankar, and L. Fei-Fei · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2014
Earlier work this paper cites.
The language of actions: Recovering the syntax and semantics of goal-directed human activities
H. Kuehne, A. B. Arslan, and T. Serre · 2014
Cited alongside, same era.
A hierarchical representation for future action prediction
T. Lan, T.-C. Chen, and S. Savarese · 2014
Cited alongside, same era.
Learning spatiotemporal features with 3d convolutional networks
D. Tran, L. Bourdev, R. Fergus, L. Torresani, and M. Paluri · 2015
Cited alongside, same era.
Structural-rnn: Deep learning on spatio-temporal graphs
A. Jain, A. R. Zamir, S. Savarese, and A. Saxena · 2016
Cited alongside, same era.
Anticipating human activities using object affordances for reactive robotic response
H. S. Koppula and A. Saxena · 2016
Cited alongside, same era.
Learning activity progression in lstms for activity detection and early detection
S. Ma, L. Sigal, and S. Sclaroff · 2016
Joint prediction of activity labels and starting times in untrimmed videos
T. Mahmud, M. Hasan, and A. K. Roy-Chowdhury · 2017
Later among the works it cites.
Online real-time multiple spatiotemporal action localisation and prediction
G. Singh, S. Saha, M. Sapienza, P. H. Torr, and F. Cuzzolin · 2017
Later among the works it cites.
Prototypical networks for few-shot learning
J. Snell, K. Swersky, and R. Zemel · 2017
Later among the works it cites.
Generating the future with adversarial transformers
C. Vondrick and A. Torralba · 2017
Later among the works it cites.
Visual forecasting by imitating dynamics in natural sequences
K.-H. Zeng, W. B. Shen, D.-A. Huang, M. Sun, and J. C. Niebles · 2017
Later among the works it cites.
Scaling egocentric vision: The epic-kitchens dataset
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Hollywood in homes: Crowdsourcing data collection for activity understanding
G. A. Sigurdsson, G. Varol, X. Wang, A. Farhadi, I. Laptev, and A. Gupta · 2016
Cited alongside, same era.
Anticipating visual representations from unlabeled video
C. Vondrick, H. Pirsiavash, and A. Torralba · 2016
Cited alongside, same era.
An uncertain future: Forecasting from static images using variational autoencoders
J. Walker, C. Doersch, A. Gupta, and M. Hebert · 2016
Cited alongside, same era.
Encouraging lstms to anticipate actions very early
M. S. Aliakbarian, F. Saleh, M. Salzmann, B. Fernando, L. Petersson, and L. Andersson · 2017
Cited alongside, same era.
Deep representation learning for human motion prediction and classification
J. Bütepage, M. J. Black, D. Kragic, and H. Kjellström · 2017
Cited alongside, same era.
Realtime multi-person 2d pose estimation using part affinity fields
Z. Cao, T. Simon, S.-E. Wei, and Y. Sheikh · 2017
Cited alongside, same era.
D. Damen, H. Doughty, G. M. Farinella, S. Fidler, A. Furnari, E. Kazakos, D. Moltisanti, J. Munro, T. Perrett, W. Price, et al · 2018
Later among the works it cites.
When will you do what?-anticipating temporal occurrences of activities
Y. A. Farha, A. Richard, and J. Gall · 2018
Later among the works it cites.
Human action recognition and prediction: A survey
Y. Kong and Y. Fu · 2018
Later among the works it cites.
Action prediction from videos via memorizing hard-to-predict samples
Y. Kong, S. Gao, B. Sun, and Y. Fu · 2018
Later among the works it cites.
Global-local temporal saliency action prediction
S. Lai, W.-S. Zheng, J.-F. Hu, and J. Zhang · 2018
Later among the works it cites.
Action anticipation by predicting future dynamic images
C. Rodriguez, B. Fernando, and H. Li · 2018
Later among the works it cites.
Action anticipation with rbf kernelized feature mapping rnn
Y. Shi, B. Fernando, and R. Hartley · 2018
Later among the works it cites.
Actor and observer: Joint modeling of first and third-person videos
G. A. Sigurdsson, A. Gupta, C. Schmid, A. Farhadi, and K. Alahari · 2018
Later among the works it cites.
G. Singh, S. Saha, and F. Cuzzolin · 2018
Later among the works it cites.
Encoding crowd interaction with deep neural network for pedestrian trajectory prediction
Y. Xu, Z. Piao, and S. Gao · 2018
Later among the works it cites.