Fetching the paper…
Reading the bibliography…
Deep learning has been demonstrated to achieve excellent results for image classification and object detection.
Event detection in crowded videos
Y. Ke, R. Sukthankar, and M. Hebert · 2007
Earlier work this paper cites.
A discriminatively trained, multiscale, deformable part model
P. Felzenszwalb, D. McAllester, and D. Ramanan · 2008
Earlier work this paper cites.
Action mach: a spatio-temporal maximum average correlation height filter for action recognition
M. Rodriguez, A. Javed, and M. Shah · 2008
Earlier work this paper cites.
Discriminative figure-centric models for joint action localization and recognition
T. Lan, Y. Wang, and G. Mori · 2011
Earlier work this paper cites.
Temporal localization of actions with actoms
A. Gaidon, Z. Harchaoui, and C. Schmid · 2013
Earlier work this paper cites.
Towards understanding action recognition
H. Jhuang, J. Gall, S. Zuffi, C. Schmid, and M. J. Black · 2013
Earlier work this paper cites.
3d convolutional neural networks for human action recognition
S. Ji, W. Xu, M. Yang, and K. Yu · 2013
Earlier work this paper cites.
THUMOS challenge: Action recognition with a large number of classes
Y.-G. Jiang, J. Liu, A. Roshan Zamir, I. Laptev, M. Piccardi, M. Shah, and R. Sukthankar · 2013
Earlier work this paper cites.
Spatiotemporal deformable part models for action detection
Y. Tian, R. Sukthankar, and M. Shah · 2013
Earlier work this paper cites.
Rich feature hierarchies for accurate object detection and semantic segmentation
R. Girshick, J. Donahue, T. Darrell, and J. Malik · 2014
Earlier work this paper cites.
Action localization with tubelets from motion
M. Jain, J. Van Gemert, H. Jégou, P. Bouthemy, and C. G. Snoek · 2014
Earlier work this paper cites.
Caffe: Convolutional architecture for fast feature embedding
Y. Jia, E. Shelhamer, J. Donahue, S. Karayev, J. Long, R. Girshick, S. Guadarrama, and T. Darrell · 2014
Cited alongside, same era.
THUMOS challenge: Action recognition with a large number of classes
Y.-G. Jiang, J. Liu, A. Roshan Zamir, G. Toderici, I. Laptev, M. Shah, and R. Sukthankar · 2014
Cited alongside, same era.
Large-scale video classification with convolutional neural networks
A. Karpathy, G. Toderici, S. Shetty, T. Leung, R. Sukthankar, and L. Fei-Fei · 2014
Cited alongside, same era.
Two-stream convolutional networks for action recognition in videos
K. Simonyan and A. Zisserman · 2014
Cited alongside, same era.
Video action detection with relational dynamic-poselets
L. Wang, Y. Qiao, and X. Tang · 2014
Cited alongside, same era.
Fast r-cnn
Action localization in videos through context walk
K. Soomro, H. Idrees, and M. Shah · 2015
Later among the works it cites.
Human action recognition using factorized spatio-temporal convolutional networks
L. Sun, K. Jia, D.-Y. Yeung, and B. E. Shi · 2015
Later among the works it cites.
Learning spatiotemporal features with 3d convolutional networks
D. Tran, L. Bourdev, R. Fergus, L. Torresani, and M. Paluri · 2015
Later among the works it cites.
Learning to track for spatio-temporal action localization
P. Weinzaepfel, Z. Harchaoui, and C. Schmid · 2015
Later among the works it cites.
Beyond short snippets: Deep networks for video classification
J. Yue-Hei Ng, M. Hausknecht, S. Vijayanarasimhan, O. Vinyals, R. Monga, and G. Toderici · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
R. Girshick · 2015
Cited alongside, same era.
Finding action tubes
G. Gkioxari and J. Malik · 2015
Cited alongside, same era.
What do 15,000 object categories tell us about classifying and localizing actions?
M. Jain, J. C. van Gemert, and C. G. Snoek · 2015
Cited alongside, same era.
Deep learning
Y. LeCun, Y. Bengio, and G. Hinton · 2015
Cited alongside, same era.
Faster r-cnn: Towards real-time object detection with region proposal networks
S. Ren, K. He, R. Girshick, and J. Sun · 2015
Cited alongside, same era.
R. Joseph and F. Ali · 2016
Later among the works it cites.
Human action recognition in videos: A survey
F. Negin and F. Bremond · 2016
Later among the works it cites.
Multi-region two-stream r-cnn for action detection
X. Peng and C. Schmid · 2016
Later among the works it cites.
Multi-region two-stream R-CNN for action detection
X. Peng and C. Schmid · 2016
Later among the works it cites.
What if we do not have multiple videos of the same action? – video action localization using web images
W. Sultani and M. Shah · 2016
Later among the works it cites.