Fetching the paper…
Reading the bibliography…
Deep ConvNets have shown its good performance in image classification tasks.
Random sample consensus: a paradigm for model fitting with applications to image analysis and automated cartography
M. A. Fischler and R. C. Bolles · 1981
Earlier work this paper cites.
Surf: Speeded up robust features
H. Bay, T. Tuytelaars, and L. Van Gool · 2006
Earlier work this paper cites.
A spatio-temporal descriptor based on 3d-gradients
A. Klaser, M. Marszałek, and C. Schmid · 2008
Earlier work this paper cites.
Learning realistic human actions from movies
I. Laptev, M. Marszałek, C. Schmid, and B. Rozenfeld · 2008
Earlier work this paper cites.
An efficient dense and scale-invariant spatio-temporal interest point detector
G. Willems, T. Tuytelaars, and L. Van Gool · 2008
Earlier work this paper cites.
Mosift: Recognizing human actions in surveillance videos
M.-y. Chen and A. Hauptmann · 2009
Earlier work this paper cites.
Evaluation of local spatio-temporal features for action recognition
H. Wang, M. M. Ullah, A. Klaser, I. Laptev, and C. Schmid · 2009
Earlier work this paper cites.
Spatial-bag-of-features
Y. Cao, C. Wang, Z. Li, L. Zhang, and L. Zhang · 2010
Earlier work this paper cites.
Aggregating local descriptors into a compact image representation
H. Jégou, M. Douze, C. Schmid, and P. Pérez · 2010
Earlier work this paper cites.
Human activity analysis: A review
J. K. Aggarwal and M. S. Ryoo · 2011
Earlier work this paper cites.
Libsvm: A library for support vector machines
C.-C. Chang and C.-J. Lin · 2011
Earlier work this paper cites.
Hmdb: a large video database for human motion recognition
H. Kuehne, H. Jhuang, E. Garrote, T. Poggio, and T. Serre · 2011
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
Action bank: A high-level representation of activity in video
S. Sadanand and J. J. Corso · 2012
Cited alongside, same era.
UCF101: A dataset of 101 human actions classes from videos in the wild
K. Soomro, A. R. Zamir, and M. Shah · 2012
Cited alongside, same era.
3d convolutional neural networks for human action recognition
S. Ji, W. Xu, M. Yang, and K. Yu · 2013
Cited alongside, same era.
Action recognition and localization by hierarchical space-time segments
S. Ma, J. Zhang, N. Ikizler-Cinbis, and S. Sclaroff · 2013
Cited alongside, same era.
Dense trajectories and motion boundary descriptors for action recognition
H. Wang, A. Kläser, C. Schmid, and C.-L. Liu · 2013
Cited alongside, same era.
Action recognition with improved trajectories
H. Wang and C. Schmid · 2013
The lear submission at thumos 2014
D. Oneata, J. Verbeek, and C. Schmid · 2014
Later among the works it cites.
Bag of visual words and fusion methods for action recognition: Comprehensive study and good practice
X. Peng, L. Wang, X. Wang, and Y. Qiao · 2014
Later among the works it cites.
M. Sapienza, F. Cuzzolin, and P. H. S. Torr · 2014
Later among the works it cites.
Two-stream convolutional networks for action recognition in videos
K. Simonyan and A. Zisserman · 2014
Later among the works it cites.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Lear-inria submission for the thumos workshop
H. Wang and C. Schmid · 2013
Cited alongside, same era.
Multi-view super vector for action recognition
Z. Cai, L. Wang, X. Peng, and Y. Qiao · 2014
Cited alongside, same era.
Decaf: A deep convolutional activation feature for generic visual recognition
J. Donahue, Y. Jia, O. Vinyals, J. Hoffman, N. Zhang, E. Tzeng, and T. Darrell · 2014
Cited alongside, same era.
University of amsterdam at thumos challenge 2014
M. Jain, J. van Gemert, and C. G. Snoek · 2014
Cited alongside, same era.
Thumos challenge: Action recognition with a large number of classes
Y. Jiang, J. Liu, A. Roshan Zamir, G. Toderici, I. Laptev, M. Shah, and R. Sukthankar · 2014
Cited alongside, same era.
Efficient feature extraction, encoding, and classification for action recognition
V. Kantorov and I. Laptev · 2014
Cited alongside, same era.
The treasure beneath convolutional layers: Cross-convolutional-layer pooling for image classification
L. Liu, C. Shen, and A. van den Hengel · 2015
Closest in time.
Beyond short snippets: Deep networks for video classification
J. Y. Ng, M. J. Hausknecht, S. Vijayanarasimhan, O. Vinyals, R. Monga, and G. Toderici · 2015
Closest in time.
Exploiting local features from deep networks for image retrieval
J. Y. Ng, F. Yang, and L. S. Davis · 2015
Closest in time.
Going deeper with convolutions
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich · 2015
Closest in time.
Action recognition with trajectory-pooled deep-convolutional descriptors
L. Wang, Y. Qiao, and X. Tang · 2015
Closest in time.
Towards good practices for very deep two-stream convnets
L. Wang, Y. Xiong, Z. Wang, and Y. Qiao · 2015
Closest in time.
A discriminative CNN video representation for event detection
Z. Xu, Y. Yang, and A. G. Hauptmann · 2015
Closest in time.