Fetching the paper…
Reading the bibliography…
Visual features are of vital importance for human action understanding in videos.
Random sample consensus: A paradigm for model fitting with applications to image analysis and automated cartography
M. A. Fischler and R. C. Bolles · 1981
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 2001
Earlier work this paper cites.
Distinctive image features from scale-invariant keypoints
D. G. Lowe · 2004
Earlier work this paper cites.
Histograms of oriented gradients for human detection
N. Dalal and B. Triggs · 2005
Earlier work this paper cites.
Behavior recognition via sparse spatio-temporal features
P. Dollár, V. Rabaud, G. Cottrell, and S. Belongie · 2005
Earlier work this paper cites.
On space-time interest points
I. Laptev · 2005
Earlier work this paper cites.
SURF: speeded up robust features
H. Bay, T. Tuytelaars, and L. J. V. Gool · 2006
Earlier work this paper cites.
A duality based approach for realtime tv-
C. Zach, T. Pock, and H. Bischof · 2007
Earlier work this paper cites.
A spatio-temporal descriptor based on 3D-gradients
A. Kläser, M. Marszalek, and C. Schmid · 2008
Earlier work this paper cites.
Learning realistic human actions from movies
I. Laptev, M. Marszalek, C. Schmid, and B. Rozenfeld · 2008
Earlier work this paper cites.
An efficient dense and scale-invariant spatio-temporal interest point detector
G. Willems, T. Tuytelaars, and L. J. V. Gool · 2008
Earlier work this paper cites.
ImageNet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L. Li, K. Li, and F. Li · 2009
Earlier work this paper cites.
Evaluation of local spatio-temporal features for action recognition
H. Wang, M. M. Ullah, A. Kläser, I. Laptev, and C. Schmid · 2009
Earlier work this paper cites.
Convolutional learning of spatio-temporal features
G. W. Taylor, R. Fergus, Y. LeCun, and C. Bregler · 2010
Earlier work this paper cites.
Human activity analysis: A review
J. K. Aggarwal and M. S. Ryoo · 2011
Cited alongside, same era.
HMDB: A large video database for human motion recognition
H. Kuehne, H. Jhuang, E. Garrote, T. Poggio, and T. Serre · 2011
Cited alongside, same era.
ImageNet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Cited alongside, same era.
Structured learning of human interactions in TV shows
A. Patron-Perez, M. Marszalek, I. Reid, and A. Zisserman · 2012
Cited alongside, same era.
Action bank: A high-level representation of activity in video
S. Sadanand and J. J. Corso · 2012
Cited alongside, same era.
UCF101: A dataset of 101 human actions classes from videos in the wild
K. Soomro, A. R. Zamir, and M. Shah · 2012
Mining motion atoms and phrases for complex action recognition
L. Wang, Y. Qiao, and X. Tang · 2013
Later among the works it cites.
Motionlets: Mid-level 3D parts for human motion recognition
L. Wang, Y. Qiao, and X. Tang · 2013
Later among the works it cites.
Action recognition with actons
J. Zhu, B. Wang, X. Yang, W. Zhang, and Z. Tu · 2013
Later among the works it cites.
Multi-view super vector for action recognition
Z. Cai, L. Wang, X. Peng, and Y. Qiao · 2014
Later among the works it cites.
Return of the devil in the details: Delving deep into convolutional nets
K. Chatfield, K. Simonyan, A. Vedaldi, and A. Zisserman · 2014
Later among the works it cites.
Large-scale video classification with convolutional neural networks
A. Karpathy, G. Toderici, S. Shetty, T. Leung, R. Sukthankar, and L. Fei-Fei · 2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
A comparative study of encoding, pooling and normalization methods for action recognition
X. Wang, L. Wang, and Y. Qiao · 2012
Cited alongside, same era.
3D convolutional neural networks for human action recognition
S. Ji, W. Xu, M. Yang, and K. Yu · 2013
Cited alongside, same era.
THUMOS challenge: Action recognition with a large number of classes, 2013
Y.-G. Jiang, J. Liu, A. Roshan Zamir, I. Laptev, M. Piccardi, M. Shah, and R. Sukthankar · 2013
Cited alongside, same era.
Image classification with the Fisher vector: Theory and practice
J. Sánchez, F. Perronnin, T. Mensink, and J. J. Verbeek · 2013
Cited alongside, same era.
Large-scale web video event classification by use of Fisher vectors
C. Sun and R. Nevatia · 2013
Cited alongside, same era.
Dense trajectories and motion boundary descriptors for action recognition
H. Wang, A. Kläser, C. Schmid, and C.-L. Liu · 2013
Cited alongside, same era.
Bag of visual words and fusion methods for action recognition: Comprehensive study and good practice
X. Peng, L. Wang, X. Wang, and Y. Qiao · 2014
Later among the works it cites.
Two-stream convolutional networks for action recognition in videos
K. Simonyan and A. Zisserman · 2014
Later among the works it cites.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2014
Later among the works it cites.
Going deeper with convolutions
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich · 2014
Later among the works it cites.
Latent hierarchical model of temporal structure for complex activity classification
L. Wang, Y. Qiao, and X. Tang · 2014
Later among the works it cites.
Video action detection with relational dynamic-poselets
L. Wang, Y. Qiao, and X. Tang · 2014
Later among the works it cites.
Visualizing and understanding convolutional networks
M. D. Zeiler and R. Fergus · 2014
Later among the works it cites.