Fetching the paper…
Reading the bibliography…
We conduct an in-depth exploration of different strategies for doing event detection in videos using convolutional neural networks (CNNs) trained for image classification.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
Probabilistic outputs for support vector machines and comparisons to regularized likelihood methods
J. C. Platt · 1999
Earlier work this paper cites.
Visual categorization with bags of keypoints
G. Csurka, C. Dance, L. Fan, J. Willamowski, and C. Bray · 2004
Earlier work this paper cites.
Distinctive image features from scale-invariant keypoints
D. Lowe · 2004
Earlier work this paper cites.
Histograms of oriented gradients for human detection
N. Dalal and B. Triggs · 2005
Earlier work this paper cites.
On space-time interest points
I. Laptev · 2005
Earlier work this paper cites.
SURF: Speeded up robust features
H. Bay, T. Tuytelaars, and L. Van Gool · 2006
Earlier work this paper cites.
Human detection using oriented histograms of flow and appearance
N. Dalal, B. Triggs, and C. Schmid · 2006
Earlier work this paper cites.
Beyond bags of features: Spatial pyramid matching for recognizing natural scene categories
S. Lazebnik, C. Schmid, and J. Ponce · 2006
Earlier work this paper cites.
Learning realistic human actions from movies
I. Laptev, M. Marszalek, C. Schmid, and B. Rozenfeld · 2008
Earlier work this paper cites.
An efficient dense and scale-invariant spatio-temporal interest point detector
G. Willems, T. Tuytelaars, and L. Van Gool · 2008
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Recognizing actions by shape-motion prototype trees
Z. Lin, Z. Jiang, and L. S. Davis · 2009
Earlier work this paper cites.
Evaluation of local spatio-temporal features for action recognition
H. Wang, M. M. Ullah, A. Kl�ser, I. Laptev, and C. Schmid · 2009
Cited alongside, same era.
Improving the Fisher kernel for large-scale image classification
F. Perronnin, J. Sánchez, and T. Mensink · 2010
Cited alongside, same era.
Action recognition by dense trajectories
H. Wang, A. Klaser, C. Schmid, and C.-L. Liu · 2011
Cited alongside, same era.
Three things everyone should know to improve object retrieval
R. Arandjelovic and A. Zisserman · 2012
Cited alongside, same era.
Trajectory-based modeling of human actions with motion reference points
Y.-G. Jiang, Q. Dai, X. Xue, W. Liu, and C.-W. Ngo · 2012
Cited alongside, same era.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Cited alongside, same era.
Multi-scale orderless pooling of deep convolutional activation features
Y. Gong, L. Wang, R. Guo, and S. Lazebnik · 2014
Later among the works it cites.
Spatial pyramid pooling in deep convolutional networks for visual recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2014
Later among the works it cites.
Large-scale video classification with convolutional neural networks
A. Karpathy, G. Toderici, S. Shetty, T. Leung, R. Sukthankar, and L. Fei-Fei · 2014
Later among the works it cites.
Beyond gaussian pyramid: Multi-skip feature stacking for action recognition
Z. Lan, M. Lin, X. Li, A. G. Hauptmann, and B. Raj · 2014
Later among the works it cites.
Learning and transferring mid-level image representations using convolutional neural networks
M. Oquab, L. Bottou, I. Laptev, and J. Sivic · 2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
UCF101: A dataset of 101 human actions classes from videos in the wild
K. Soomro, A. R. Zamir, and M. Shah · 2012
Cited alongside, same era.
Better exploiting motion for better action recognition
M. Jain, H. Jégou, and P. Bouthemy · 2013
Cited alongside, same era.
3d convolutional neural networks for human action recognition
S. Ji, W. Xu, M. Yang, and K. Yu · 2013
Cited alongside, same era.
Dense trajectories and motion boundary descriptors for action recognition
H. Wang, A. Kläser, C. Schmid, and C.-L. Liu · 2013
Cited alongside, same era.
Action recognition with improved trajectories
H. Wang and C. Schmid · 2013
Cited alongside, same era.
Bing: Binarized normed gradients for objectness estimation at 300fps
M.-M. Cheng, Z. Zhang, W.-Y. Lin, and P. Torr · 2014
Cited alongside, same era.
TRECVID 2014 – an overview of the goals, tasks, data, evaluation mechanisms and metrics
P. Over, G. Awad, M. Michel, J. Fiscus, G. Sanders, W. Kraaij, A. F. Smeaton, and G. Quéenot · 2014
Later among the works it cites.
CNN features off-the-shelf: an astounding baseline for recognition
A. S. Razavian, H. Azizpour, J. Sullivan, and S. Carlsson · 2014
Later among the works it cites.
Two-stream convolutional networks for action recognition in videos
K. Simonyan and A. Zisserman · 2014
Later among the works it cites.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2014
Later among the works it cites.
Dropout: A simple way to prevent neural networks from overfitting
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov · 2014
Later among the works it cites.
Going deeper with convolutions
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich · 2014
Later among the works it cites.
A discriminative CNN video representation for event detection
Z. Xu, Y. Yang, and A. G. Hauptmann · 2014
Later among the works it cites.
Beyond short snippets: Deep networks for video classification
J. Y.-H. Ng, M. Hausknecht, S. Vijayanarasimhan, O. Vinyals, R. Monga, and G. Toderici · 2015
Closest in time.