Fetching the paper…
Reading the bibliography…
In this paper, we propose a discriminative video representation for event detection over a large scale video dataset when only limited hardware resources are available.
Video google: A text retrieval approach to object matching in videos
J. Sivic and A. Zisserman · 2003
Earlier work this paper cites.
Distinctive image features from scale-invariant keypoints
D. G. Lowe · 2004
Earlier work this paper cites.
Histograms of oriented gradients for human detection
N. Dalal and B. Triggs · 2005
Earlier work this paper cites.
On space-time interest points
I. Laptev · 2005
Earlier work this paper cites.
Early versus late fusion in semantic video analysis
C. G. Snoek, M. Worring, and A. W. Smeulders · 2005
Earlier work this paper cites.
Mosift: Recognizing human actions in surveillance videos
M.-Y. Chen and A. Hauptmann · 2009
Earlier work this paper cites.
Aggregating local descriptors into a compact image representation
H. Jégou, M. Douze, C. Schmid, and P. Pérez · 2010
Earlier work this paper cites.
Improving the fisher kernel for large-scale image classification
F. Perronnin, J. Sánchez, and T. Mensink · 2010
Earlier work this paper cites.
Evaluating color descriptors for object and scene recognition
K. E. Van De Sande, T. Gevers, and C. G. Snoek · 2010
Earlier work this paper cites.
Vlfeat: An open and portable library of computer vision algorithms
A. Vedaldi and B. Fulkerson · 2010
Earlier work this paper cites.
Libsvm: a library for support vector machines
C.-C. Chang and C.-J. Lin · 2011
Earlier work this paper cites.
Product quantization for nearest neighbor search
H. Jegou, M. Douze, and C. Schmid · 2011
Earlier work this paper cites.
Action recognition by dense trajectories
H. Wang, A. Klaser, C. Schmid, and C.-L. Liu · 2011
Earlier work this paper cites.
Aggregating local image descriptors into compact codes
H. Jégou, F. Perronnin, M. Douze, J. Sánchez, P. Pérez, and C. Schmid · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Earlier work this paper cites.
Chebyshev approximations to the histogram
F. Li, G. Lebanon, and C. Sminchisescu · 2012
Cited alongside, same era.
Multimodal feature fusion for robust event detection in web videos
P. Natarajan, S. Wu, S. Vitaladevuni, X. Zhuang, S. Tsakalidis, U. Park, and R. Prasad · 2012
Cited alongside, same era.
AXES at TRECVid 2012: KIS, INS, and MED
D. Oneata, M. Douze, J. Revaud, S. Jochen, D. Potapov, H. Wang, Z. Harchaoui, J. Verbeek, C. Schmid, R. Aly, et al · 2012
Cited alongside, same era.
Evaluation of low-level features and their combinations for complex event detection in open source videos
A. Tamrakar, S. Ali, Q. Yu, J. Liu, O. Javed, A. Divakaran, H. Cheng, and H. Sawhney · 2012
Cited alongside, same era.
Efficient additive kernels via explicit feature maps
A. Vedaldi and A. Zisserman · 2012
Cited alongside, same era.
The AXES submissions at TrecVid 2013
R. Aly, R. Arandjelovic, K. Chatfield, M. Douze, B. Fernando, Z. Harchaoui, K. McGuinness, N. E. O’Connor, D. Oneata, O. M. Parkhi, et al · 2013
Action and event recognition with Fisher vectors on a compact feature set
D. Oneata, J. Verbeek, and C. Schmid · 2013
Later among the works it cites.
Image classification with the fisher vector: Theory and practice
J. Sánchez, F. Perronnin, T. Mensink, and J. Verbeek · 2013
Later among the works it cites.
Action recognition with improved trajectories
H. Wang and C. Schmid · 2013
Later among the works it cites.
Feature weighting via optimal thresholding for video analysis
Z. Xu, Y. Yang, I. Tsang, N. Sebe, and A. G. Hauptmann · 2013
Later among the works it cites.
How related exemplars help complex event detection in web videos?
Y. Yang, Z. Ma, Z. Xu, S. Yan, and A. G. Hauptmann · 2013
Later among the works it cites.
Return of the devil in the details: Delving deep into convolutional nets
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
All about VLAD
R. Arandjelović and A. Zisserman · 2013
Cited alongside, same era.
Segmentation driven object detection with Fisher vectors
R. G. Cinbis, J. Verbeek, and C. Schmid · 2013
Cited alongside, same era.
Stable hyper-pooling and query expansion for event detection
M. Douze, J. Revaud, C. Schmid, and H. Jégou · 2013
Cited alongside, same era.
Caffe: An open source convolutional architecture for fast feature embedding
Y. Jia · 2013
Cited alongside, same era.
CMU-Informedia at TRECVID 2013 Multimedia Event Detection
Z.-Z. Lan, L. Jiang, S.-I. Yu, et al · 2013
Cited alongside, same era.
M. Lin, Q. Chen, and S. Yan · 2013
Cited alongside, same era.
K. Chatfield, K. Simonyan, A. Vedaldi, and A. Zisserman · 2014
Closest in time.
Rich feature hierarchies for accurate object detection and semantic segmentation
R. Girshick, J. Donahue, T. Darrell, and J. Malik · 2014
Closest in time.
Multi-scale orderless pooling of deep convolutional activation features
Y. Gong, L. Wang, R. Guo, and S. Lazebnik · 2014
Closest in time.
Spatial pyramid pooling in deep convolutional networks for visual recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2014
Closest in time.
Efficient feature extraction, encoding and classification for action recognition
V. Kantorov and I. Laptev · 2014
Closest in time.
Large-scale video classification with convolutional neural networks
A. Karpathy, G. Toderici, S. Shetty, T. Leung, R. Sukthankar, and L. Fei-Fei · 2014
Closest in time.
Improved audio features for large-scale multimedia event detection
F. Metze, S. Rawat, and Y. Wang · 2014
Closest in time.
Bag of visual words and fusion methods for action recognition: Comprehensive study and good practice
X. Peng, L. Wang, X. Wang, and Y. Qiao · 2014
Closest in time.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2014
Closest in time.