Fetching the paper…
Reading the bibliography…
Temporal action detection is a very important yet challenging problem, since videos in real applications are usually long, untrimmed and contain multiple action instances.
Fast temporal activity proposals for efficient detection of human actions in untrimmed videos. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition
F. Caba Heilbron, J. Carlos Niebles, and B. Ghanem. 2016 · 1923
Earlier work this paper cites.
Convolutional two-stream network fusion for video action recognition. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition
C. Feichtenhofer, A. Pinz, and A. Zisserman. 2016 · 1941
Earlier work this paper cites.
Learning activity progression in LSTMs for activity detection and early detection. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition
S. Ma, L. Sigal, and S. Sclaroff. 2016 · 1950
Earlier work this paper cites.
A multi-stream bi-directional recurrent neural network for fine-grained action detection. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition
B. Singh, T. K. Marks, M. Jones, O. Tuzel, and M. Shao. 2016 · 1970
Earlier work this paper cites.
ImageNet: A large-scale hierarchical image database
J. Deng, W. Dong, R. Socher, L. Li, K. Li, and L. Feifei. 2009 · 2009
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks.. In Aistats
X. Glorot and Y. Bengio. 2010 · 2010
Earlier work this paper cites.
Deep Sparse Rectifier Neural Networks.. In Aistats
X. Glorot, A. Bordes, and Y. Bengio. 2011 · 2011
Earlier work this paper cites.
Action recognition by dense trajectories. In Computer Vision and Pattern Recognition (CVPR), 2011 IEEE Conference on
H. Wang, A. Kläser, C. Schmid, and C.-L. Liu. 2011 · 2011
Earlier work this paper cites.
UCF101: A dataset of 101 human actions classes from videos in the wild
K. Soomro, A. R. Zamir, and M. Shah. 2012 · 2012
Earlier work this paper cites.
HMDB51: A large video database for human motion recognition
H. Kuehne, H. Jhuang, R. Stiefelhagen, and T. Serre. 2013 · 2013
Earlier work this paper cites.
Action recognition with improved trajectories. In Proceedings of the IEEE International Conference on Computer Vision
H. Wang and C. Schmid. 2013 · 2013
Earlier work this paper cites.
Rich feature hierarchies for accurate object detection and semantic segmentation. In Proceedings of the IEEE conference on computer vision and pattern recognition
R. Girshick, J. Donahue, T. Darrell, and J. Malik. 2014 · 2014
Earlier work this paper cites.
Caffe: Convolutional architecture for fast feature embedding. In Proceedings of the 22nd ACM international conference on Multimedia
Y. Jia, E. Shelhamer, J. Donahue, S. Karayev, J. Long, R. Girshick, S. Guadarrama, and T. Darrell. 2014 · 2014
Earlier work this paper cites.
THUMOS challenge: Action recognition with a large number of classes. In ECCV Workshop
Y. G. Jiang, J. Liu, A. R. Zamir, G. Toderici, I. Laptev, M. Shah, and R. Sukthankar. 2014 · 2014
Earlier work this paper cites.
Fast saliency based pooling of fisher encoded dense trajectories. In ECCV THUMOS Workshop
S. Karaman, L. Seidenari, and A. Del Bimbo. 2014 · 2014
Earlier work this paper cites.
Large-scale video classification with convolutional neural networks. In Proceedings of the IEEE conference on Computer Vision and Pattern Recognition
A. Karpathy, G. Toderici, S. Shetty, T. Leung, R. Sukthankar, and L. Fei-Fei. 2014 · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
D. Kingma and J. Ba. 2014 · 2014
Cited alongside, same era.
The LEAR submission at Thumos 2014
D. Oneata, J. Verbeek, and C. Schmid. 2014 · 2014
Cited alongside, same era.
Two-stream convolutional networks for action recognition in videos. In Advances in Neural Information Processing Systems
K. Simonyan and A. Zisserman. 2014 · 2014
Cited alongside, same era.
Action recognition and detection by combining motion and appearance features
L. Wang, Y. Qiao, and X. Tang. 2014 · 2014
Cited alongside, same era.
Daps: Deep action proposals for action understanding. In European Conference on Computer Vision
V. Escorcia, F. C. Heilbron, J. C. Niebles, and B. Ghanem. 2016 · 2016
Later among the works it cites.
Deep residual learning for image recognition. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition
K. He, X. Zhang, S. Ren, and J. Sun. 2016 · 2016
Later among the works it cites.
Temporal Convolutional Networks: A Unified Approach to Action Segmentation. In Computer Vision–ECCV 2016 Workshops
C. Lea, R. Vidal, A. Reiter, and G. D. Hager. 2016 · 2016
Later among the works it cites.
SSD: Single shot multibox detector. In European Conference on Computer Vision
W. Liu, D. Anguelov, D. Erhan, C. Szegedy, S. Reed, C. Fu, and A. C. Berg. 2016 · 2016
Later among the works it cites.
Spot on: Action localization from pointly-supervised proposals. In European Conference on Computer Vision
P. Mettes, J. C. van Gemert, and C. G. Snoek. 2016 · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
MEXaction2
2015 · 2015
Cited alongside, same era.
Apt: Action localization proposals from dense trajectories
J. Gemert, M. Jain, E. Gati, C. G. Snoek, and others. 2015 · 2015
Cited alongside, same era.
Fast r-cnn. In Proceedings of the IEEE International Conference on Computer Vision
R. Girshick. 2015 · 2015
Cited alongside, same era.
Faster r-cnn: Towards real-time object detection with region proposal networks. In Advances in neural information processing systems
S. Ren, K. He, R. Girshick, and J. Sun. 2015 · 2015
Cited alongside, same era.
Very Deep Convolutional Networks for Large-Scale Image Recognition. In International Conference on Learning Representations
K. Simonyan and A. Zisserman. 2015 · 2015
Cited alongside, same era.
Going deeper with convolutions. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich. 2015 · 2015
Cited alongside, same era.
Learning spatiotemporal features with 3d convolutional networks. In Proceedings of the IEEE International Conference on Computer Vision
D. Tran, L. Bourdev, R. Fergus, L. Torresani, and M. Paluri. 2015 · 2015
Cited alongside, same era.
Deep Quantization: Encoding Convolutional Activations with Deep Generative Model
Z. Qiu, T. Yao, and T. Mei. 2016 · 2016
Later among the works it cites.
You only look once: Unified, real-time object detection. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition
J. Redmon, S. Divvala, R. Girshick, and A. Farhadi. 2016 · 2016
Later among the works it cites.
YOLO9000: Better, Faster, Stronger
J. Redmon and A. Farhadi. 2016 · 2016
Later among the works it cites.
Temporal action detection using a statistical language model. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition
A. Richard and J. Gall. 2016 · 2016
Later among the works it cites.
Temporal action localization in untrimmed videos via multi-stage cnns. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition
Z. Shou, D. Wang, and S.-F. Chang. 2016 · 2016
Later among the works it cites.
Untrimmed Video Classification for Activity Detection: submission to ActivityNet Challenge
G. Singh and F. Cuzzolin. 2016 · 2016
Later among the works it cites.
UTS at activitynet 2016
R. Wang and D. Tao. 2016 · 2016
Later among the works it cites.
End-to-end learning of action detection from frame glimpses in videos. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition
S. Yeung, O. Russakovsky, G. Mori, and L. Fei-Fei. 2016 · 2016
Later among the works it cites.
Temporal Action Localization with Pyramid of Score Distribution Features. In IEEE Conference on Computer Vision and Pattern Recognition
J. Yuan, B. Ni, X. Yang, and A. A. Kassim. 2016 · 2016
Later among the works it cites.
Efficient Action Detection in Untrimmed Videos via Multi-Task Learning
Y. Zhu and S. Newsam. 2016 · 2016
Later among the works it cites.