Fetching the paper…
Reading the bibliography…
This paper studies the joint learning of action recognition and temporal localization in long, untrimmed videos.
Simple Statistical Gradient-following Algorithms for Connectionist Reinforcement Learning
R. J. Williams · 1992
Earlier work this paper cites.
Long Short-Term Memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Discriminative Figure-Centric Models for Joint Action Localization and Recognition
T. Lan, Y. Wang, , and G. Mori · 2001
Earlier work this paper cites.
ImageNet: A Large-Scale Hierarchical Image Database
J. Deng, W. Dong, R. Socher, L.-J. Li, K. Li, and L. Fei-Fei · 2009
Earlier work this paper cites.
Discriminative Video Pattern Search for Efficient Action Detection
J. Yuan, Z. Liu, and Y. Wu · 2011
Earlier work this paper cites.
ImageNet Classification with Deep Convolutional Neural Networks
A. Krizhevsky, I. Sutskever, and G. Hinton · 2012
Earlier work this paper cites.
Reconstructing 3D Human Pose from 2D Image Landmarks
V. Ramakrishna, T. Kanade, and Y. Sheikh · 2012
Earlier work this paper cites.
A Database for Fine Grained Activity Detection of Cooking Activities
M. Rohrbach, S. Amin, M. Andriluka, and B. Schiele · 2012
Earlier work this paper cites.
Action and Event Recognition with Fisher Vectors on a Compact Feature Set
D. Oneata, J. Verbeek, and C. Schmid · 2013
Earlier work this paper cites.
Action Recognition with Improved Trajectories
H. Wang and C. Schmid · 2013
Earlier work this paper cites.
Actionness Ranking with Lattice Conditional Ordinal Random Fields
W. Chen, C. Xiong, R. Xu, and J. J. Corso · 2014
Earlier work this paper cites.
Action Localization with Tubelets from Motion
M. Jain, J. van Gemert, H. Jégou, P. Bouthemy, and C. G. M. Snoek · 2014
Earlier work this paper cites.
THUMOS challenge: Action recognition with a large number of classes
Y.-G. Jiang, J. Liu, A. Roshan Zamir, G. Toderici, I. Laptev, M. Shah, and R. Sukthankar · 2014
Earlier work this paper cites.
Fast Saliency Based Pooling of Fisher Encoded Dense Trajectories, 2014
S. Karaman, L. Seidenari, and A. D. Bimbo · 2014
Earlier work this paper cites.
Large-scale Video Classification with Convolutional Neural Networks
A. Karpathy, G. Toderici, S. Shetty, T. Leung, R. Sukthankar, and L. Fei-Fei · 2014
Earlier work this paper cites.
The LEAR Submission at THUMOS 2014, 2014
D. Oneata, J. Verbeek, and C. Schmid · 2014
Earlier work this paper cites.
Two-Stream Convolutional Networks for Action Recognition in Videos
K. Simonyan and A. Zisserman · 2014
Cited alongside, same era.
Action Recognition and Detection by Combining Motion and Appearance Features
L. Wang, Y. Qiao, and X. Tang · 2014
Cited alongside, same era.
Facial Landmark Detection by Deep Multi-task Learning
Z. Zhang, P. Luo, C. C. Loy, and X. Tang · 2014
Cited alongside, same era.
DeepStereo: Learning to Predict New Views from the World’s Imagery
J. Flynn, I. Neulander, J. Philbin, and N. Snavely · 2015
Cited alongside, same era.
Domain Generalization for Object Recognition with Multi-task Autoencoders
M. Ghifary, W. B. Kleijn, M. Zhang, and D. Balduzzi · 2015
Cited alongside, same era.
Finding Action Tubes
G. Gkioxari and J. Malik · 2015
Towards Good Practices for Very Deep Two-Stream ConvNets
L. Wang, Y. Xiong, Z. Wang, and Y. Qiao · 2015
Later among the works it cites.
Learning to Track for Spatio-Temporal Action Localization
P. Weinzaepfel, Z. Harchaoui, and C. Schmid · 2015
Later among the works it cites.
3D ShapeNets: A Deep Representation for Volumetric Shapes
Z. Wu, S. Song, A. Khosla, F. Yu, L. Zhang, X. Tang, and J. Xiao · 2015
Later among the works it cites.
Fast Action Proposals for Human Action Detection and Search
G. Yu and J. Yuan · 2015
Later among the works it cites.
ADSC Submission at THUMOS Challenge 2015
J. Yuan, Y. Pei, B. Ni, P. Moulin, and A. Kassim · 2015
Later among the works it cites.
Instance-aware Semantic Segmentation via Multi-task Network Cascades
J. Dai, K. He, and J. Sun · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Spatial Pyramid Pooling in Deep Convolutional Networks for Visual Recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2015
Cited alongside, same era.
ActivityNet: A Large-Scale Video Benchmark for Human Activity Understanding
F. C. Heilbron, V. Escorcia, B. Ghanem, and J. C. Niebles · 2015
Cited alongside, same era.
Objects2action: Classifying and Localizing Actions without Any Video Example
M. Jain, J. C. van Gemert, T. Mensink, and C. G. M. Snoek · 2015
Cited alongside, same era.
What do 15,000 Object Categories Tell Us about Classifying and Localizing Actions?
M. Jain, J. C. van Gemert, and C. G. M. Snoek · 2015
Cited alongside, same era.
Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks
S. Ren, K. He, R. Girshick, and J. Sun · 2015
Cited alongside, same era.
Action Localization in Videos through Context Walk
K. Soomro, H. Idrees, and M. Shah · 2015
Cited alongside, same era.
Closest in time.
Deep Residual Learning for Image Recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Closest in time.
Fast Temporal Activity Proposals for Efficient Detection of Human Actions in Untrimmed Videos
F. C. Heilbron, J. C. Niebles, and B. Ghanem · 2016
Closest in time.
DAP3D-Net: Where, What and How Actions Occur in Videos?
L. Liu, Y. Zhou, and L. Shao · 2016
Closest in time.
Multi-task Sequence to Sequence Learning
M.-T. Luong, Q. V. Le, I. Sutskever, O. Vinyals, and L. Kaiser · 2016
Closest in time.
Temporal Action Detection using a Statistical Language Model
A. Richard and J. Gall · 2016
Closest in time.
Temporal Action Localization in Untrimmed Videos via Multi-stage CNNs
Z. Shou, D. Wang, and S.-F. Chang · 2016
Closest in time.
Deep End2End Voxel2Voxel Prediction
D. Tran, L. Bourdev, R. Fergus, L. Torresani, and M. Paluri · 2016
Closest in time.
Improving Human Action Recognition by Non-action Classification
Y. Wang and M. Hoai · 2016
Closest in time.
End-to-end Learning of Action Detection from Frame Glimpses in Videos
S. Yeung, O. Russakovsky, G. Mori, and L. Fei-Fei · 2016
Closest in time.
Depth2Action: Exploring Embedded Depth for Large-Scale Action Recognition
Y. Zhu and S. Newsam · 2016
Closest in time.