Fetching the paper…
Reading the bibliography…
Temporal action proposal generation is an important task, aiming to localize the video segments containing human actions in an untrimmed video.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Rectified linear units improve restricted boltzmann machines
V. Nair and G. E. Hinton · 2010
Earlier work this paper cites.
Trajectory-based modeling of human actions with motion reference points
Y. Jiang, Q. Dai, X. Xue, W. Liu, and C. Ngo · 2012
Earlier work this paper cites.
Mining spatiotemporal video patterns towards robust action retrieval
L. Cao, R. Ji, Y. Gao, W. Liu, and Q. Tian · 2013
Earlier work this paper cites.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
K. Cho, B. Van Merriënboer, C. Gulcehre, D. Bahdanau, F. Bougares, H. Schwenk, and Y. Bengio · 2014
Earlier work this paper cites.
Rich feature hierarchies for accurate object detection and semantic segmentation
R. Girshick, J. Donahue, T. Darrell, and J. Malik · 2014
Earlier work this paper cites.
THUMOS challenge: Action recognition with a large number of classes
Y.-G. Jiang, J. Liu, A. Roshan Zamir, G. Toderici, I. Laptev, M. Shah, and R. Sukthankar · 2014
Earlier work this paper cites.
Two-stream convolutional networks for action recognition in videos
K. Simonyan and A. Zisserman · 2014
Earlier work this paper cites.
Action recognition and detection by combining motion and appearance features
L. Wang, Y. Qiao, and X. Tang · 2014
Earlier work this paper cites.
Activitynet: A large-scale video benchmark for human activity understanding
F. Caba Heilbron, V. Escorcia, B. Ghanem, and J. Carlos Niebles · 2015
Earlier work this paper cites.
Fast r-cnn
R. Girshick · 2015
Earlier work this paper cites.
Human action recognition in unconstrained videos by explicit motion modeling
Y. Jiang, Q. Dai, W. Liu, X. Xue, and C. Ngo · 2015
Earlier work this paper cites.
Bilinear cnn models for fine-grained visual recognition
T.-Y. Lin, A. RoyChowdhury, and S. Maji · 2015
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
S. Ren, K. He, R. Girshick, and J. Sun · 2015
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
O. Ronneberger, P. Fischer, and T. Brox · 2015
Earlier work this paper cites.
Learning spatiotemporal features with 3d convolutional networks
D. Tran, L. Bourdev, R. Fergus, L. Torresani, and M. Paluri · 2015
Earlier work this paper cites.
Describing videos by exploiting temporal structure
L. Yao, A. Torabi, K. Cho, N. Ballas, C. Pal, H. Larochelle, and A. Courville · 2015
Earlier work this paper cites.
Daps: Deep action proposals for action understanding
V. Escorcia, F. C. Heilbron, J. C. Niebles, and B. Ghanem · 2016
Cited alongside, same era.
Temporal action detection using a statistical language model
A. Richard and J. Gall · 2016
Cited alongside, same era.
Temporal action localization in untrimmed videos via multi-stage cnns
Z. Shou, D. Wang, and S.-F. Chang · 2016
Cited alongside, same era.
Highlight detection with pairwise deep ranking for first-person video summarization
T. Yao, T. Mei, and Y. Rui · 2016
Cited alongside, same era.
End-to-end learning of action detection from frame glimpses in videos
S. Yeung, O. Russakovsky, G. Mori, and L. Fei-Fei · 2016
Cited alongside, same era.
Temporal action localization with pyramid of score distribution features
J. Yuan, B. Ni, X. Yang, and A. A. Kassim · 2016
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin · 2017
Later among the works it cites.
Untrimmednets for weakly supervised action recognition and detection
L. Wang, Y. Xiong, D. Lin, and L. V. Gool · 2017
Later among the works it cites.
A pursuit of temporal accuracy in general activity detection
Y. Xiong, Y. Zhao, L. Wang, D. Lin, and X. Tang · 2017
Later among the works it cites.
R-c3d: region convolutional 3d network for temporal activity detection
H. Xu, A. Das, and K. Saenko · 2017
Later among the works it cites.
Msr asia msm at activitynet challenge 2017: Trimmed action recognition, temporal action proposals and densecaptioning events in videos
T. Yao, Y. Li, Z. Qiu, F. Long, Y. Pan, D. Li, and T. Mei · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Sst: Single-stream temporal action proposals
S. Buch, V. Escorcia, C. Shen, B. Ghanem, and J. C. Niebles · 2017
Cited alongside, same era.
Quo vadis, action recognition? a new model and the kinetics dataset
J. Carreira and A. Zisserman · 2017
Cited alongside, same era.
Temporal context network for activity localization in videos
X. Dai, B. Singh, G. Zhang, L. S. Davis, and Y. Q. Chen · 2017
Cited alongside, same era.
Dssd: Deconvolutional single shot detector
C.-Y. Fu, W. Liu, A. Ranga, A. Tyagi, and A. C. Berg · 2017
Cited alongside, same era.
Turn tap: Temporal unit regression network for temporal action proposals
J. Gao, Z. Yang, C. Sun, K. Chen, and R. Nevatia · 2017
Cited alongside, same era.
Convolutional sequence to sequence learning
J. Gehring, M. Auli, D. Grangier, D. Yarats, and Y. N. Dauphin · 2017
Cited alongside, same era.
Temporal action localization by structured maximal sums
Z.-H. Yuan, J. C. Stroud, T. Lu, and J. Deng · 2017
Later among the works it cites.
Temporal action detection with structured segment networks
Y. Zhao, Y. Xiong, L. Wang, Z. Wu, X. Tang, and D. Lin · 2017
Later among the works it cites.
Rethinking the faster r-cnn architecture for temporal action localization
Y.-W. Chao, S. Vijayanarasimhan, B. Seybold, D. A. Ross, J. Deng, and R. Sukthankar · 2018
Closest in time.
Temporally grounding natural sentence in video
J. Chen, X. Chen, L. Ma, Z. Jie, and T.-S. Chua · 2018
Closest in time.
Video re-localization
Y. Feng, L. Ma, W. Liu, T. Zhang, and J. Luo · 2018
Closest in time.
Ctap: Complementary temporal action proposal generation
J. Gao, K. Chen, and R. Nevatia · 2018
Closest in time.
Bsn: Boundary sensitive network for temporal action proposal generation
T. Lin, X. Zhao, H. Su, C. Wang, and M. Yang · 2018
Closest in time.
Autoloc: Weaklysupervised temporal action localization in untrimmed videos
Z. Shou, H. Gao, L. Zhang, K. Miyazawa, and S.-F. Chang · 2018
Closest in time.
Reconstruction network for video captioning
B. Wang, L. Ma, W. Zhang, and W. Liu · 2018
Closest in time.
Bidirectional attentive fusion with context gating for dense video captioning
J. Wang, W. Jiang, L. Ma, W. Liu, and Y. Xu · 2018
Closest in time.
Localizing natural language in videos
J. Chen, L. Ma, X. Chen, Z. Jie, and J. Luo · 2019
Closest in time.
Towards efficient action recognition: Principal backpropagation for training two-stream networks
W. Huang, L. Fan, M. Harandi, L. Ma, H. Liu, W. Liu, and C. Gan · 2019
Closest in time.