Fetching the paper…
Reading the bibliography…
In this work we introduce a fully end-to-end approach for action detection in videos that learns to directly predict the temporal bounds of actions.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
R. J. Williams · 1992
Earlier work this paper cites.
Recognizing multitasked activities from video using stochastic context-free grammar
D. Moore and I. Essa · 2002
Earlier work this paper cites.
Detecting unusual activity in video
H. Zhong, J. Shi, and M. Visontai · 2004
Earlier work this paper cites.
Actions as space-time shapes
M. Blank, L. Gorelick, E. Shechtman, M. Irani, and R. Basri · 2005
Earlier work this paper cites.
Learning temporal sequence model from partially labeled data
Y. Shi, A. Bobick, and I. Essa · 2006
Earlier work this paper cites.
Event detection in crowded videos
Y. Ke, R. Sukthankar, and M. Hebert · 2007
Earlier work this paper cites.
Learning realistic human actions from movies
I. Laptev, M. Marszałek, C. Schmid, and B. Rozenfeld · 2008
Earlier work this paper cites.
Understanding videos, constructing plots learning a visually grounded storyline model from annotated videos
A. Gupta, P. Srinivasan, J. Shi, and L. S. Davis · 2009
Earlier work this paper cites.
A survey on vision-based human action recognition
R. Poppe · 2010
Earlier work this paper cites.
A survey of vision-based methods for action representation, segmentation and recognition
D. Weinland, R. Ronfard, and E. Boyer · 2010
Earlier work this paper cites.
A hough transform-based voting framework for action recognition
A. Yao, J. Gall, and L. Van Gool · 2010
Earlier work this paper cites.
Discriminative figure-centric models for joint action localization and recognition
T. Lan, Y. Wang, and G. Mori · 2011
Earlier work this paper cites.
Incremental activity modeling in multiple disjoint cameras
C. C. Loy, T. Xiang, and S. Gong · 2012
Earlier work this paper cites.
A database for fine grained activity detection of cooking activities
M. Rohrbach, S. Amin, M. Andriluka, and B. Schiele · 2012
Earlier work this paper cites.
Lecture 6.5-rmsprop: Divide the gradient by a running average of its recent magnitude., 2012
T. Tieleman and G. E. Hinton · 2012
Earlier work this paper cites.
Towards understanding action recognition
H. Jhuang, J. Gall, S. Zuffi, C. Schmid, and M. J. Black · 2013
Earlier work this paper cites.
Multi-agent event detection: Localization and role assignment
S. Kwak, B. Han, and J. H. Han · 2013
Cited alongside, same era.
Overfeat: Integrated recognition, localization and detection using convolutional networks
P. Sermanet, D. Eigen, X. Zhang, M. Mathieu, R. Fergus, and Y. LeCun · 2013
Cited alongside, same era.
Spatiotemporal deformable part models for action detection
Y. Tian, R. Sukthankar, and M. Shah · 2013
Cited alongside, same era.
Action recognition with improved trajectories
H. Wang and C. Schmid · 2013
Cited alongside, same era.
Multiple object recognition with visual attention
J. Ba, V. Mnih, and K. Kavukcuoglu · 2014
Cited alongside, same era.
Long-term recurrent convolutional networks for visual recognition and description
Attention for fine-grained categorization
P. Sermanet, A. Frome, and E. Real · 2014
Later among the works it cites.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2014
Later among the works it cites.
Scalable, high-quality object detection
C. Szegedy, S. Reed, D. Erhan, and D. Anguelov · 2014
Later among the works it cites.
Activitynet: A large-scale video benchmark for human activity understanding
F. Caba Heilbron, V. Escorcia, B. Ghanem, and J. Carlos Niebles · 2015
Closest in time.
R. Girshick · 2015
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Donahue, L. A. Hendricks, S. Guadarrama, M. Rohrbach, S. Venugopalan, K. Saenko, and T. Darrell · 2014
Cited alongside, same era.
Scalable object detection using deep neural networks
D. Erhan, C. Szegedy, A. Toshev, and D. Anguelov · 2014
Cited alongside, same era.
G. Gkioxari and J. Malik · 2014
Cited alongside, same era.
Action localization with tubelets from motion
M. Jain, J. Van Gemert, H. Jégou, P. Bouthemy, and C. G. Snoek · 2014
Cited alongside, same era.
THUMOS challenge: Action recognition with a large number of classes
Y.-G. Jiang, J. Liu, A. Roshan Zamir, G. Toderici, I. Laptev, M. Shah, and R. Sukthankar · 2014
Cited alongside, same era.
Efficient feature extraction, encoding, and classification for action recognition
V. Kantorov and I. Laptev · 2014
Cited alongside, same era.
Recurrent models of visual attention
V. Mnih, N. Heess, A. Graves, et al · 2014
Cited alongside, same era.
You only look once: Unified, real-time object detection
J. Redmon, S. Divvala, R. Girshick, and A. Farhadi · 2015
Closest in time.
Faster r-cnn: Towards real-time object detection with region proposal networks
S. Ren, K. He, R. Girshick, and J. Sun · 2015
Closest in time.
Joint inference of groups, events and human roles in aerial videos
T. Shu, D. Xie, B. Rothrock, S. Todorovic, and S.-C. Zhu · 2015
Closest in time.
Temporal localization of fine-grained actions in videos by domain transfer from web images
C. Sun, S. Shetty, R. Sukthankar, and R. Nevatia · 2015
Closest in time.
Learning to track for spatio-temporal action localization
P. Weinzaepfel, Z. Harchaoui, and C. Schmid · 2015
Closest in time.
Show, attend and tell: Neural image caption generation with visual attention
K. Xu, J. Ba, R. Kiros, A. Courville, R. Salakhutdinov, R. Zemel, and Y. Bengio · 2015
Closest in time.
Fast action proposals for human action detection and search
G. Yu and J. Yuan · 2015
Closest in time.
Adsc submission at thumos challenge 2015
J. Yuan, Y. Pei, B. Ni, P. Moulin, and A. Kassim · 2015
Closest in time.
Reinforcement learning neural turing machines
W. Zaremba and I. Sutskever · 2015
Closest in time.
Exploiting image-trained cnn architectures for unconstrained video classification
S. Zha, F. Luisier, W. Andrews, N. Srivastava, and R. Salakhutdinov · 2015
Closest in time.
Y. Zhu, R. Kiros, R. Zemel, R. Salakhutdinov, R. Urtasun, A. Torralba, and S. Fidler · 2015
Closest in time.