Fetching the paper…
Reading the bibliography…
In this paper, we present a novel Single Shot multi-Span Detector for temporal activity detection in long, untrimmed videos using a simple end-to-end fully three-dimensional convolutional (Conv3D) network.
Aggregating local descriptors into a compact image representation
Hervé Jégou, Matthijs Douze, Cordelia Schmid, and Patrick Pérez · 2010
Earlier work this paper cites.
Improving the fisher kernel for large-scale image classification
Florent Perronnin, Jorge Sánchez, and Thomas Mensink · 2010
Earlier work this paper cites.
Action recognition by dense trajectories
Heng Wang, Alexander Kläser, Cordelia Schmid, and Cheng-Lin Liu · 2011
Earlier work this paper cites.
Temporal localization of actions with actoms
Adrien Gaidon, Zaid Harchaoui, and Cordelia Schmid · 2013
Earlier work this paper cites.
Action and event recognition with fisher vectors on a compact feature set
Dan Oneata, Jakob Verbeek, and Cordelia Schmid · 2013
Earlier work this paper cites.
Combining the right features for complex event recognition
Kevin Tang, Bangpeng Yao, Li Fei-Fei, and Daphne Koller · 2013
Earlier work this paper cites.
Action recognition with improved trajectories
Heng Wang and Cordelia Schmid · 2013
Earlier work this paper cites.
Rich feature hierarchies for accurate object detection and semantic segmentation
Ross Girshick, Jeff Donahue, Trevor Darrell, and Jitendra Malik · 2014
Earlier work this paper cites.
Action localization with tubelets from motion
Mihir Jain, Jan Van Gemert, Hervé Jégou, Patrick Bouthemy, and Cees GM Snoek · 2014
Earlier work this paper cites.
Thumos challenge: Action recognition with a large number of classes, 2014
YG Jiang, J Liu, A Roshan Zamir, G Toderici, I Laptev, M Shah, and R Sukthankar · 2014
Earlier work this paper cites.
The lear submission at thumos 2014
Dan Oneata, Jakob Verbeek, and Cordelia Schmid · 2014
Earlier work this paper cites.
Action recognition and detection by combining motion and appearance features
Limin Wang, Yu Qiao, and Xiaoou Tang · 2014
Earlier work this paper cites.
Fast r-cnn
Ross Girshick · 2015
Earlier work this paper cites.
Bag-of-fragments: Selecting and encoding video fragments for event detection and recounting
Pascal Mettes, Jan C van Gemert, Spencer Cappallo, Thomas Mensink, and Cees GM Snoek · 2015
Cited alongside, same era.
Faster r-cnn: Towards real-time object detection with region proposal networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun · 2015
Cited alongside, same era.
Learning spatiotemporal features with 3d convolutional networks
Du Tran, Lubomir Bourdev, Rob Fergus, Lorenzo Torresani, and Manohar Paluri · 2015
Cited alongside, same era.
Apt: Action localization proposals from dense trajectories
Jan C van Gemert, Mihir Jain, Ella Gati, Cees GM Snoek, et al · 2015
Cited alongside, same era.
Towards good practices for very deep two-stream convnets
Limin Wang, Yuanjun Xiong, Zhe Wang, and Yu Qiao · 2015
Cited alongside, same era.
Actionness estimation using hybrid fully convolutional networks
Limin Wang, Yu Qiao, Xiaoou Tang, and Luc Van Gool · 2016
Later among the works it cites.
End-to-end learning of action detection from frame glimpses in videos
Serena Yeung, Olga Russakovsky, Greg Mori, and Li Fei-Fei · 2016
Later among the works it cites.
Sst: Single-stream temporal action proposals
Shyamal Buch, Victor Escorcia, Chuanqi Shen, Bernard Ghanem, and Juan Carlos Niebles · 2017
Later among the works it cites.
Quo vadis, action recognition? a new model and the kinetics dataset
Joao Carreira and Andrew Zisserman · 2017
Later among the works it cites.
Stylebank: An explicit representation for neural image style transfer
Dongdong Chen, Lu Yuan, Jing Liao, Nenghai Yu, and Gang Hua · 2017
Later among the works it cites.
Fason: First and second order information fusion network for texture recognition
Xiyang Dai, Joe Yue-Hei Ng, and Larry S Davis · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Fast action proposals for human action detection and search
Gang Yu and Junsong Yuan · 2015
Cited alongside, same era.
Fast temporal activity proposals for efficient detection of human actions in untrimmed videos
Fabian Caba Heilbron, Juan Carlos Niebles, and Bernard Ghanem · 2016
Cited alongside, same era.
Daps: Deep action proposals for action understanding
Victor Escorcia, Fabian Caba Heilbron, Juan Carlos Niebles, and Bernard Ghanem · 2016
Cited alongside, same era.
Convolutional two-stream network fusion for video action recognition
Christoph Feichtenhofer, Axel Pinz, and Andrew Zisserman · 2016
Cited alongside, same era.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Cited alongside, same era.
Ssd: Single shot multibox detector
Wei Liu, Dragomir Anguelov, Dumitru Erhan, Christian Szegedy, Scott Reed, Cheng-Yang Fu, and Alexander C Berg · 2016
Cited alongside, same era.
Spot on: Action localization from pointly-supervised proposals
Pascal Mettes, Jan C van Gemert, and Cees GM Snoek · 2016
Cited alongside, same era.
Later among the works it cites.
Single shot temporal action detection
Tianwei Lin, Xu Zhao, and Zheng Shou · 2017
Later among the works it cites.
Cdc: Convolutional-de-convolutional networks for precise temporal action localization in untrimmed videos
Zheng Shou, Jonathan Chan, Alireza Zareian, Kazuyuki Miyazawa, and Shih-Fu Chang · 2017
Later among the works it cites.
R-c3d: Region convolutional 3d network for temporal activity detection
Huijuan Xu, Abir Das, and Kate Saenko · 2017
Later among the works it cites.
Deep reinforcement learning for visual object tracking in videos
Da Zhang, Hamid Maei, Xin Wang, and Yuan-Fang Wang · 2017
Later among the works it cites.
Temporal action detection with structured segment networks
Yue Zhao, Yuanjun Xiong, Limin Wang, Zhirong Wu, Xiaoou Tang, and Dahua Lin · 2017
Later among the works it cites.
Deep exemplar-based colorization
Mingming He, Dongdong Chen, Jing Liao, Pedro V Sander, and Lu Yuan · 2018
Closest in time.