Fetching the paper…
Reading the bibliography…
With the knowledge of action moments (i.e., trimmed video clips that each contains an action instance), humans could routinely localize an action temporally in an untrimmed video.
van der Maaten, L., Hinton, G.: Visualizing Data using t-SNE. JMLR (2008)
2008
Earlier work this paper cites.
Gaidon, A., Harchaoui, Z., Schmid, C.: Temporal Localization of Actions with Actoms. IEEE Trans. on PAMI 35
2013
Earlier work this paper cites.
Maas, A.L., Hannun, A.Y., Ng, A.Y.: Rectifier Nonlinearities Improve Neural Network Acoustic Models. In: ICML (2013)
2013
Earlier work this paper cites.
Oneata, D., Verbeek, J., Schmid, C.: Action and Event Recognition with Fisher Vectors on a Compact Feature Set. In: ICCV (2013)
2013
Earlier work this paper cites.
Tang, K., Yao, B., Fei-Fei, L., Koller, D.: Combining the Right Features for Complex Event Recognition. In: ICCV (2013)
2013
Earlier work this paper cites.
Goodfellow, I.J., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A., Bengio, Y.: Generative Adversarial Nets. In: NIPS (2014)
2014
Earlier work this paper cites.
Hoffman, J., Guadarrama, S., Tzeng, E., Hu, R., Donahue, J.: LSDA: Large Scale Detection through Adaptation. In: NIPS (2014)
2014
Earlier work this paper cites.
Jiang, Y.G., Liu, J., R.Zamir, A., Toderici, G.: THUMOS challenge: Action recognition with a large number of classes. http://crcv.ucf.edu/THUMOS14 (2014)
2014
Earlier work this paper cites.
Lu, S., Wang, Z., Mei, T., Guan, G., Feng, D.D.: A Bag-of-Importance Model With Locality-Constrained Coding Based Feature Learning for Video Summarization. IEEE Trans. on Multimedia 16
2014
Earlier work this paper cites.
Ganin, Y., Lempitsky, V.: Unsupervised Domain Adaptation by Backpropagation. In: ICML (2015)
2015
Earlier work this paper cites.
Girshick, R.: Fast R-CNN. In: ICCV (2015)
2015
Earlier work this paper cites.
Heilbron, F.C., Escorcia, V., Ghanem, B., Niebles, J.C.: ActivityNet: A Large-Scale Video Benchmark for Human Activity Understanding. In: CVPR (2015)
2015
Earlier work this paper cites.
Kingma, D.P., Ba, J.: Adam: A Method for Stochastic Optimization. In: ICLR (2015)
2015
Earlier work this paper cites.
Abadi, M., Barham, P., Chen, J., Chen, Z., Davis, A., Dean, J., Devin, M., Ghemawat, S., Irving, G., Isard, M., Kudlur, M., Levenberg, J., Monga, R., Moore, S., Murray, D.G., Steiner, B., Tucker, P., Vasudevan, V., Warden, P., Wicke, M., Yu, Y., Zheng, X.: TensorFlow: A System for Large-scale Machine Learning. In: OSDI (2016)
2016
Earlier work this paper cites.
Clevert, D.A., Unterthiner, T., Hochreiter, S.: Fast and Accurate Deep Network Learning by Exponential Linear Units (ELUs). In: ICLR (2016)
2016
Earlier work this paper cites.
Escorcia, V., Heilbron, F.C., Niebles, J.C., Ghanem, B.: DAPs: Deep Action Proposals for Action Understanding. In: ECCV (2016)
2016
Earlier work this paper cites.
Geest, R.D., Gavves, E., Ghodrati, A., Li, Z., Snoek, C., Tuytelaars, T.: Online Action Detection. In: ECCV (2016)
2016
Earlier work this paper cites.
Shou, Z., Wang, D., Chang, S.F.: Temporal Action Localization in Untrimmed Videos via Multi-stage CNNs. In: CVPR (2016)
2016
Earlier work this paper cites.
Singh, B., Marks, T.K., Jones, M., Tuzel, O., Shao, M.: A Multi-Stream Bi-Directional Recurrent Neural Network for Fine-Grained Action Detection. In: CVPR (2016)
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
Tang, Y., Wang, J., Gao, B., Dellandrea, E., Gaizauskas, R., Chen, L.: Large Scale Semi-supervised Object Detection using Visual and Semantic Knowledge Transfer. In: CVPR (2016)
2016
Cited alongside, same era.
Wang, R., Tao, D.: UTS at activitynet 2016. In: CVPR ActivityNet Challenge Workshop (2016)
2016
Cited alongside, same era.
Yeung, S., Russakovsky, O., Mori, G., Fei-Fei, L.: End-to-end Learning of Action Detection from Frame Glimpses in Videos. In: CVPR (2016)
2016
Cited alongside, same era.
Yuan, J., Ni, B., Yang, X., A.Kassim, A.: Temporal Action Localization With Pyramid of Score Distribution Features. In: CVPR (2016)
2016
Cited alongside, same era.
Buch, S., Escorcia, V., Ghanem, B., Fei-Fei, L., Niebles, J.C.: End-to-End, Single-Stream Temporal Action Detection in Untrimmed Videos. In: BMVC (2017)
2017
Cited alongside, same era.
Zhao, Y., Xiong, Y., Wang, L., Wu, Z., Tang, X., Lin, D.: Temporal Action Detection with Structured Segment Networks. In: ICCV (2017)
2017
Later among the works it cites.
Chao, Y.W., Vijayanarasimhan, S., Seybold, B., Ross, D.A., Deng, J., Sukthankar, R.: Rethinking the Faster R-CNN Architecture for Temporal Action Localization. In: CVPR (2018)
2018
Later among the works it cites.
Gao, J., Chen, K., Nevatia, R.: CTAP: Complementary Temporal Action Proposal Generation. In: ECCV (2018)
2018
Later among the works it cites.
2018
Later among the works it cites.
Hu, R., Dollar, P., He, K., Darell, T., Girshick, R.: Learning to Segment Every Thing. In: CVPR (2018)
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Buch, S., Escorcia, V., Shen, C., Ghanem, B., Niebles, J.C.: SST: Single-Stream Temporal Action Proposals. In: CVPR (2017)
2017
Cited alongside, same era.
Gao, J., Yang, Z., Sun, C., Chen, K., Nevatia, R.: TURN TAP: Temporal Unit Regression Network for Temporal Action Proposals. In: ICCV (2017)
2017
Cited alongside, same era.
He, K., Gkioxari, G., Dollar, P., Girshick, R.: Mask R-CNN. In: ICCV (2017)
2017
Cited alongside, same era.
Heilbron, F.C., Barrios, W., Escorica, V., Ghanem, B.: SCC: Semantic Context Cascade for Efficient Action Detection. In: CVPR (2017)
2017
Cited alongside, same era.
Isola, P., Zhu, J.Y., Zhou, T., Efros, A.A.: Image-to-Image Translation with Conditional Adversarial Networks. In: CVPR (2017)
2017
Cited alongside, same era.
Lea, C., Michael D. Flynn, R.V., Reiter, A., Hager, G.D.: Temporal Convolutional Netowrk for Action Segmentation and Detection. In: CVPR (2017)
2017
Cited alongside, same era.
Lin, T., Zhao, X., Shou, Z.: Single Shot Temporal Action Detection. In: ACM MM (2017)
2017
Cited alongside, same era.
2018
Later among the works it cites.
Li, D., Qiu, Z., Dai, Q., Yao, T., Mei, T.: Recurrent Tubelet Proposal and Recognition Networks for Action Detection. In: ECCV (2018)
2018
Later among the works it cites.
Lin, T., Zhao, X., Su, H., Wang, C., Yang, M.: BSN: Boundary Sensitive Network for Temporal Action Proposal Generation. In: ECCV (2018)
2018
Later among the works it cites.
Nguyen, P., Liu, T., Prasad, G., Han, B.: Weakly Supervised Action Localization by Sparse Temporal Pooling Network. In: CVPR (2018)
2018
Later among the works it cites.
Shou, Z., Gao, H., Zhang, L., Miyazawa, K., Chang, S.F.: AutoLoc: Weakly-supervised Temporal Action Localization in Untrimmed Videos. In: ECCV (2018)
2018
Later among the works it cites.
Shou, Z., Pan, J., Chan, J., Miyazawa, K., Mansour, H., Vetro, A., i Nieto, X.G., Chang, S.F.: Online Detection of Action Start in Untrimmed, Streaming Videos. In: ECCV (2018)
2018
Later among the works it cites.
Kuen, J., Perazzi, F., Lin, Z., Zhang, J., Tan, Y.P.: Scaling Object Detection by Transferring Classification Weights. In: ICCV (2019)
2019
Later among the works it cites.
Li, D., Yao, T., Qiu, Z., Li, H., Mei, T.: Long Short-Term Relation Networks for Video Action Detection. In: ACM MM (2019)
2019
Later among the works it cites.
Liu, D., Jiang, T., Wang, Y.: Completeness Modeling and Context Separation for Weakly Supervised Temporal Action Localization. In: CVPR (2019)
2019
Later among the works it cites.
Long, F., Yao, T., Qiu, Z., Tian, X., Luo, J., Mei, T.: Gaussian Temporal Awareness Networks for Action Localization. In: CVPR (2019)
2019
Later among the works it cites.
Nguyen, P.X., Ramanan, D., Fowlkes, C.C.: Weakly-supervised Action Localization with Background Modeling. In: ICCV (2019)
2019
Later among the works it cites.
Qiu, Z., Yao, T., Ngo, C.W., Tian, X., Mei, T.: Learning Spatio-Temporal Representation with Local and Global Diffusion. In: CVPR (2019)
2019
Later among the works it cites.
Zeng, R., Huang, W., Tan, M., Rong, Y., Zhao, P., Huang, J., Gan, C.: Graph Convolutional Networks for Temporal Action Localization. In: ICCV (2019)
2019
Later among the works it cites.
Long, F., Yao, T., Qiu, Z., Tian, X., Mei, T., Luo, J.: Coarse-to-Fine Localization of Temporal Action Proposals. IEEE Trans. on Multimedia 22
2020
Closest in time.
Shi, B., Dai, Q., Mu, Y., Wang, J.: Weakly-Supervised Action Localization by Generative Attention Modeling. In: CVPR (2020)
2020
Closest in time.