Fetching the paper…
Reading the bibliography…
Temporal Activity Detection aims to predict activity classes per frame, in contrast to video-level predictions in Activity Classification (i.e., Activity Recognition).
ImageNet: A Large-Scale Hierarchical Image Database
Deng, J.; Dong, W.; Socher, R.; Li, L.-J.; Li, K.; and Fei-Fei, L. 2009 · 2009
Earlier work this paper cites.
THUMOS Challenge: Action Recognition with a Large Number of Classes
Jiang, Y.-G.; Liu, J.; Zamir, A. R.; Toderici, G.; Laptev, I.; Shah, M.; and Sukthankar, R. 2014 · 2014
Earlier work this paper cites.
Two-Stream Convolutional Networks for Action Recognition in Videos
Simonyan, K.; and Zisserman, A. 2014 · 2014
Earlier work this paper cites.
Activitynet: A large-scale video benchmark for human activity understanding
Caba Heilbron, F.; Escorcia, V.; Ghanem, B.; and Carlos Niebles, J. 2015 · 2015
Earlier work this paper cites.
Temporal Localization of Fine-Grained Actions in Videos by Domain Transfer from Web Images
Sun, C.; Shetty, S.; Sukthankar, R.; and Nevatia, R. 2015 · 2015
Earlier work this paper cites.
Learning Spatiotemporal Features with 3D Convolutional Networks
Tran, D.; Bourdev, L.; Fergus, R.; Torresani, L.; and Paluri, M. 2015 · 2015
Earlier work this paper cites.
Beyond Short Snippets: Deep Networks for Video Classification
Yue-Hei Ng, J.; Hausknecht, M.; Vijayanarasimhan, S.; Vinyals, O.; Monga, R.; and Toderici, G. 2015 · 2015
Earlier work this paper cites.
DAPs: Deep Action Proposals for Action Understanding
Escorcia, V.; Heilbron, F. C.; Niebles, J. C.; and Ghanem, B. 2016 · 2016
Earlier work this paper cites.
Convolutional Two-Stream Network Fusion for Video Action Recognition
Feichtenhofer, C.; Pinz, A.; and Zisserman, A. 2016 · 2016
Earlier work this paper cites.
Deep Residual Learning for Image Recognition
He, K.; Zhang, X.; Ren, S.; and Sun, J. 2016 · 2016
Earlier work this paper cites.
Shuffle and Learn: Unsupervised Learning using Temporal Order Verification
Misra, I.; Zitnick, C. L.; and Hebert, M. 2016 · 2016
Earlier work this paper cites.
Temporal Action Localization in Untrimmed Videos via Multi-stage CNNs
Shou, Z.; Wang, D.; and Chang, S.-F. 2016 · 2016
Earlier work this paper cites.
Hollywood in Homes: Crowdsourcing Data Collection for Activity Understanding
Sigurdsson, G. A.; Varol, G.; Wang, X.; Farhadi, A.; Laptev, I.; and Gupta, A. 2016 · 2016
Earlier work this paper cites.
End-to-End Learning of Action Detection from Frame Glimpses in Videos
Yeung, S.; Russakovsky, O.; Mori, G.; and Fei-Fei, L. 2016 · 2016
Earlier work this paper cites.
SST: Single-Stream Temporal Action Proposals
Buch, S.; Escorcia, V.; Shen, C.; Ghanem, B.; and Carlos Niebles, J. 2017 · 2017
Earlier work this paper cites.
Quo Vadis, Action Recognition? A New Model and the Kinetics Dataset
Carreira, J.; and Zisserman, A. 2017 · 2017
Earlier work this paper cites.
Weakly Supervised Action Learning with RNN based Fine-to-coarse Modeling
Richard, A.; Kuehne, H.; and Gall, J. 2017 · 2017
Earlier work this paper cites.
CDC: Convolutional-De-Convolutional Networks for Precise Temporal Action Localization in Untrimmed Videos
Shou, Z.; Chan, J.; Zareian, A.; Miyazawa, K.; and Chang, S.-F. 2017 · 2017
Earlier work this paper cites.
Long-term Temporal Convolutions for Action Recognition
Varol, G.; Laptev, I.; and Schmid, C. 2017 · 2017
Earlier work this paper cites.
R-C3D: Region Convolutional 3D Network for Temporal Activity Detection
Xu, H.; Das, A.; and Saenko, K. 2017 · 2017
Earlier work this paper cites.
Object Level Visual Reasoning in Videos
Baradel, F.; Neverova, N.; Wolf, C.; Mille, J.; and Mori, G. 2018 · 2018
Earlier work this paper cites.
Exploring the Limits of Weakly Supervised Pretraining
Mahajan, D.; Girshick, R.; Ramanathan, V.; He, K.; Paluri, M.; Li, Y.; Bharambe, A.; and Van Der Maaten, L. 2018 · 2018
Earlier work this paper cites.
Weakly Supervised Action Localization by Sparse Temporal Pooling Network
Nguyen, P.; Liu, T.; Prasad, G.; and Han, B. 2018 · 2018
Earlier work this paper cites.
Learning Latent Super-Events to Detect Multiple Activities in Videos
Piergiovanni, A.; and Ryoo, M. S. 2018 · 2018
Cited alongside, same era.
Unsupervised Learning and Segmentation of Complex Activities From Video
Sener, F.; and Yao, A. 2018 · 2018
Cited alongside, same era.
Learning and Using the Arrow of Time
Wei, D.; Lim, J. J.; Zisserman, A.; and Freeman, W. T. 2018 · 2018
Cited alongside, same era.
Every Moment Counts: Dense Detailed Labeling of Actions in Complex Videos
Yeung, S.; Russakovsky, O.; Jin, N.; Andriluka, M.; Mori, G.; and Fei-Fei, L. 2018 · 2018
Cited alongside, same era.
mixup: Beyond Empirical Risk Minimization
Zhang, H.; Cisse, M.; Dauphin, Y. N.; and Lopez-Paz, D. 2018 · 2018
Cited alongside, same era.
SlowFast Networks for Video Recognition
Feichtenhofer, C.; Fan, H.; Malik, J.; and He, K. 2019 · 2019
Cited alongside, same era.
Representation Learning on Visual-Symbolic Graphs for Video Understanding
Mavroudi, E.; Haro, B. B.; and Vidal, R. 2020 · 2020
Later among the works it cites.
Aligning Videos in Space and Time
Purushwalkam, S.; Ye, T.; Gupta, S.; and Gupta, A. 2020 · 2020
Later among the works it cites.
AssembleNet: Searching for Multi-Stream Neural Connectivity in Video Architectures
Ryoo, M.; Piergiovanni, A.; Tan, M.; and Angelova, A. 2020 · 2020
Later among the works it cites.
Weakly-Supervised Action Localization by Generative Attention Modeling
Shi, B.; Dai, Q.; Mu, Y.; and Wang, J. 2020 · 2020
Later among the works it cites.
Learning Actionness via Long-range Temporal Order Verification
Zhukov, D.; Alayrac, J.-B.; Laptev, I.; and Sivic, J. 2020 · 2020
Later among the works it cites.
Exploring the Limits of Large Scale Pre-training
Abnar, S.; Dehghani, M.; Neyshabur, B.; and Sedghi, H. 2021 · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Large-Scale Weakly-Supervised Pre-Training for Video Action Recognition
Ghadiyaram, D.; Tran, D.; and Mahajan, D. 2019 · 2019
Cited alongside, same era.
Learning Temporal Action Proposals With Fewer Labels
Ji, J.; Cao, K.; and Niebles, J. C. 2019 · 2019
Cited alongside, same era.
Unsupervised Learning of Action Classes With Continuous Temporal Embedding
Kukleva, A.; Kuehne, H.; Sener, F.; and Gall, J. 2019 · 2019
Cited alongside, same era.
Completeness Modeling and Context Separation for Weakly Supervised Temporal Action Localization
Liu, D.; Jiang, T.; and Wang, Y. 2019 · 2019
Cited alongside, same era.
Weakly Supervised Temporal Action Localization Through Contrast Based Evaluation Networks
Liu, Z.; Wang, L.; Zhang, Q.; Gao, Z.; Niu, Z.; Zheng, N.; and Hua, G. 2019 · 2019
Cited alongside, same era.
Temporal Gaussian Mixture Layer for Videos
Piergiovanni, A.; and Ryoo, M. S. 2019 · 2019
Cited alongside, same era.
Closest in time.
Tsp: Temporally-sensitive pretraining of video encoders for localization tasks
Alwassel, H.; Giancola, S.; and Ghanem, B. 2021 · 2021
Closest in time.
On the Opportunities and Risks of Foundation Models
Bommasani, R.; Hudson, D. A.; Adeli, E.; Altman, R.; Arora, S.; von Arx, S.; Bernstein, M. S.; Bohg, J.; Bosselut, A.; Brunskill, E.; et al. 2021 · 2021
Closest in time.
Augmented Transformer with Adaptive Graph for Temporal Action Proposal Generation
Chang, S.; Wang, P.; Wang, F.; Li, H.; and Feng, J. 2021 · 2021
Closest in time.
Exploring Simple Siamese Representation Learning
Chen, X.; and He, K. 2021 · 2021
Closest in time.
Multiscale vision transformers
Fan, H.; Xiong, B.; Mangalam, K.; Li, Y.; Yan, Z.; Malik, J.; and Feichtenhofer, C. 2021 · 2021
Closest in time.
Coarse-Fine Networks for Temporal Activity Detection in Videos
Kahatapitiya, K.; and Ryoo, M. S. 2021 · 2021
Closest in time.
Acsnet: Action-context separation network for weakly supervised temporal action localization
Liu, Z.; Wang, L.; Zhang, Q.; Tang, W.; Yuan, J.; Zheng, N.; and Hua, G. 2021 · 2021
Closest in time.
Activity Graph Transformer for Temporal Action Localization
Nawhal, M.; and Mori, G. 2021 · 2021
Closest in time.
Broaden Your Views for Self-Supervised Video Learning
Recasens, A.; Luc, P.; Alayrac, J.-B.; Wang, L.; Strub, F.; Tallec, C.; Malinowski, M.; Pătrăucean, V.; Altché, F.; Valko, M.; Grill, J.-B.; van den Oord, A.; and Zisserman, A. 2021 · 2021
Closest in time.
Pretraining Representations for Data-Efficient Reinforcement Learning
Schwarzer, M.; Rajkumar, N.; Noukhovitch, M.; Anand, A.; Charlin, L.; Hjelm, D.; Bachman, P.; and Courville, A. 2021 · 2021
Closest in time.
Modeling multi-label action dependencies for temporal action localization
Tirupattur, P.; Duarte, K.; Rawat, Y. S.; and Shah, M. 2021 · 2021
Closest in time.
Action coherence network for weakly-supervised temporal action localization
Zhai, Y.; Wang, L.; Tang, W.; Zhang, Q.; Zheng, N.; and Hua, G. 2021 · 2021
Closest in time.
Video Self-Stitching Graph Network for Temporal Action Localization
Zhao, C.; Thabet, A. K.; and Ghanem, B. 2021 · 2021
Closest in time.
Semi-Supervised and Unsupervised Deep Visual Learning: A Survey
Chen, Y.; Mancini, M.; Zhu, X.; and Akata, Z. 2022 · 2022
Closest in time.
MS-TCT: Multi-Scale Temporal ConvTransformer for Action Detection
Dai, R.; Das, S.; Kahatapitiya, K.; Ryoo, M. S.; and Bremond, F. 2022 · 2022
Closest in time.
Uncertainty-Based Spatial-Temporal Attention for Online Action Detection
Guo, H.; Ren, Z.; Wu, Y.; Hua, G.; and Ji, Q. 2022 · 2022
Closest in time.
Unsupervised pre-training for temporal action localization tasks
Zhang, C.; Yang, T.; Weng, J.; Cao, M.; Wang, J.; and Zou, Y. 2022 · 2022
Closest in time.