Fetching the paper…
Reading the bibliography…
Existing Temporal Action Detection (TAD) methods typically take a pre-processing step in converting an input varying-length video into a fixed-length snippet representation sequence, before temporal boundary estimation and action classification.
Activitynet: A large-scale video benchmark for human activity understanding
Fabian Caba Heilbron, Victor Escorcia, Bernard Ghanem, and Juan Carlos Niebles · 2015
Earlier work this paper cites.
Faster r-cnn: towards real-time object detection with region proposal networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun · 2016
Earlier work this paper cites.
Temporal segment networks: Towards good practices for deep action recognition
Limin Wang, Yuanjun Xiong, Zhe Wang, Yu Qiao, Dahua Lin, Xiaoou Tang, and Luc Van Gool · 2016
Earlier work this paper cites.
Sst: Single-stream temporal action proposals
Shyamal Buch, Victor Escorcia, Chuanqi Shen, Bernard Ghanem, and Juan Carlos Niebles · 2017
Earlier work this paper cites.
Turn tap: Temporal unit regression network for temporal action proposals
Jiyang Gao, Zhenheng Yang, Kan Chen, Chen Sun, and Ram Nevatia · 2017
Earlier work this paper cites.
The thumos challenge on action recognition for videos “in the wild”
Haroon Idrees, Amir R Zamir, Yu-Gang Jiang, Alex Gorban, Ivan Laptev, Rahul Sukthankar, and Mubarak Shah · 2017
Earlier work this paper cites.
R-c3d: Region convolutional 3d network for temporal activity detection
Huijuan Xu, Abir Das, and Kate Saenko · 2017
Earlier work this paper cites.
Temporal action detection with structured segment networks
Yue Zhao, Yuanjun Xiong, Limin Wang, Zhirong Wu, Xiaoou Tang, and Dahua Lin · 2017
Earlier work this paper cites.
Diagnosing error in temporal action detectors
Humam Alwassel, Fabian Caba Heilbron, Victor Escorcia, and Bernard Ghanem · 2018
Earlier work this paper cites.
Rethinking the faster r-cnn architecture for temporal action localization
Yu-Wei Chao, Sudheendra Vijayanarasimhan, Bryan Seybold, David A Ross, Jia Deng, and Rahul Sukthankar · 2018
Earlier work this paper cites.
BSN: Boundary sensitive network for temporal action proposal generation
Tianwei Lin, Xu Zhao, Haisheng Su, Chongjing Wang, and Ming Yang · 2018
Earlier work this paper cites.
Bmn: Boundary-matching network for temporal action proposal generation
Tianwei Lin, Xiao Liu, Xin Li, Errui Ding, and Shilei Wen · 2019
Earlier work this paper cites.
Gaussian temporal awareness networks for action localization
Fuchen Long, Ting Yao, Zhaofan Qiu, Xinmei Tian, Jiebo Luo, and Tao Mei · 2019
Earlier work this paper cites.
Graph convolutional networks for temporal action localization
Runhao Zeng, Wenbing Huang, Mingkui Tan, Yu Rong, Peilin Zhao, Junzhou Huang, and Chuang Gan · 2019
Earlier work this paper cites.
Progressive boundary refinement network for temporal action detection
Qinying Liu and Zilei Wang · 2020
Cited alongside, same era.
Haisheng Su, Weihao Gan, Wei Wu, Yu Qiao, and Junjie Yan · 2020
Cited alongside, same era.
Boundary-sensitive pre-training for temporal localization in videos
Mengmeng Xu, Juan-Manuel Pérez-Rúa, Victor Escorcia, Brais Martinez, Xiatian Zhu, Li Zhang, Bernard Ghanem, and Tao Xiang · 2020
Cited alongside, same era.
G-tad: Sub-graph localization for temporal action detection
Mengmeng Xu, Chen Zhao, David S Rojas, Ali Thabet, and Bernard Ghanem · 2020
Cited alongside, same era.
Distribution-aware coordinate representation for human pose estimation
Feng Zhang, Xiatian Zhu, Hanbin Dai, Mao Ye, and Ce Zhu · 2020
Cited alongside, same era.
Relaxed transformer decoders for direct action proposal generation
Jing Tan, Jiaqi Tang, Limin Wang, and Gangshan Wu · 2021
Later among the works it cites.
Self-supervised learning for semi-supervised temporal action proposal
Xiang Wang, Shiwei Zhang, Zhiwu Qing, Yuanjie Shao, Changxin Gao, and Nong Sang · 2021
Later among the works it cites.
Boundary-sensitive pre-training for temporal localization in videos
Mengmeng Xu, Juan-Manuel Pérez-Rúa, Victor Escorcia, Brais Martinez, Xiatian Zhu, Li Zhang, Bernard Ghanem, and Tao Xiang · 2021
Later among the works it cites.
Low-fidelity end-to-end video encoder pre-training for temporal action localization
Mengmeng Xu, Juan-Manuel Perez-Rua, Xiatian Zhu, Bernard Ghanem, and Brais Martinez · 2021
Later among the works it cites.
Cola: Weakly-supervised temporal action localization with snippet contrastive learning
Can Zhang, Meng Cao, Dongming Yang, Jie Chen, and Yuexian Zou · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cross-modal consensus network for weakly supervised temporal action localization
Fa-Ting Hong, Jia-Chang Feng, Dan Xu, Ying Shan, and Wei-Shi Zheng · 2021
Cited alongside, same era.
Learning salient boundary feature for anchor-free temporal action localization
Chuming Lin, Chengming Xu, Donghao Luo, Yabiao Wang, Ying Tai, Chengjie Wang, Jilin Li, Feiyue Huang, and Yanwei Fu · 2021
Cited alongside, same era.
Multi-shot temporal event localization: a benchmark
Xiaolong Liu, Yao Hu, Song Bai, Fei Ding, Xiang Bai, and Philip HS Torr · 2021
Cited alongside, same era.
The blessings of unlabeled background in untrimmed videos
Yuan Liu, Jingyuan Chen, Zhenfang Chen, Bing Deng, Jianqiang Huang, and Hanwang Zhang · 2021
Cited alongside, same era.
Weakly supervised action selection learning in video
Junwei Ma, Satya Krishna Gorti, Maksims Volkovs, and Guangwei Yu · 2021
Cited alongside, same era.
Few-shot temporal action localization with query adaptive transformer
Sauradip Nag, Xiatian Zhu, and Tao Xiang · 2021
Cited alongside, same era.
Temporal context aggregation network for temporal action proposal refinement
Zhiwu Qing, Haisheng Su, Weihao Gan, Dongliang Wang, Wei Wu, Xiang Wang, Yu Qiao, Junjie Yan, Changxin Gao, and Nong Sang · 2021
Cited alongside, same era.
Video self-stitching graph network for temporal action localization
Chen Zhao, Ali K Thabet, and Bernard Ghanem · 2021
Later among the works it cites.
Dcan: improving temporal action detection via dual context aggregation
Guo Chen, Yin-Dong Zheng, Limin Wang, and Tong Lu · 2022
Closest in time.
Dual-evidential learning for weakly-supervised temporal action localization
Mengyuan Chen, Junyu Gao, Shicai Yang, and Changsheng Xu · 2022
Closest in time.
Asm-loc: Action-aware segment modeling for weakly-supervised temporal action localization
Bo He, Xitong Yang, Le Kang, Zhiyu Cheng, Xin Zhou, and Abhinav Shrivastava · 2022
Closest in time.
Proposal-free temporal action detection via global segmentation mask learning
Sauradip Nag, Xiatian Zhu, Yi-zhe Song, and Tao Xiang · 2022
Closest in time.
Semi-supervised temporal action detection with proposal-free masking
Sauradip Nag, Xiatian Zhu, Yi-zhe Song, and Tao Xiang · 2022
Closest in time.
Zero-shot temporal action detection via vision-language prompting
Sauradip Nag, Xiatian Zhu, Yi-zhe Song, and Tao Xiang · 2022
Closest in time.
React: Temporal action detection with relational queries
Dingfeng Shi, Yujie Zhong, Qiong Cao, Jing Zhang, Lin Ma, Jia Li, and Dacheng Tao · 2022
Closest in time.
Actionformer: Localizing moments of actions with transformers
Chen-Lin Zhang, Jianxin Wu, and Yin Li · 2022
Closest in time.