Fetching the paper…
Reading the bibliography…
Weakly-supervised temporal action localization aims to localize actions in untrimmed videos with only video-level action category labels.
L. Devroye, “Sample-based non-uniform random variate generation,” in Proceedings of the 18th conference on Winter simulation . ACM, 1986, pp. 260–265
1986
Earlier work this paper cites.
T. G. Dietterich et al. , “Ensemble learning,” The handbook of brain theory and neural networks , vol. 2, pp. 110–125, 2002
2002
Earlier work this paper cites.
L. Wolf, M. Guttmann, and D. Cohen-Or, “Non-homogeneous content-driven video-retargeting,” in Proceedings of the IEEE International Conference on Computer Vision (ICCV) . IEEE, 2007, pp. 1–6
2007
Earlier work this paper cites.
Z. Karni, D. Freedman, and C. Gotsman, “Energy-based image deformation,” in Computer Graphics Forum , vol. 28, no. 5. Wiley Online Library, 2009, pp. 1257–1268
2009
Earlier work this paper cites.
A. Wedel, T. Pock, C. Zach, H. Bischof, and D. Cremers, “An improved algorithm for tv-l 1 optical flow,” in Statistical and geometrical approaches to visual motion analysis . Springer, 2009, pp. 23–45
2009
Earlier work this paper cites.
Z.-H. Zhou, “Ensemble learning.” Encyclopedia of biometrics , vol. 1, pp. 270–273, 2009
2009
Earlier work this paper cites.
R. Polikar, “Ensemble learning,” in Ensemble machine learning . Springer, 2012, pp. 1–34
2012
Earlier work this paper cites.
Y.-G. Jiang, J. Liu, A. R. Zamir, G. Toderici, I. Laptev, M. Shah, and R. Sukthankar, “Thumos challenge: Action recognition with a large number of classes,” 2014
2014
Earlier work this paper cites.
F. Caba Heilbron, V. Escorcia, B. Ghanem, and J. Carlos Niebles, “Activitynet: A large-scale video benchmark for human activity understanding,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2015, pp. 961–970
2015
Earlier work this paper cites.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,” in Proceedings of the International Conference on Learning Representations (ICLR) , 2015
2015
Earlier work this paper cites.
Z. Shou, D. Wang, and S.-F. Chang, “Temporal action localization in untrimmed videos via multi-stage cnns,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2016, pp. 1049–1058
2016
Earlier work this paper cites.
B. Zhou, A. Khosla, A. Lapedriza, A. Oliva, and A. Torralba, “Learning deep features for discriminative localization,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2016, pp. 2921–2929
2016
Earlier work this paper cites.
J. Gao, Z. Yang, K. Chen, C. Sun, and R. Nevatia, “Turn tap: Temporal unit regression network for temporal action proposals,” in Proceedings of the IEEE International Conference on Computer Vision (ICCV) , 2017, pp. 3628–3636
2017
Earlier work this paper cites.
Y. Zhao, Y. Xiong, L. Wang, Z. Wu, X. Tang, and D. Lin, “Temporal action detection with structured segment networks,” in Proceedings of the IEEE International Conference on Computer Vision (ICCV) , 2017, pp. 2914–2923
2017
Earlier work this paper cites.
Z. Shou, J. Chan, A. Zareian, K. Miyazawa, and S.-F. Chang, “Cdc: Convolutional-de-convolutional networks for precise temporal action localization in untrimmed videos,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2017, pp. 5734–5743
2017
Earlier work this paper cites.
T. Lin, X. Zhao, and Z. Shou, “Single shot temporal action detection,” in Proceedings of the ACM international conference on Multimedia (MM) , 2017, pp. 988–996
2017
Earlier work this paper cites.
H. Xu, A. Das, and K. Saenko, “R-c3d: Region convolutional 3d network for temporal activity detection,” in Proceedings of the IEEE International Conference on Computer Vision (ICCV) , 2017, pp. 5783–5792
2017
Earlier work this paper cites.
L. Wang, Y. Xiong, D. Lin, and L. Van Gool, “Untrimmednets for weakly supervised action recognition and detection,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2017, pp. 4325–4334
2017
Earlier work this paper cites.
J. Carreira and A. Zisserman, “Quo vadis, action recognition? a new model and the kinetics dataset,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2017, pp. 6299–6308
2017
Earlier work this paper cites.
A. Paszke, S. Gross, S. Chintala, G. Chanan, E. Yang, Z. DeVito, Z. Lin, A. Desmaison, L. Antiga, and A. Lerer, “Automatic differentiation in pytorch,” in Advances in neural information processing systems (NIPS) , 2017
2017
Earlier work this paper cites.
Y.-W. Chao, S. Vijayanarasimhan, B. Seybold, D. A. Ross, J. Deng, and R. Sukthankar, “Rethinking the faster r-cnn architecture for temporal action localization,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 1130–1139
2018
Earlier work this paper cites.
T. Lin, X. Zhao, H. Su, C. Wang, and M. Yang, “Bsn: Boundary sensitive network for temporal action proposal generation,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2018, pp. 3–19
2018
Earlier work this paper cites.
P. Nguyen, T. Liu, G. Prasad, and B. Han, “Weakly supervised action localization by sparse temporal pooling network,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 6752–6761
2018
Earlier work this paper cites.
S. Paul, S. Roy, and A. K. Roy-Chowdhury, “W-talc: Weakly-supervised temporal activity localization and classification,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2018, pp. 563–579
2018
Cited alongside, same era.
Z. Shou, H. Gao, L. Zhang, K. Miyazawa, and S.-F. Chang, “Autoloc: Weakly-supervised temporal action localization in untrimmed videos,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2018, pp. 154–171
2018
Cited alongside, same era.
H. Su, X. Zhao, and T. Lin, “Cascaded pyramid mining network for weakly supervised temporal action localization,” in Proceedings of the Asian Conference on Computer Vision (ACCV) . Springer, 2018, pp. 558–574
2018
Cited alongside, same era.
H. Song, X. Wu, B. Zhu, Y. Wu, M. Chen, and Y. Jia, “Temporal action localization in untrimmed videos using action pattern trees,” IEEE Transactions on Multimedia (T-MM) , vol. 21, no. 3, pp. 717–730, 2018
2018
Cited alongside, same era.
Y. Xu, C. Zhang, Z. Cheng, J. Xie, Y. Niu, S. Pu, and F. Wu, “Segregated temporal assembly recurrent networks for weakly supervised multiple action detection,” in Proceedings of the AAAI Conference on Artificial Intelligence (AAAI) , vol. 34, 2019, pp. 9070–9078
2019
Later among the works it cites.
S. Narayan, H. Cholakkal, F. S. Khan, and L. Shao, “3c-net: Category count and center loss for weakly-supervised action localization,” in Proceedings of the IEEE International Conference on Computer Vision (ICCV) , 2019, pp. 8679–8687
2019
Later among the works it cites.
C. Zhang, Y. Xu, Z. Cheng, Y. Niu, S. Pu, F. Wu, and F. Zou, “Adversarial seeded sequence growing for weakly-supervised temporal action localization,” in Proceedings of the ACM international conference on Multimedia (MM) . ACM, 2019, pp. 738–746
2019
Later among the works it cites.
T. Yu, Z. Ren, Y. Li, E. Yan, N. Xu, and J. Yuan, “Temporal structure mining for weakly supervised action detection,” in Proceedings of the IEEE International Conference on Computer Vision (ICCV) , 2019, pp. 5522–5531
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D. Guo, W. Li, and X. Fang, “Fully convolutional network for multiscale temporal action proposals,” IEEE Transactions on Multimedia (T-MM) , vol. 20, no. 12, pp. 3428–3438, 2018
2018
Cited alongside, same era.
M. Schlichtkrull, T. N. Kipf, P. Bloem, R. Van Den Berg, I. Titov, and M. Welling, “Modeling relational data with graph convolutional networks,” in European semantic web conference (ESWC) . Springer, 2018, pp. 593–607
2018
Cited alongside, same era.
J. Gao, K. Chen, and R. Nevatia, “Ctap: Complementary temporal action proposal generation,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2018, pp. 68–83
2018
Cited alongside, same era.
A. Recasens, P. Kellnhofer, S. Stent, W. Matusik, and A. Torralba, “Learning to zoom: a saliency-based sampling layer for neural networks,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2018, pp. 51–66
2018
Cited alongside, same era.
T. Lin, X. Liu, X. Li, E. Ding, and S. Wen, “Bmn: Boundary-matching network for temporal action proposal generation,” in Proceedings of the IEEE International Conference on Computer Vision (ICCV) , 2019, pp. 3889–3898
2019
Cited alongside, same era.
C. Lin, J. Li, Y. Wang, Y. Tai, D. Luo, Z. Cui, C. Wang, J. Li, F. Huang, and R. Ji, “Fast learning of temporal action proposal via dense boundary generator,” 2019
2019
Cited alongside, same era.
R. Zeng, C. Gan, P. Chen, W. Huang, Q. Wu, and M. Tan, “Breaking winner-takes-all: Iterative-winners-out networks for weakly supervised temporal action localization,” IEEE Transactions on Image Processing (T-IP) , vol. 28, no. 12, pp. 5797–5808, 2019
2019
Cited alongside, same era.
D. Liu, T. Jiang, and Y. Wang, “Completeness modeling and context separation for weakly supervised temporal action localization,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2019, pp. 1298–1307
2019
Cited alongside, same era.
2019
Later among the works it cites.
P. Lee, Y. Uh, and H. Byun, “Background suppression network for weakly-supervised temporal action localization,” in Proceedings of the AAAI Conference on Artificial Intelligence (AAAI) , vol. 34, 2020, pp. 11 320–11 327
2020
Later among the works it cites.
B. Shi, Q. Dai, Y. Mu, and J. Wang, “Weakly-supervised action localization by generative attention modeling,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2020, pp. 1009–1019
2020
Later among the works it cites.
Y. Zhai, L. Wang, W. Tang, Q. Zhang, J. Yuan, and G. Hua, “Two-stream consensus network for weakly-supervised temporal action localization,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2020, pp. 37–54
2020
Later among the works it cites.
Z. Luo, D. Guillory, B. Shi, W. Ke, F. Wan, T. Darrell, and H. Xu, “Weakly-supervised action localization with expectation-maximization multi-instance learning,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2020, pp. 729–745
2020
Later among the works it cites.
Y. Zhou, R. Wang, H. Li, and S. Y. Kung, “Temporal action localization using long short-term dependency,” IEEE Transactions on Multimedia (T-MM) , 2020
2020
Later among the works it cites.
P. Zhao, L. Xie, C. Ju, Y. Zhang, Y. Wang, and Q. Tian, “Bottom-up temporal action localization with mutual regularization,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2020, pp. 539–555
2020
Later among the works it cites.
Y. Bai, Y. Wang, Y. Tong, Y. Yang, Q. Liu, and J. Liu, “Boundary content graph neural network for temporal action proposal generation,” in Proceedings of the European Conference on Computer Vision (ECCV) . Springer, 2020, pp. 121–137
2020
Later among the works it cites.
M. Xu, C. Zhao, D. S. Rojas, A. Thabet, and B. Ghanem, “G-tad: Sub-graph localization for temporal action detection,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2020, pp. 10 156–10 165
2020
Later among the works it cites.
J. Gao, Z. Shi, G. Wang, J. Li, Y. Yuan, S. Ge, and X. Zhou, “Accurate temporal action proposal generation with relation-aware pyramid network,” in Proceedings of the AAAI Conference on Artificial Intelligence (AAAI) , vol. 34, no. 07, 2020, pp. 10 810–10 817
2020
Later among the works it cites.
G. Chen, C. Zhang, and Y. Zou, “Afnet: Temporal locality-aware network with dual structure for accurate and fast action detection,” IEEE Transactions on Multimedia (T-MM) , 2020
2020
Later among the works it cites.
L. Yang, H. Peng, D. Zhang, J. Fu, and J. Han, “Revisiting anchor mechanisms for temporal action localization,” IEEE Transactions on Image Processing (T-IP) , vol. 29, pp. 8535–8548, 2020
2020
Later among the works it cites.
Q. Liu and Z. Wang, “Progressive boundary refinement network for temporal action detection,” in Proceedings of the AAAI Conference on Artificial Intelligence (AAAI) , vol. 34, no. 07, 2020, pp. 11 612–11 619
2020
Later among the works it cites.
K. Min and J. J. Corso, “Adversarial background-aware loss for weakly-supervised temporal activity localization,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2020, pp. 283–299
2020
Later among the works it cites.
H. Su, X. Zhao, T. Lin, S. Liu, and Z. Hu, “Transferable knowledge-based multi-granularity fusion network for weakly supervised temporal action detection,” IEEE Transactions on Multimedia (T-MM) , 2020
2020
Later among the works it cites.
P. Lee, J. Wang, Y. Lu, and H. Byun, “Background modeling via uncertainty estimation for weakly-supervised action localization,” 2020
2020
Later among the works it cites.
A. Pardo, H. Alwassel, F. Caba, A. Thabet, and B. Ghanem, “Refineloc: Iterative refinement for weakly-supervised action localization,” in Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) , 2021, pp. 3319–3328
2021
Closest in time.
C. Sun, H. Song, X. Wu, Y. Jia, and J. Luo, “Exploiting informative video segments for temporal action localization,” IEEE Transactions on Multimedia (T-MM) , 2021
2021
Closest in time.