Fetching the paper…
Reading the bibliography…
This paper simultaneously addresses three limitations associated with conventional skeleton-based action recognition; skeleton detection and tracking errors, poor variety of the targeted actions, as well as person-wise and frame-wise action recognition.
Solving the Multiple Instance Problem with Axis-parallel Rectangles
Thomas G. Dietterich, Richard H. Lathrop, and Tomás Lozano-Pérez · 1997
Earlier work this paper cites.
Violence Detection in Video Using Computer Vision Techniques
Enrique Bermejo Nievas, Oscar Deniz Suarez, Gloria Bueno García, and Rahul Sukthankar · 2011
Earlier work this paper cites.
HMDB: A Large Video Database for Human Motion Recognition
H. Kuehne, H. Jhuang, E. Garrote, T. Poggio, and T. Serre · 2011
Earlier work this paper cites.
Movies Fight Detection Dataset
Enrique Bermejo Nievas, Oscar Deniz Suarez, Gloria Bueno Garcia, and Rahul Sukthankar · 2011
Earlier work this paper cites.
Violent Flows: Real-Time Detection of Violent Crowd Behavior
Tal Hassner, Yossi Itcher, and Orit Kliper-Gross · 2012
Earlier work this paper cites.
UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild
Khurram Soomro, Amir Roshan Zamir, and Mubarak Shah · 2012
Earlier work this paper cites.
Microsoft COCO: Common Objects in Context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C. Lawrence Zitnick · 2014
Earlier work this paper cites.
Two-Stream Convolutional Networks for Action Recognition in Videos
Karen Simonyan and Andrew Zisserman · 2014
Earlier work this paper cites.
Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks
Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun · 2015
Earlier work this paper cites.
Learning Spatiotemporal Features With 3D Convolutional Networks
Du Tran, Lubomir Bourdev, Rob Fergus, Lorenzo Torresani, and Manohar Paluri · 2015
Earlier work this paper cites.
Beyond Short Snippets: Deep Networks for Video Classification
Joe Yue-Hei Ng, Matthew Hausknecht, Sudheendra Vijayanarasimhan, Oriol Vinyals, Rajat Monga, and George Toderici · 2015
Earlier work this paper cites.
Spatiotemporal Residual Networks for Video Action Recognition
Christoph Feichtenhofer, Axel Pinz, and Richard P. Wildes · 2016
Earlier work this paper cites.
Multimodal Human Action Recognition in Assistive Human-robot Interaction
I. Rodomagoulakis, N. Kardaris, V. Pitsikalis, E. Mavroudi, A. Katsamanis, A. Tsiami, and P. Maragos · 2016
Earlier work this paper cites.
Realtime Multi-Person 2D Pose Estimation Using Part Affinity Fields
Zhe Cao, Tomas Simon, Shih-En Wei, and Yaser Sheikh · 2017
Earlier work this paper cites.
Quo Vadis, Action Recognition? A New Model and the Kinetics Dataset
Joao Carreira and Andrew Zisserman · 2017
Earlier work this paper cites.
PointNet: Deep Learning on Point Sets for 3D Classification and Segmentation
Charles R. Qi, Hao Su, Kaichun Mo, and Leonidas J. Guibas · 2017
Earlier work this paper cites.
PointNet++: Deep Hierarchical Feature Learning on Point Sets in a Metric Space
Charles R. Qi, Li Yi, Hao Su, and Leonidas J Guibas · 2017
Earlier work this paper cites.
Online Real-Time Multiple Spatiotemporal Action Localisation and Prediction
Gurkirt Singh, Suman Saha, Michael Sapienza, Philip H. S. Torr, and Fabio Cuzzolin · 2017
Earlier work this paper cites.
Attention is All you Need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Ł ukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
A Flexible Model for Training Action Localization with Varying Levels of Supervision
Guilhem Chéron, Jean-Baptiste Alayrac, Ivan Laptev, and Cordelia Schmid · 2018
Earlier work this paper cites.
Pose Proposal Networks
Taiki Sekii · 2018
Earlier work this paper cites.
A Closer Look at Spatiotemporal Convolutions for Action Recognition
Du Tran, Heng Wang, Lorenzo Torresani, Jamie Ray, Yann LeCun, and Manohar Paluri · 2018
Cited alongside, same era.
Pelee: A Real-Time Object Detection System on Mobile Devices
Robert J. Wang, Xiang Li, and Charles X. Ling · 2018
Cited alongside, same era.
Rethinking Spatiotemporal Feature Learning: Speed-Accuracy Trade-offs in Video Classification
Saining Xie, Chen Sun, Jonathan Huang, Zhuowen Tu, and Kevin Murphy · 2018
Cited alongside, same era.
Spatial Temporal Graph Convolutional Networks for Skeleton-Based Action Recognition
Sijie Yan, Yuanjun Xiong, and Dahua Lin · 2018
Cited alongside, same era.
Temporal Attentive Alignment for Large-Scale Video Domain Adaptation
Min-Hung Chen, Zsolt Kira, Ghassan AlRegib, Jaekwon Yoo, Ruxin Chen, and Jian Zheng · 2019
Cited alongside, same era.
Why Can’t I Dance in the Mall? Learning to Mitigate Scene Bias in Action Recognition
Self-supervised Keypoint Correspondences for Multi-person Pose Estimation and Tracking in Videos
Umer Rafi, Andreas Doering, Bastian Leibe, and Juergen Gall · 2020
Later among the works it cites.
15 Keypoints Is All You Need
Michael Snower, Asim Kadav, Farley Lai, and Hans Peter Graf · 2020
Later among the works it cites.
Human Interaction Learning on 3D Skeleton Point Clouds for Video Violence Recognition
Yukun Su, Guosheng Lin, Jinhui Zhu, and Qingyao Wu · 2020
Later among the works it cites.
Semantics-Guided Neural Networks for Efficient Skeleton-Based Human Action Recognition
Pengfei Zhang, Cuiling Lan, Wenjun Zeng, Junliang Xing, Jianru Xue, and Nanning Zheng · 2020
Later among the works it cites.
ViViT: A Video Vision Transformer
Anurag Arnab, Mostafa Dehghani, Georg Heigold, Chen Sun, Mario Lučić, and Cordelia Schmid · 2021
Later among the works it cites.
JOLO-GCN: Mining Joint-Centered Light-Weight Information for Skeleton-Based Action Recognition
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jinwoo Choi, Chen Gao, C. E. Joseph Messou, and Jia-Bin Huang · 2019
Cited alongside, same era.
SlowFast Networks for Video Recognition
Christoph Feichtenhofer, Haoqi Fan, Jitendra Malik, and Kaiming He · 2019
Cited alongside, same era.
Video Action Transformer Network
Rohit Girdhar, Joao Carreira, Carl Doersch, and Andrew Zisserman · 2019
Cited alongside, same era.
A Model-Based Human Activity Recognition for Human–Robot Collaboration
Sang Uk Lee, Andreas Hofmann, and Brian Williams · 2019
Cited alongside, same era.
Actional-Structural Graph Convolutional Networks for Skeleton-Based Action Recognition
Maosen Li, Siheng Chen, Xu Chen, Ya Zhang, Yanfeng Wang, and Qi Tian · 2019
Cited alongside, same era.
Two-Stream Adaptive Graph Convolutional Networks for Skeleton-Based Action Recognition
Lei Shi, Yifan Zhang, Jian Cheng, and Hanqing Lu · 2019
Cited alongside, same era.
Deep High-Resolution Representation Learning for Human Pose Estimation
Ke Sun, Bin Xiao, Dong Liu, and Jingdong Wang · 2019
Cited alongside, same era.
Jinmiao Cai, Nianjuan Jiang, Xiaoguang Han, Kui Jia, and Jiangbo Lu · 2021
Later among the works it cites.
Learning Multi-Granular Spatio-Temporal Graph Network for Skeleton-Based Action Recognition
Tailin Chen, Desen Zhou, Jian Wang, Shidong Wang, Yu Guan, Xuming He, and Errui Ding · 2021
Later among the works it cites.
Multi-Scale Spatial Temporal Graph Convolutional Network for Skeleton-Based Action Recognition
Zhan Chen, Sicheng Li, Bing Yang, Qinghan Li, and Hong Liu · 2021
Later among the works it cites.
RWF-2000: An Open Large Scale Video Database for Violence Detection
Ming Cheng, Kunjing Cai, and Ming Li · 2021
Later among the works it cites.
Efficient Two-Stream Network for Violence Detection Using Separable Convolutional LSTM
Zahidul Islam, Mohammad Rukonuzzaman, Raiyan Ahmed, Md. Hasanul Kabir, and Moshiur Farazi · 2021
Later among the works it cites.
IntegralAction: Pose-Driven Feature Integration for Robust Human Action Recognition in Videos
Gyeongsik Moon, Heeseung Kwon, Kyoung Mu Lee, and Minsu Cho · 2021
Later among the works it cites.
Actor-Context-Actor Relation Network for Spatio-Temporal Action Localization
Junting Pan, Siyu Chen, Mike Zheng Shou, Yu Liu, Jing Shao, and Hongsheng Li · 2021
Later among the works it cites.
Object Detection Method and Object Detection Device
Taiki Sekii · 2021
Later among the works it cites.
Mimetics: Towards Understanding Human Actions out of Context
Philippe Weinzaepfel and Grégory Rogez · 2021
Later among the works it cites.
Point Transformer
Hengshuang Zhao, Li Jiang, Jiaya Jia, Philip H.S. Torr, and Vladlen Koltun · 2021
Later among the works it cites.
InfoGCN: Representation Learning for Human Skeleton-Based Action Recognition
Hyung-gun Chi, Myoung Hoon Ha, Seunggeun Chi, Sang Wan Lee, Qixing Huang, and Karthik Ramani · 2022
Later among the works it cites.
Dual-Head Contrastive Domain Adaptation for Video Action Recognition
Victor G. Turrisi da Costa, Giacomo Zara, Paolo Rota, Thiago Oliveira-Santos, Nicu Sebe, Vittorio Murino, and Elisa Ricci · 2022
Later among the works it cites.
Revisiting Skeleton-based Action Recognition
Haodong Duan, Yue Zhao, Kai Chen, Dian Shao, Dahua Lin, and Bo Dai · 2022
Later among the works it cites.
End-to-End Semi-Supervised Learning for Video Action Detection
Akash Kumar and Yogesh Singh Rawat · 2022
Later among the works it cites.
Video Swin Transformer
Ze Liu, Jia Ning, Yue Cao, Yixuan Wei, Zheng Zhang, Stephen Lin, and Han Hu · 2022
Later among the works it cites.