Fetching the paper…
Reading the bibliography…
Video violence recognition based on deep learning concerns accurate yet scalable human violence recognition.
L. Van der Maaten, G. Hinton, Visualizing data using t-sne., Journal of machine learning research 9 (11) (2008)
2008
Earlier work this paper cites.
Y. Gong, W. Wang, S. Jiang, Q. Huang, W. Gao, Detecting violent scenes in movies by auditory and visual cues, in: Advances in Multimedia Information Processing-PCM 2008: 9th Pacific Rim Conference on Multimedia, Tainan, Taiwan, December 9-13, 2008. Proceedings 9, Springer, 2008, pp. 317–326
2008
Earlier work this paper cites.
O. Barnich, M. Van Droogenbroeck, Vibe: a powerful random technique to estimate the background in video sequences, in: 2009 IEEE international conference on acoustics, speech and signal processing, IEEE, 2009, pp. 945–948
2009
Earlier work this paper cites.
L. Pessoa, R. Adolphs, Emotion processing and the amygdala: from a’low road’to’many roads’ of evaluating biological significance, Nature reviews neuroscience 11 (11) (2010) 773–782
2010
Earlier work this paper cites.
E. Bermejo Nievas, O. Deniz Suarez, G. Bueno García, R. Sukthankar, Violence detection in video using computer vision techniques, in: Computer Analysis of Images and Patterns: 14th International Conference, CAIP 2011, Seville, Spain, August 29-31, 2011, Proceedings, Part II 14, Springer, 2011, pp. 332–339
2011
Earlier work this paper cites.
T. Hassner, Y. Itcher, O. Kliper-Gross, Violent flows: Real-time detection of violent crowd behavior, in: 2012 IEEE computer society conference on computer vision and pattern recognition workshops, IEEE, 2012, pp. 1–6
2012
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, et al., Human-level control through deep reinforcement learning, nature 518 (7540) (2015) 529–533
2015
Earlier work this paper cites.
X. Wang, L. Gao, J. Song, H. Shen, Beyond frame-level cnn: saliency-aware 3-d cnn with lstm for video action recognition, IEEE signal processing letters 24 (4) (2016) 510–514
2016
Earlier work this paper cites.
J. Carreira, A. Zisserman, Quo vadis, action recognition? a new model and the kinetics dataset, in: proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, 2017, pp. 6299–6308
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, I. Polosukhin, Attention is all you need, Advances in neural information processing systems 30 (2017)
2017
Earlier work this paper cites.
T. Suzuki, T. Itazuri, K. Hara, H. Kataoka, Learning spatiotemporal 3d convolution with video order self-supervision, in: Proceedings of the European Conference on Computer Vision (ECCV) Workshops, 2018, pp. 0–0
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
X. Wang, R. Girshick, A. Gupta, K. He, Non-local neural networks, in: Proceedings of the IEEE conference on computer vision and pattern recognition, 2018, pp. 7794–7803
2018
Earlier work this paper cites.
H. Ge, Z. Yan, W. Yu, L. Sun, An attention mechanism based convolutional lstm network for video action recognition, Multimedia Tools and Applications 78 (2019) 20533–20556
2019
Earlier work this paper cites.
L. Wang, P. Koniusz, D. Q. Huynh, Hallucinating idt descriptors and i3d optical flow features for action recognition with cnns, in: Proceedings of the IEEE/CVF International Conference on Computer Vision, 2019, pp. 8698–8708
2019
Earlier work this paper cites.
P. Ramachandran, N. Parmar, A. Vaswani, I. Bello, A. Levskaya, J. Shlens, Stand-alone self-attention in vision models, Advances in neural information processing systems 32 (2019)
2019
Earlier work this paper cites.
Z. Qiu, T. Yao, C.-W. Ngo, X. Tian, T. Mei, Learning spatio-temporal representation with local and global diffusion, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2019, pp. 12056–12065
2019
Earlier work this paper cites.
C. Feichtenhofer, H. Fan, J. Malik, K. He, Slowfast networks for video recognition, in: Proceedings of the IEEE/CVF international conference on computer vision, 2019, pp. 6202–6211
2019
Earlier work this paper cites.
L. Sevilla-Lara, Y. Liao, F. Güney, V. Jampani, A. Geiger, M. J. Black, On the integration of optical flow and action recognition, in: Pattern Recognition: 40th German Conference, GCPR 2018, Stuttgart, Germany, October 9-12, 2018, Proceedings 40, Springer, 2019, pp. 281–297
2019
Earlier work this paper cites.
J. Hwangbo, J. Lee, A. Dosovitskiy, D. Bellicoso, V. Tsounis, V. Koltun, M. Hutter, Learning agile and dynamic motor skills for legged robots, Science Robotics 4 (26) (2019) eaau5872
2019
Earlier work this paper cites.
Y.-D. Zheng, Z. Liu, T. Lu, L. Wang, Dynamic sampling networks for efficient action recognition in videos, IEEE transactions on image processing 29 (2020) 7970–7983
2020
Earlier work this paper cites.
Y. Su, G. Lin, J. Zhu, Q. Wu, Human interaction learning on 3d skeleton point clouds for video violence recognition, in: Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part IV 16, Springer, 2020, pp. 74–90
2020
Earlier work this paper cites.
Z. Zhang, Z. Lv, C. Gan, Q. Zhu, Human action recognition using convolutional lstm and fully-connected lstm with different attentions, Neurocomputing 410 (2020) 304–316
2020
Earlier work this paper cites.
2020
Cited alongside, same era.
X. Yang, An overview of the attention mechanisms in computer vision, in: Journal of Physics: Conference Series, Vol. 1693, IOP Publishing, 2020, p. 012173
2020
Cited alongside, same era.
J. Sun, J. Jiang, Y. Liu, An introductory survey on attention mechanisms in computer vision problems, in: 2020 6th International Conference on Big Data and Information Analytics (BigDIA), IEEE, 2020, pp. 295–300
2020
Cited alongside, same era.
K. Wang, B. Kang, J. Shao, J. Feng, Improving generalization in reinforcement learning with mixture regularization, Advances in Neural Information Processing Systems 33 (2020) 7968–7978
2020
Cited alongside, same era.
A. Raffin, A. Hill, A. Gleave, A. Kanervisto, M. Ernestus, N. Dormann, Stable-baselines3: Reliable reinforcement learning implementations , Journal of Machine Learning Research 22 (268) (2021) 1–8. URL http://jmlr.org/papers/v22/20-1364.html
2021
Later among the works it cites.
N. Mumtaz, N. Ejaz, S. Habib, S. M. Mohsin, P. Tiwari, S. S. Band, N. Kumar, An overview of violence detection techniques: current challenges and future directions, Artificial intelligence review (2022) 1–26
2022
Later among the works it cites.
Y. Kong, Y. Fu, Human action recognition and prediction: A survey, International Journal of Computer Vision 130 (5) (2022) 1366–1401
2022
Later among the works it cites.
Y. Hao, J. Li, N. Wang, X. Wang, X. Gao, Spatiotemporal consistency-enhanced network for video anomaly detection, Pattern Recognition 121 (2022) 108232
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. Wolf, L. Debut, V. Sanh, J. Chaumond, C. Delangue, A. Moi, P. Cistac, T. Rault, R. Louf, M. Funtowicz, J. Davison, S. Shleifer, P. von Platen, C. Ma, Y. Jernite, J. Plu, C. Xu, T. L. Scao, S. Gugger, M. Drame, Q. Lhoest, A. M. Rush, Transformers: State-of-the-art natural language processing , in: Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing: System Demonstrations, Association for Computational Linguistics, Online, 2020, pp. 38–45. URL https://www.aclweb.org/anthology/2020.emnlp-demos.6
2020
Cited alongside, same era.
Z. Islam, M. Rukonuzzaman, R. Ahmed, M. H. Kabir, M. Farazi, Efficient two-stream network for violence detection using separable convolutional lstm, in: 2021 International Joint Conference on Neural Networks (IJCNN), IEEE, 2021, pp. 1–8
2021
Cited alongside, same era.
M. Cheng, K. Cai, M. Li, Rwf-2000: an open large scale video database for violence detection, in: 2020 25th International Conference on Pattern Recognition (ICPR), IEEE, 2021, pp. 4183–4190
2021
Cited alongside, same era.
R. Portsev, A. Makarenko, Comparative analysis of 3d convolutional and lstm neural networks in the action recognition task by video data, in: Journal of Physics: Conference Series, Vol. 1864, IOP Publishing, 2021, p. 012015
2021
Cited alongside, same era.
B. T. Hung, V. B. Semwal, N. Gaud, V. Bijalwan, Violent video detection by pre-trained model and cnn-lstm approach, in: Proceedings of Integrated Intelligence Enable Networks and Computing: IIENC 2020, Springer, 2021, pp. 979–989
2021
Cited alongside, same era.
N. Honarjoo, A. Abdari, A. Mansouri, Violence detection using pre-trained models, in: 2021 5th International Conference on Pattern Recognition and Image Analysis (IPRIA), IEEE, 2021, pp. 1–4
2021
Cited alongside, same era.
A. Vaswani, P. Ramachandran, A. Srinivas, N. Parmar, B. Hechtman, J. Shlens, Scaling local self-attention for parameter efficient visual backbones, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2021, pp. 12894–12904
2021
Cited alongside, same era.
2021
Cited alongside, same era.
H. Duan, Y. Zhao, K. Chen, D. Lin, B. Dai, Revisiting skeleton-based action recognition, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2022, pp. 2969–2978
2022
Later among the works it cites.
X. Pan, C. Ge, R. Lu, S. Song, G. Chen, Z. Huang, G. Huang, On the integration of self-attention and convolution, in: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, 2022, pp. 815–825
2022
Later among the works it cites.
S. Khan, M. Naseer, M. Hayat, S. W. Zamir, F. S. Khan, M. Shah, Transformers in vision: A survey, ACM computing surveys (CSUR) 54 (10s) (2022) 1–41
2022
Later among the works it cites.
2022
Later among the works it cites.
Y. Sun, C. Wang, A computation-efficient cnn system for high-quality brain tumor segmentation, Biomedical Signal Processing and Control 74 (2022) 103475
2022
Later among the works it cites.
X. Zhai, A. Kolesnikov, N. Houlsby, L. Beyer, Scaling vision transformers, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2022, pp. 12104–12113
2022
Later among the works it cites.
2022
Later among the works it cites.
J. Pan, A. Bulat, F. Tan, X. Zhu, L. Dudziak, H. Li, G. Tzimiropoulos, B. Martinez, Edgevits: Competing light-weight cnns on mobile devices with vision transformers, in: Computer Vision–ECCV 2022: 17th European Conference, Tel Aviv, Israel, October 23–27, 2022, Proceedings, Part XI, Springer, 2022, pp. 294–311
2022
Later among the works it cites.
K. He, X. Chen, S. Xie, Y. Li, P. Dollár, R. Girshick, Masked autoencoders are scalable vision learners, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2022, pp. 16000–16009
2022
Later among the works it cites.
Y. Zhou, T. Lei, H. Liu, N. Du, Y. Huang, V. Zhao, A. M. Dai, Q. V. Le, J. Laudon, et al., Mixture-of-experts with expert choice routing, Advances in Neural Information Processing Systems 35 (2022) 7103–7114
2022
Later among the works it cites.
N. Du, Y. Huang, A. M. Dai, S. Tong, D. Lepikhin, Y. Xu, M. Krikun, Y. Zhou, A. W. Yu, O. Firat, et al., Glam: Efficient scaling of language models with mixture-of-experts, in: International Conference on Machine Learning, PMLR, 2022, pp. 5547–5569
2022
Later among the works it cites.
W. Fedus, B. Zoph, N. Shazeer, Switch transformers: Scaling to trillion parameter models with simple and efficient sparsity, The Journal of Machine Learning Research 23 (1) (2022) 5232–5270
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
N. Le, V. S. Rathour, K. Yamazaki, K. Luu, M. Savvides, Deep reinforcement learning in computer vision: a comprehensive survey, Artificial Intelligence Review (2022) 1–87
2022
Later among the works it cites.
H. Mohammadi, E. Nazerfard, Video violence recognition and localization using a semi-supervised hard attention model, Expert Systems with Applications 212 (2023) 118791
2023
Closest in time.
X. Huang, Z. Cai, A review of video action recognition based on 3d convolution, Computers and Electrical Engineering 108 (2023) 108713
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.