Fetching the paper…
Reading the bibliography…
Though multimodal emotion recognition has achieved significant progress over recent years, the potential of rich synergic relationships across the modalities is not fully exploited.
Three dimensions of emotion
H. Schlosberg · 1954
Earlier work this paper cites.
An argument for basic emotions
Paul Ekman · 1992
Earlier work this paper cites.
Affective multimodal human-computer interaction
Maja Pantic, Nicu Sebe, Jeffrey F. Cohn, and Thomas Huang · 2005
Earlier work this paper cites.
Continuous prediction of spontaneous affect from multiple cues and modalities in valence-arousal space
Mihalis A. Nicolaou, Hatice Gunes, and Maja Pantic · 2011
Earlier work this paper cites.
Lstm-modeling of continuous emotions in an a-v affect recognition framework
M Wöllmer, M Kaiser, F Eyben, B Schuller, and G Rigoll · 2013
Earlier work this paper cites.
Distilling the knowledge in a neural network
Geoffrey Hinton, Oriol Vinyals, and Jeff Dean · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Training deep networks for facial expression recognition with crowd-sourced label distribution
Emad Barsoum, Cha Zhang, Cristian Canton Ferrer, and Zhengyou Zhang · 2016
Earlier work this paper cites.
Ms-celeb-1m: A dataset and benchmark for large-scale face recognition
Yandong Guo, Lei Zhang, Yuxiao Hu, Xiaodong He, and Jianfeng Gao · 2016
Earlier work this paper cites.
End-to-end multimodal er using deep neural networks
Panagiotis Tzirakis, George Trigeorgis, Mihalis A. Nicolaou, Björn W. Schuller, and Stefanos Zafeiriou · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Ł ukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Cnn architectures for large-scale audio classification
Shawn Hershey, Sourish Chaudhuri, Daniel PW Ellis, Jort F Gemmeke, Aren Jansen, R Channing Moore, Manoj Plakal, Devin Platt, Rif A Saurous, Bryan Seybold, et al · 2017
Earlier work this paper cites.
Cnn architectures for large-scale audio classification
S Hershey, S Chaudhuri, D Ellis, J F Gemmeke, A Jansen, C Moore, M Plakal, D Platt, R A Saurous, B Seybold, M Slaney, R Weiss, and K Wilson · 2017
Earlier work this paper cites.
Aff-wild: Valence and arousal ‘in-the-wild’challenge
Stefanos Zafeiriou, Dimitrios Kollias, Mihalis A Nicolaou, Athanasios Papaioannou, Guoying Zhao, and Irene Kotsia · 2017
Earlier work this paper cites.
A closer look at spatiotemporal convolutions for action recognition
D. Tran, H. Wang, L. Torresani, J. Ray, Y. LeCun, and M. Paluri · 2018
Earlier work this paper cites.
Voxceleb2: Deep speaker recognition
J. S. Chung, A. Nagrani, and A. Zisserman · 2018
Earlier work this paper cites.
Emotion recognition using fusion of audio and video features
Juan D. S. Ortega, Patrick Cardinal, and Alessandro L. Koerich · 2019
Earlier work this paper cites.
Face behavior a la carte: Expressions, affect and action units in a single network
Dimitrios Kollias, Viktoriia Sharmanska, and Stefanos Zafeiriou · 2019
Earlier work this paper cites.
Deep affect prediction in-the-wild: Aff-wild database and challenge, deep architectures, and beyond
D Kollias, P Tzirakis, M A Nicolaou, A Papaioannou, G Zhao, B Schuller, I Kotsia, and S Zafeiriou · 2019
Earlier work this paper cites.
Expression, affect, action unit recognition: Aff-wild2, multi-task learning and arcface
Dimitrios Kollias and Stefanos Zafeiriou · 2019
Earlier work this paper cites.
Deep weakly supervised domain adaptation for pain localization in videos
R Gnana Praveen, Eric Granger, and Patrick Cardinal · 2020
Earlier work this paper cites.
Two-stream aural-visual affect analysis in the wild
Felix Kuhnke, Lars Rumberg, and Jörn Ostermann · 2020
Earlier work this paper cites.
Multimodal transformer fusion for continuous emotion recognition
Jian Huang, Jianhua Tao, Bin Liu, Zheng Lian, and Mingyue Niu · 2020
Earlier work this paper cites.
Analysing affective behavior in the first abaw 2020 competition
Dimitrios Kollias, Attila Schulc, Elnar Hajiyev, and Stefanos Zafeiriou · 2020
Cited alongside, same era.
Multi-modal continuous dimensional emotion recognition using recurrent neural network and self-attention mechanism
Licai Sun, Zheng Lian, Jianhua Tao, Bin Liu, and Mingyue Niu · 2020
Cited alongside, same era.
Deep domain adaptation with ordinal regression for pain assessment using weakly-labeled videos
Gnana Praveen Rajasekhar, Eric Granger, and Patrick Cardinal · 2021
Cited alongside, same era.
Leveraging recent advances in deep learning for audio-visual emotion recognition
L Schoneveld, A Othmani, and H Abdelkawy · 2021
Cited alongside, same era.
Detecting expressions with multimodal transformers
Srinivas Parthasarathy and Shiva Sundaram · 2021
Cited alongside, same era.
Deep multimodal complementarity learning
Daheng Wang, Tong Zhao, Wenhao Yu, Nitesh V. Chawla, and Meng Jiang · 2023
Later among the works it cites.
Abaw5 challenge: A facial affect recognition approach utilizing transformer encoder and audiovisual fusion
Ziyang Zhang, Liuwei An, Zishun Cui, Ao Xu, Tengteng Dong, Yueqi Jiang, Jingyi Shi, Xin Liu, Xiao Sun, and Meng Wang · 2023
Later among the works it cites.
Audio-visual speaker verification via joint cross-attention
Gnana Praveen Rajasekhar and Jahangir Alam · 2023
Later among the works it cites.
Audio–visual fusion for emotion recognition in the valence–arousal space using joint cross-attention
R. Gnana Praveen, Patrick Cardinal, and Eric Granger · 2023
Later among the works it cites.
Recursive joint attention for audio-visual fusion in regression based emotion recognition
R Gnana Praveen, Eric Granger, and Patrick Cardinal · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jun-Tae Lee, Mihir Jain, Hyoungwoo Park, and Sungrack Yun · 2021
Cited alongside, same era.
Cross attentional audio-visual fusion for dimensional emotion recognition
R. Gnana Praveen, Eric Granger, and Patrick Cardinal · 2021
Cited alongside, same era.
Audio-visual event localization via recursive fusion by joint co-attention
Bin Duan, Hao Tang, Wei Wang, Ziliang Zong, Guowei Yang, and Yan Yan · 2021
Cited alongside, same era.
Affect analysis in-the-wild: Valence-arousal, expressions, action units and a unified framework
Dimitrios Kollias and Stefanos Zafeiriou · 2021
Cited alongside, same era.
Distribution matching for heterogeneous multi-task learning: a large-scale face study
Dimitrios Kollias, Viktoriia Sharmanska, and Stefanos Zafeiriou · 2021
Cited alongside, same era.
A multi-task mean teacher for semi-supervised facial affective behavior analysis
Lingfeng Wang, Shisen Wang, Jin Qi, and Kenji Suzuki · 2021
Cited alongside, same era.
Iterative distillation for better uncertainty estimates in multitask emotion recognition
Didan Deng, Liang Wu, and Bertram E. Shi · 2021
Cited alongside, same era.
Privileged knowledge distillation for dimensional emotion recognition in the wild
Muhammad Haseeb Aslam, Muhammad Osama Zeeshan, Marco Pedersoli, Alessandro L. Koerich, Simon Bacon, and Eric Granger · 2023
Later among the works it cites.
Decoupled multimodal distilling for emotion recognition
Yong Li, Yuanzhi Wang, and Zhen Cui · 2023
Later among the works it cites.
Leveraging tcn and transformer for effective visual-audio fusion in continuous emotion recognition
Weiwei Zhou, Jiada Lu, Zhaolong Xiong, and Weifeng Wang · 2023
Later among the works it cites.
Multi-modal facial affective analysis based on masked autoencoder
Wei Zhang, Bowen Ma, Feng Qiu, and Yu Ding · 2023
Later among the works it cites.
Multimodal continuous emotion recognition: A technical report for abaw5
Su Zhang, Ziyuan Zhao, and Cuntai Guan · 2023
Later among the works it cites.
Abaw: learning from synthetic data & multi-task learning challenges
Dimitrios Kollias · 2023
Later among the works it cites.
Abaw: Valence-arousal estimation, expression recognition, action unit detection & emotional reaction intensity estimation challenges
Dimitrios Kollias, Panagiotis Tzirakis, Alice Baird, Alan Cowen, and Stefanos Zafeiriou · 2023
Later among the works it cites.
Audio-visual person verification based on recursive fusion of joint cross-attention
R Gnana Praveen and Jahangir Alam · 2024
Closest in time.
Challenges and opportunities of text-based emotion detection: A survey
Abdullah Al Maruf, Fahima Khanam, Md. Mahmudul Haque, Zakaria Masud Jiyad, M. F. Mridha, and Zeyar Aung · 2024
Closest in time.
The 6th affective behavior analysis in-the-wild (abaw) competition
Dimitrios Kollias, Panagiotis Tzirakis, Alan Cowen, Stefanos Zafeiriou, Chunchang Shao, and Guanyu Hu · 2024
Closest in time.
Affective behaviour analysis via integrating multi-modal knowledge
Wei Zhang, Feng Qiu, Chen Liu, Lincheng Li, Heming Du, Tiancheng Guo, and Xin Yu · 2024
Closest in time.
Weiwei Zhou, Jiada Lu, Chenkun Ling, Weifeng Wang, and Shaowei Liu · 2024
Closest in time.
Denis Dresvyanskiy, Maxim Markitantov, Jiawei Yu, Peitong Li, Heysem Kaya, and Alexey Karpov · 2024
Closest in time.
Jun Yu, Gongpeng Zhao, Yongqi Wan, Zhihong Wei, Yang Zheng, Zerui Zhang, Zhongpeng Cai, Guochen Xie, Jichao Zhu, and Wangyuan Zhu · 2024
Closest in time.
Andrey V Savchenko · 2024
Closest in time.
Joint multimodal transformer for dimensional emotional recognition in the wild
Paul Waligora, Osama Zeeshan, Haseeb Aslam, Soufiane Belharbi, Alessandro Lameiras Koerich, Marco Pedersoli, Simon Bacon, and Eric Granger · 2024
Closest in time.
Emotion recognition using transformers with masked learning
Seongjae Min, Junseok Yang, Sangjun Lim, Junyong Lee, Sangwon Lee, and Sejoon Lim · 2024
Closest in time.