Fetching the paper…
Reading the bibliography…
Multimodal emotion recognition has recently gained much attention since it can leverage diverse and complementary relationships over multiple modalities (e.g., audio, visual, biosignals, etc.), and can provide some robustness to noisy modalities.
Three dimensions of emotion
H. Schlosberg · 1954
Earlier work this paper cites.
An argument for basic emotions
Paul Ekman · 1992
Earlier work this paper cites.
More evidence for the universality of a contempt expression
D. Matsumoto · 1992
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Xavier Glorot and Yoshua Bengio · 2010
Earlier work this paper cites.
Continuous prediction of spontaneous affect from multiple cues and modalities in valence-arousal space
Mihalis A. Nicolaou, Hatice Gunes, and Maja Pantic · 2011
Earlier work this paper cites.
The semaine database: Annotated multimodal records of emotionally colored conversations between a person and a limited agent
G McKeown, M Valstar, R Cowie, M Pantic, and M Schroder · 2012
Earlier work this paper cites.
Introducing the recola multimodal corpus of remote collaborative and affective interactions
F. Ringeval, A. Sonderegger, J. Sauer, and D. Lalanne · 2013
Earlier work this paper cites.
Lstm-modeling of continuous emotions in an a-v affect recognition framework
M Wöllmer, M Kaiser, F Eyben, B Schuller, and G Rigoll · 2013
Earlier work this paper cites.
Emotion Recognition and Its Applications
A. Kołakowska, A. Landowska, M. Szwoch, W. Szwoch, and M. R. Wróbel · 2014
Earlier work this paper cites.
Multimodal affective dimension prediction using deep bidirectional long short-term memory recurrent neural networks
Lang He, Dongmei Jiang, Le Yang, Ercheng Pei, Peng Wu, and Hichem Sahli · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
K He, X Zhang, S Ren, and J Sun · 2016
Earlier work this paper cites.
Quo vadis, action recognition? a new model and the kinetics dataset
J Carreira and A Zisserman · 2017
Earlier work this paper cites.
Nonverbal Communication
Albert Mehrabian · 2017
Earlier work this paper cites.
End-to-end multimodal er using deep neural networks
Panagiotis Tzirakis, George Trigeorgis, Mihalis A. Nicolaou, Björn W. Schuller, and Stefanos Zafeiriou · 2017
Earlier work this paper cites.
Aff-wild: Valence and arousal ‘in-the-wild’challenge
Stefanos Zafeiriou, Dimitrios Kollias, Mihalis A Nicolaou, Athanasios Papaioannou, Guoying Zhao, and Irene Kotsia · 2017
Earlier work this paper cites.
Audio-visual attention networks for emotion recognition
Jiyoung Lee, Sunok Kim, Seungryong Kim, and Kwanghoon Sohn · 2018
Earlier work this paper cites.
Emotion recognition from variable-length speech segments using deep learning on spectrograms
Xi Ma, Zhiyong Wu, Jia Jia, Mingxing Xu, Helen Meng, and Lianhong Cai · 2018
Earlier work this paper cites.
A closer look at spatiotemporal convolutions for action recognition
Du Tran, Heng Wang, Lorenzo Torresani, Jamie Ray, Yann LeCun, and Manohar Paluri · 2018
Cited alongside, same era.
Cross attention network for few-shot classification
Ruibing Hou, Hong Chang, Bingpeng MA, Shiguang Shan, and Xilin Chen · 2019
Cited alongside, same era.
Face behavior a la carte: Expressions, affect and action units in a single network
Dimitrios Kollias, Viktoriia Sharmanska, and Stefanos Zafeiriou · 2019
Cited alongside, same era.
Deep affect prediction in-the-wild: Aff-wild database and challenge, deep architectures, and beyond
D Kollias, P Tzirakis, M A Nicolaou, A Papaioannou, G Zhao, B Schuller, I Kotsia, and S Zafeiriou · 2019
Cited alongside, same era.
Expression, affect, action unit recognition: Aff-wild2, multi-task learning and arcface
Dimitrios Kollias and Stefanos Zafeiriou · 2019
Sewa db: A rich database for a-v emotion and sentiment research in the wild
J Kossaifi, R Walecki, Y Panagakis, J Shen, M Schmitt, F Ringeval, J Han, V Pandit, A Toisoul, B Schuller, K Star, E Hajiyev, and M Pantic · 2021
Later among the works it cites.
Cross-attentional audio-visual fusion for weakly-supervised action localization
Jun-Tae Lee, Mihir Jain, Hyoungwoo Park, and Sungrack Yun · 2021
Later among the works it cites.
Attention bottlenecks for multimodal fusion
A Nagrani, S Yang, A Arnab, C Schmid, and C Sun · 2021
Later among the works it cites.
Deep auto-encoders with sequential learning for multimodal dimensional emotion recognition
Dung Nguyen, Duc Thanh Nguyen, Rui Zeng, Thanh Thi Nguyen, Son Tran, Thin Khac Nguyen, S. Sridharan, and Clinton Fookes · 2021
Later among the works it cites.
Detecting expressions with multimodal transformers
Srinivas Parthasarathy and Shiva Sundaram · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Affectnet: A database for facial expression, valence, and arousal computing in the wild
Ali Mollahosseini, Behzad Hasani, and Mohammad H. Mahoor · 2019
Cited alongside, same era.
Emotion recognition using fusion of audio and video features
Juan D. S. Ortega, Patrick Cardinal, and Alessandro L. Koerich · 2019
Cited alongside, same era.
Deep weakly supervised domain adaptation for pain localization in videos
R Gnana Praveen, E. Granger, and P. Cardinal · 2020
Cited alongside, same era.
Analysing affective behavior in the first abaw 2020 competition
D Kollias, A Schulc, E Hajiyev, and S Zafeiriou · 2020
Cited alongside, same era.
Two-stream aural-visual affect analysis in the wild
F Kuhnke, L Rumberg, and J Ostermann · 2020
Cited alongside, same era.
What makes training multi-modal classification networks hard?
W Wang, D Tran, and M Feiszli · 2020
Cited alongside, same era.
Multi-modal continuous valence-arousal estimation in the wild
Yuan-Hang Zhang, Rulin Huang, Jiabei Zeng, and Shiguang Shan · 2020
Cited alongside, same era.
R. Gnana Praveen, Eric Granger, and Patrick Cardinal · 2021
Later among the works it cites.
Leveraging recent advances in deep learning for audio-visual emotion recognition
L Schoneveld, A Othmani, and H Abdelkawy · 2021
Later among the works it cites.
End-to-end multimodal affect recognition in real-world environments
P Tzirakis, J Chen, S Zafeiriou, and B Schuller · 2021
Later among the works it cites.
A multi-task mean teacher for semi-supervised facial affective behavior analysis
Lingfeng Wang, Shisen Wang, Jin Qi, and Kenji Suzuki · 2021
Later among the works it cites.
Continuous emotion recognition with audio-visual leader-follower attentive fusion
S Zhang, Y Ding, Z Wei, and C Guan · 2021
Later among the works it cites.
Continuous-time audiovisual fusion with recurrence vs. attention for in-the-wild affect recognition
Vincent Karas, Mani Kumar Tellamekala, Adria Mallol-Ragolta, Michel Valstar, and Björn W. Schuller · 2022
Closest in time.
Dimitrios Kollias · 2022
Closest in time.
Multi-modal emotion estimation for in-the-wild videos, 2022
Liyu Meng, Yuchen Liu, Xiaolong Liu, Zhaopei Huang, Yuan Cheng, Meng Wang, Chuanhe Liu, and Qin Jin · 2022
Closest in time.
An ensemble approach for facial expression analysis in video, 2022
Hong-Hai Nguyen, Van-Thong Huynh, and Soo-Hyung Kim · 2022
Closest in time.
Frame-level prediction of facial expressions, valence, arousal and action units for mobile devices, 2022
Andrey V. Savchenko · 2022
Closest in time.
Continuous emotion recognition using visual-audio-linguistic information: A technical report for abaw3, 2022
Su Zhang, Ruyi An, Yi Ding, and Cuntai Guan · 2022
Closest in time.