Fetching the paper…
Reading the bibliography…
Multimodal emotion recognition (MMER) systems typically outperform unimodal systems by leveraging the inter- and intra-modal relationships between, e.g., visual, textual, physiological, and auditory modalities.
Face behavior a la carte: Expressions, affect and action units in a single network
Kollias, D., Sharmanska, V., and Zafeiriou, S. (2019a) · 1910
Earlier work this paper cites.
Expression, affect, action unit recognition: Aff-wild2, multi-task learning and arcface
Kollias, D. and Zafeiriou, S. (2019) · 1910
Earlier work this paper cites.
Aff-wild: Valence and arousal ‘in-the-wild’challenge
Zafeiriou, S., Kollias, D., Nicolaou, M. A., Papaioannou, A., Zhao, G., and Kotsia, I. (2017) · 1987
Earlier work this paper cites.
A concordance correlation coefficient to evaluate reproducibility
Lawrence, I. and Lin, K. (1989) · 1989
Earlier work this paper cites.
An argument for basic emotions
Ekman, P. (1992) · 1992
Earlier work this paper cites.
Features and classifiers for emotion recognition from speech: a survey from 2000 to 2011
Anagnostopoulos, C., Iliou, T., and Giannoukos, I. (2015) · 2011
Earlier work this paper cites.
Multimodal deep learning
Ngiam, J., Khosla, A., Kim, M., Nam, J., Lee, H., and Ng, A. Y. (2011) · 2011
Earlier work this paper cites.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
Cho, K., Van Merriënboer, B., Gulcehre, C., Bahdanau, D., Bougares, F., Schwenk, H., and Bengio, Y. (2014) · 2014
Earlier work this paper cites.
Automatic pain recognition from video and biomedical signals
Werner, P., Al-Hamadi, A., Niese, R., Walter, S., Gruss, S., and Traue, H. C. (2014a) · 2014
Earlier work this paper cites.
Automatic pain recognition from video and biomedical signals
Werner, P., Al-Hamadi, A., Niese, R., Walter, S., Gruss, S., and Traue, H. C. (2014b) · 2014
Earlier work this paper cites.
Bio-visual fusion for person-independent recognition of pain intensity
Kächele, M., Werner, P., Al-Hamadi, A., Palm, G., Walter, S., and Schwenker, F. (2015) · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S., and Sun, J. (2016) · 2016
Earlier work this paper cites.
Methods for person-centered continuous pain intensity assessment from bio-physiological channels
Kächele, M., Thiam, P., Amirian, M., Schwenker, F., and Palm, G. (2016) · 2016
Earlier work this paper cites.
Automatic pain assessment with facial activity descriptors
Werner, P., Al-Hamadi, A., Limbrecht-Ecklundt, K., Walter, S., Gruss, S., and Traue, H. C. (2016) · 2016
Earlier work this paper cites.
Multi-task neural networks for personalized pain recognition from physiological signals
Lopez-Martinez, D. and Picard, R. (2017) · 2017
Earlier work this paper cites.
End-to-end multimodal emotion recognition using deep neural networks
Tzirakis, P., Trigeorgis, G., Nicolaou, M. A., Schuller, B. W., and Zafeiriou, S. (2017) · 2017
Earlier work this paper cites.
Continuous pain intensity estimation from autonomic signals with recurrent neural networks
Lopez-Martinez, D. and Picard, R. (2018) · 2018
Earlier work this paper cites.
Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks
Lu, J., Batra, D., Parikh, D., and Lee, S. (2019) · 2019
Earlier work this paper cites.
Emotion recognition using fusion of audio and video features
Ortega, J. D. S., Cardinal, P., and Koerich, A. L. (2019) · 2019
Earlier work this paper cites.
Exploring deep physiological models for nociceptive pain recognition
Thiam, P., Bellmann, P., Kestler, H. A., and Schwenker, F. (2019) · 2019
Cited alongside, same era.
Automatic subject independent pain intensity estimation using a deep learning approach
Dragomir, M.-C., Florea, C., and Pupezescu, V. (2020) · 2020
Cited alongside, same era.
Multimodal transformer fusion for continuous emotion recognition
Huang, J., Tao, J., Liu, B., Lian, Z., and Niu, M. (2020) · 2020
Cited alongside, same era.
Analysing affective behavior in the first abaw competition
Kollias, D., Schulc, A., Hajiyev, E., and Zafeiriou, S. (2020) · 2020
Cited alongside, same era.
Two-stream aural-visual affect analysis in the wild
Kuhnke, F., Rumberg, L., and Ostermann, J. (2020b) · 2020
Cited alongside, same era.
Training strategies to handle missing modalities for audio-visual expression recognition
End-to-end multimodal affect recognition in real-world environments
Tzirakis, P., Chen, J., Zafeiriou, S., and Schuller, B. (2021) · 2021
Later among the works it cites.
Continuous emotion recognition with audio-visual leader-follower attentive fusion
Zhang, S., Ding, Y., Wei, Z., and Guan, C. (2021) · 2021
Later among the works it cites.
Multimodal-based stream integrated neural networks for pain assessment
Zhi, R., Zhou, C., Yu, J., Li, T., and Zamzmi, G. (2021) · 2021
Later among the works it cites.
Continuous-time audiovisual fusion with recurrence vs. attention for in-the-wild affect recognition
Karas, V., Tellamekala, M. K., Mallol-Ragolta, A., Valstar, M., and Schuller, B. W. (2022) · 2022
Later among the works it cites.
Abaw: Valence-arousal estimation, expression recognition, action unit detection & multi-task learning challenges
Kollias, D. (2022) · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Parthasarathy, S. and Sundaram, S. (2020) · 2020
Cited alongside, same era.
Hybrid rnn-ann based deep physiological network for pain recognition
Wang, R., Xu, K., Feng, H., and Chen, W. (2020) · 2020
Cited alongside, same era.
Multi-modality cross attention network for image and sentence matching
Wei, X., Zhang, T., Li, Y., Zhang, Y., and Wu, F. (2020) · 2020
Cited alongside, same era.
Multi-modal continuous valence-arousal estimation in the wild
Zhang, Y.-H., Huang, R., Zeng, J., and Shan, S. (2020) · 2020
Cited alongside, same era.
Mdn: A deep maximization-differentiation network for spatio-temporal depression detection
de Melo, W. C., Granger, E., and Lopez, M. B. (2021) · 2021
Cited alongside, same era.
Iterative distillation for better uncertainty estimates in multitask emotion recognition
Deng, D., Wu, L., and Shi, B. E. (2021) · 2021
Cited alongside, same era.
Audio-visual event localization via recursive fusion by joint co-attention
Duan, B., Tang, H., Wang, W., Zong, Z., Yang, G., and Yan, Y. (2021) · 2021
Cited alongside, same era.
Meng, L., Liu, Y., Liu, X., Huang, Z., Jiang, W., Zhang, T., Deng, Y., Li, R., Wu, Y., Zhao, J., et al. (2022) · 2022
Later among the works it cites.
Pain detection from facial expressions based on transformers and distillation
Morabit, S. E. and Rivenq, A. (2022) · 2022
Later among the works it cites.
An ensemble approach for facial expression analysis in video
Nguyen, H.-H., Huynh, V.-T., and Kim, S.-H. (2022) · 2022
Later among the works it cites.
Frame-level prediction of facial expressions, valence, arousal and action units for mobile devices
Savchenko, A. V. (2022) · 2022
Later among the works it cites.
A pre-trained audio-visual transformer for emotion recognition
Tran, M. and Soleymani, M. (2022) · 2022
Later among the works it cites.
Continuous emotion recognition using visual-audio-linguistic information: A technical report for abaw3
Zhang, S., An, R., Ding, Y., and Guan, C. (2022) · 2022
Later among the works it cites.
Privileged knowledge distillation for dimensional emotion recognition in the wild
Aslam, M. H., Zeeshan, O., Pedersoli, M., Koerich, A. L., Bacon, S., and Granger, E. (2023) · 2023
Later among the works it cites.
Abaw: Valence-arousal estimation, expression recognition, action unit detection & emotional reaction intensity estimation challenges
Kollias, D., Tzirakis, P., Baird, A., Cowen, A., and Zafeiriou, S. (2023) · 2023
Later among the works it cites.
Multi-label multimodal emotion recognition with transformer-based fusion and emotion-level representation learning
Le, H.-D., Lee, G.-S., Kim, S.-H., Kim, S., and Yang, H.-J. (2023) · 2023
Later among the works it cites.
Lu, Z., Ozek, B., and Kamarthi, S. (2023) · 2023
Later among the works it cites.
Pain recognition with physiological signals using multi-level context information
Phan, K. N., Iyortsuun, N. K., Pant, S., Yang, H.-J., and Kim, S.-H. (2023) · 2023
Later among the works it cites.
Leveraging tcn and transformer for effective visual-audio fusion in continuous emotion recognition
Zhou, W., Lu, J., Xiong, Z., and Wang, W. (2023) · 2023
Later among the works it cites.
Distilling privileged multimodal information for expression recognition using optimal transport
Aslam, M. H., Zeeshan, M. O., Belharbi, S., Pedersoli, M., Koerich, A. L., Bacon, S., and Granger, E. (2024) · 2024
Closest in time.
The 6th affective behavior analysis in-the-wild (abaw) competition
Kollias, D., Tzirakis, P., Cowen, A., Zafeiriou, S., Shao, C., and Hu, G. (2024) · 2024
Closest in time.