Fetching the paper…
Reading the bibliography…
The Multimodal Emotion Recognition challenge MER2024 focuses on recognizing emotions using audio, language, and visual signals.
Improved baselines with momentum contrastive learning
Chen, X.; Fan, H.; Girshick, R.; and He, K. 2020b · 2003
Earlier work this paper cites.
Openface: an open source facial behavior analysis toolkit
Baltrušaitis, T.; Robinson, P.; and Morency, L.-P. 2016 · 2016
Earlier work this paper cites.
Representation learning with contrastive predictive coding
Oord, A. v. d.; Li, Y.; and Vinyals, O. 2018 · 2018
Earlier work this paper cites.
Attending to Emotional Narratives
Wu, Z.; Zhang, X.; Zhi-Xuan, T.; Zaki, J.; and Ong, D. C. 2019 · 2019
Earlier work this paper cites.
Denoising diffusion probabilistic models
Ho, J.; Jain, A.; and Abbeel, P. 2020 · 2020
Earlier work this paper cites.
Multimodal sentimental analysis for social media applications: A comprehensive review
Chandrasekaran, G.; Nguyen, T. N.; and Hemanth D, J. 2021 · 2021
Earlier work this paper cites.
Hubert: Self-supervised speech representation learning by masked prediction of hidden units
Hsu, W.-N.; Bolte, B.; Tsai, Y.-H. H.; Lakhotia, K.; Salakhutdinov, R.; and Mohamed, A. 2021 · 2021
Earlier work this paper cites.
[Retracted] Multimodal Emotion Recognition Model Based on a Deep Neural Network with Multiobjective Optimization
Li, M.; Qiu, X.; Peng, S.; Tang, L.; Li, Q.; Yang, W.; and Ma, Y. 2021 · 2021
Earlier work this paper cites.
CTNet: Conversational transformer network for emotion recognition
Lian, Z.; Liu, B.; and Tao, J. 2021 · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Radford, A.; Kim, J. W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; et al. 2021 · 2021
Cited alongside, same era.
Wenet: Production oriented streaming and non-streaming end-to-end speech recognition toolkit
Yao, Z.; Wu, D.; Wang, X.; Zhang, B.; Yu, F.; Yang, C.; Peng, Z.; Chen, X.; Xie, L.; and Lei, X. 2021 · 2021
Cited alongside, same era.
Former-dfer: Dynamic facial expression recognition transformer
Zhao, Z.; and Liu, Q. 2021 · 2021
Cited alongside, same era.
Analyzing modality robustness in multimodal sentiment analysis
Hazarika, D.; Li, Y.; Cheng, B.; Zhao, S.; Zimmermann, R.; and Poria, S. 2022 · 2022
Cited alongside, same era.
Everything at once-multi-modal fusion transformer for video retrieval
Shvetsova, N.; Chen, B.; Rouditchenko, A.; Thomas, S.; Kingsbury, B.; Feris, R. S.; Harwath, D.; Glass, J.; and Kuehne, H. 2022 · 2022
Mer 2023: Multi-label learning, modality robustness, and semi-supervised learning
Lian, Z.; Sun, H.; Sun, L.; Chen, K.; Xu, M.; Wang, K.; Xu, K.; He, Y.; Li, Y.; Zhao, J.; et al. 2023 · 2023
Later among the works it cites.
Emotion recognition framework using multiple modalities for an effective human–computer interaction
Moin, A.; Aadil, F.; Ali, Z.; and Kang, D. 2023 · 2023
Later among the works it cites.
ZeroPrompt: Streaming Acoustic Encoders are Zero-Shot Masked LMs
Song, X.; Wu, D.; Zhang, B.; Peng, Z.; Dang, B.; Pan, F.; and Wu, Z. 2023 · 2023
Later among the works it cites.
Self-Training with Label-Feature-Consistency for Domain Adaptation
Xin, Y.; Luo, S.; Jin, P.; Du, Y.; and Wang, C. 2023 · 2023
Later among the works it cites.
Exploiting modality-invariant feature for robust multimodal emotion recognition with missing modalities
Zuo, H.; Liu, R.; Zhao, J.; Gao, G.; and Li, H. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
A measurement method for mental health based on dynamic multimodal feature recognition
Xu, H.; Wu, X.; and Liu, X. 2022 · 2022
Cited alongside, same era.
Baichuan 2: Open Large-scale Language Models
Baichuan. 2023 · 2023
Cited alongside, same era.
Fan, Q.; Zuo, H.; Liu, R.; Lian, Z.; and Gao, G. 2023 · 2023
Cited alongside, same era.
Mining High-quality Samples from Raw Data and Majority Voting Method for Multimodal Emotion Recognition
Li, Q.; Gao, Y.; and Li, Y. 2023 · 2023
Cited alongside, same era.
A simple framework for contrastive learning of visual representations
Chen, T.; Kornblith, S.; Norouzi, M.; and Hinton, G. 2020a
Cited in the paper.
Enhanced detection classification via clustering svm for various robot collaboration task
Liu, R.; Xu, X.; Shen, Y.; Zhu, A.; Yu, C.; Chen, T.; and Zhang, Y. 2024a
Cited in the paper.
Contrastive Learning based Modality-Invariant Feature Acquisition for Robust Multimodal Emotion Recognition with Missing Modalities
Liu, R.; Zuo, H.; Lian, Z.; Schuller, B. W.; and Li, H. 2024b
Cited in the paper.
MIE-Net: Motion Information Enhancement Network for Fine-Grained Action Recognition Using RGB Sensors
Li, Y.; Ma, M.; Wu, J.; Yang, K.; Pei, Z.; and Ren, J. 2024 · 2024
Closest in time.
Lian, Z.; Sun, H.; Sun, L.; Wen, Z.; Zhang, S.; Chen, S.; Gu, H.; Zhao, J.; et al. 2024 · 2024
Closest in time.
Harnessing XGBoost for Robust Biomarker Selection of Obsessive-Compulsive Disorder (OCD) from Adolescent Brain Cognitive Development (ABCD) data
Shen, X.; Zhang, Q.; Zheng, H.; and Qi, W. 2024 · 2024
Closest in time.
Credit card fraud detection using advanced transformer model
Yu, C.; Xu, Y.; Cao, J.; Zhang, Y.; Jin, Y.; and Zhu, M. 2024 · 2024
Closest in time.