Fetching the paper…
Reading the bibliography…
Multimodal emotion recognition (MER) in practical scenarios is significantly challenged by the presence of missing or incomplete data across different modalities.
Creating artificial neural networks that generalize
Jocelyn Sietsma and Robert JF Dow. 1991 · 1991
Earlier work this paper cites.
Adaptive robust impulse noise filtering
Seong Rag Kim and Adam Efron. 1995 · 1995
Earlier work this paper cites.
Long Short-Term Memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Pattern recognition and machine learning . Vol. 4
Christopher M Bishop and Nasser M Nasrabadi. 2006 · 2006
Earlier work this paper cites.
IEMOCAP: Interactive emotional dyadic motion capture database
Carlos Busso, Murtaza Bulut, Chi-Chun Lee, Abe Kazemzadeh, Emily Mower, Samuel Kim, Jeannette N Chang, Sungbok Lee, and Shrikanth S Narayanan. 2008 · 2008
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database. In 2009 IEEE conference on computer vision and pattern recognition . Ieee, 248–255
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei. 2009 · 2009
Earlier work this paper cites.
The elements of statistical learning: data mining, inference, and prediction . Vol. 2
Trevor Hastie, Robert Tibshirani, Jerome H Friedman, and Jerome H Friedman. 2009 · 2009
Earlier work this paper cites.
Opensmile: the munich versatile and fast open-source audio feature extractor. In Proceedings of the 18th ACM international conference on Multimedia . 1459–1462
Florian Eyben, Martin Wöllmer, and Björn Schuller. 2010 · 2010
Earlier work this paper cites.
COVAREP—A collaborative voice analysis repository for speech technologies. In 2014 ieee international conference on acoustics, speech and signal processing (icassp) . IEEE, 960–964
Gilles Degottex, John Kane, Thomas Drugman, Tuomo Raitio, and Stefan Scherer. 2014 · 2014
Earlier work this paper cites.
Convolutional Neural Networks for Sentence Classification. In EMNLP . ACL, 1746–1751
Yoon Kim. 2014 · 2014
Earlier work this paper cites.
Auto-Encoding Variational Bayes. In ICLR
Diederik P. Kingma and Max Welling. 2014 · 2014
Earlier work this paper cites.
Glove: Global vectors for word representation. In Proceedings of the 2014 conference on empirical methods in natural language processing (EMNLP) . 1532–1543
Jeffrey Pennington, Richard Socher, and Christopher D Manning. 2014 · 2014
Earlier work this paper cites.
Learning structured output representation using deep conditional generative models
Kihyuk Sohn, Honglak Lee, and Xinchen Yan. 2015 · 2015
Earlier work this paper cites.
Deep residual learning for image recognition. In Proceedings of the IEEE conference on computer vision and pattern recognition . 770–778
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Earlier work this paper cites.
Densely Connected Convolutional Networks. In 2017 IEEE Conference on Computer Vision and Pattern Recognition, CVPR 2017, Honolulu, HI, USA, July 21-26, 2017 . 2261–2269
Gao Huang, Zhuang Liu, Laurens van der Maaten, and Kilian Q. Weinberger. 2017 · 2017
Cited alongside, same era.
Adversarial Training Methods for Semi-Supervised Text Classification. In International Conference on Learning Representations
Takeru Miyato, Andrew M. Dai, and Ian Goodfellow. 2017 · 2017
Cited alongside, same era.
Deep adversarial learning for multi-modality missing data completion. In Proceedings of the 24th ACM SIGKDD international conference on knowledge discovery & data mining . 1158–1166
Lei Cai, Zhengyang Wang, Hongyang Gao, Dinggang Shen, and Shuiwang Ji. 2018 · 2018
Cited alongside, same era.
Semi-supervised deep generative modelling of incomplete multi-modality emotional data. In Proceedings of the 26th ACM international conference on Multimedia . 108–116
Changde Du, Changying Du, Hao Wang, Jinpeng Li, Wei-Long Zheng, Bao-Liang Lu, and Huiguang He. 2018 · 2018
Cited alongside, same era.
Hgmf: heterogeneous graph-based fusion for multimodal data with incompleteness. In Proceedings of the 26th ACM SIGKDD international conference on knowledge discovery & data mining . 1295–1305
Jiayi Chen and Aidong Zhang. 2020 · 2020
Later among the works it cites.
Pitch-synchronous single frequency filtering spectrogram for speech emotion recognition
Shruti Gupta, Md. Shah Fahad, and Akshay Deepak. 2020 · 2020
Later among the works it cites.
Misa: Modality-invariant and-specific representations for multimodal sentiment analysis. In Proceedings of the 28th ACM international conference on multimedia . 1122–1131
Devamanyu Hazarika, Roger Zimmermann, and Soujanya Poria. 2020 · 2020
Later among the works it cites.
Denoising diffusion probabilistic models
Jonathan Ho, Ajay Jain, and Pieter Abbeel. 2020 · 2020
Later among the works it cites.
Deep transformer models for time series forecasting: The influenza prevalence case
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Multi-attention recurrent network for human communication comprehension. In Thirty-Second AAAI Conference on Artificial Intelligence
Amir Zadeh, Paul Pu Liang, Soujanya Poria, Prateek Vij, Erik Cambria, and Louis-Philippe Morency. 2018 · 2018
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In NAACL-HLT (1) . Association for Computational Linguistics, 4171–4186
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Implicit fusion by joint audiovisual training for emotion recognition in mono modality. In ICASSP 2019-2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 5861–5865
Jing Han, Zixing Zhang, Zhao Ren, and Björn Schuller. 2019 · 2019
Cited alongside, same era.
Impulsive noise recovery and elimination: A sparse machine learning based approach
Sicong Liu, Liang Xiao, Lianfen Huang, and Xianbin Wang. 2019 · 2019
Cited alongside, same era.
Decoupled Weight Decay Regularization. In International Conference on Learning Representations
Ilya Loshchilov and Frank Hutter. 2019 · 2019
Cited alongside, same era.
A comprehensive survey on impulse and Gaussian denoising filters for digital images
Mehdi Mafi, Harold Martin, Mercedes Cabrerizo, Jean Andrian, Armando Barreto, and Malek Adjouadi. 2019 · 2019
Cited alongside, same era.
Found in translation: Learning robust joint representations by cyclic translations between modalities. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 33. 6892–6899
Hai Pham, Paul Pu Liang, Thomas Manzini, Louis-Philippe Morency, and Barnabás Póczos. 2019 · 2019
Cited alongside, same era.
Metric Learning on Healthcare Data with Incomplete Modalities.. In IJCAI . 3534–3540
Qiuling Suo, Weida Zhong, Fenglong Ma, Ye Yuan, Jing Gao, and Aidong Zhang. 2019 · 2019
Cited alongside, same era.
Neo Wu, Bradley Green, Xue Ben, and Shawn O’Banion. 2020 · 2020
Later among the works it cites.
MultiBench: Multiscale Benchmarks for Multimodal Representation Learning. In Thirty-fifth Conference on Neural Information Processing Systems Datasets and Benchmarks Track (Round 1)
Paul Pu Liang, Yiwei Lyu, Xiang Fan, Zetian Wu, Yun Cheng, Jason Wu, Leslie Yufan Chen, Peter Wu, Michelle A Lee, Yuke Zhu, et al · 2021
Later among the works it cites.
Missing modality imagination network for emotion recognition with uncertain missing modalities. In Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers) . 2608–2618
Jinming Zhao, Ruichen Li, and Qin Jin. 2021 · 2021
Later among the works it cites.
Seen and unseen emotional style transfer for voice conversion with a new emotional speech dataset. In ICASSP 2021-2021 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 920–924
Kun Zhou, Berrak Sisman, Rui Liu, and Haizhou Li. 2021 · 2021
Later among the works it cites.
Accurate Emotion Strength Assessment for Seen and Unseen Speech Based on Data-Driven Deep Learning. In Interspeech 2022, 23rd Annual Conference of the International Speech Communication Association, Incheon, Korea, 18-22 September 2022 . ISCA, 5493–5497
Rui Liu, Berrak Sisman, Björn W. Schuller, Guanglai Gao, and Haizhou Li. 2022 · 2022
Later among the works it cites.
Mitigating Inconsistencies in Multimodal Sentiment Analysis under Uncertain Missing Modalities. In Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing . 2924–2934
Jiandian Zeng, Jiantao Zhou, and Tianyi Liu. 2022 · 2022
Later among the works it cites.
GCNet: graph completion network for incomplete multimodal learning in conversation
Zheng Lian, Lan Chen, Licai Sun, Bin Liu, and Jianhua Tao. 2023 · 2023
Closest in time.
Exploiting modality-invariant feature for robust multimodal emotion recognition with missing modalities. In ICASSP 2023-2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 1–5
Haolin Zuo, Rui Liu, Jinming Zhao, Guanglai Gao, and Haizhou Li. 2023 · 2023
Closest in time.
Contrastive Learning based Modality-Invariant Feature Acquisition for Robust Multimodal Emotion Recognition with Missing Modalities
Rui Liu, Haolin Zuo, Zheng Lian, Bjorn W Schuller, and Haizhou Li. 2024 · 2024
Closest in time.