Fetching the paper…
Reading the bibliography…
Multimodal sentiment analysis has been studied under the assumption that all modalities are available.
IEMOCAP: Interactive emotional dyadic motion capture database
Carlos Busso, Murtaza Bulut, Chi-Chun Lee, Abe Kazemzadeh, Emily Mower, Samuel Kim, Jeannette N Chang, Sungbok Lee, and Shrikanth S Narayanan. 2008 · 2008
Earlier work this paper cites.
Visualizing data using t-SNE
Laurens van der Maaten and Geoffrey Hinton. 2008 · 2008
Earlier work this paper cites.
Extracting and composing robust features with denoising autoencoders. In Proceedings of the 25th International Conference on Machine Learning . 1096–1103
Pascal Vincent, Hugo Larochelle, Yoshua Bengio, and Pierre-Antoine Manzagol. 2008 · 2008
Earlier work this paper cites.
Spectral regularization algorithms for learning large incomplete matrices
Rahul Mazumder, Trevor Hastie, and Robert Tibshirani. 2010 · 2010
Earlier work this paper cites.
Autoencoders, unsupervised learning, and deep architectures. In Proceedings of ICML Workshop on Unsupervised and Transfer Learning . 37–49
Pierre Baldi. 2012 · 2012
Earlier work this paper cites.
Clustering on multiple incomplete datasets via collective kernel learning. In 2013 IEEE 13th International Conference on Data Mining . IEEE, 1181–1186
Weixiang Shao, Xiaoxiao Shi, and S Yu Philip. 2013 · 2013
Earlier work this paper cites.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
Auto-encoding variational bayes. In Proceedings of the International Conference on Learning Representations . 1–14
Diederik P Kingma and Max Welling. 2014 · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization. In Proceedings of the 3rd International Conference on Learning Representations . 1–15
Diederik P Kingma and Jimmy Ba. 2015 · 2015
Earlier work this paper cites.
librosa: Audio and music signal analysis in python. In Proceedings of the 14th Python in Science Conference , Vol. 8. 18–25
Brian McFee, Colin Raffel, Dawen Liang, Daniel PW Ellis, Matt McVicar, Eric Battenberg, and Oriol Nieto. 2015 · 2015
Earlier work this paper cites.
Deep residual learning for image recognition. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 770–778
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Earlier work this paper cites.
Multimodal sentiment intensity analysis in videos: Facial gestures and verbal messages
Amir Zadeh, Rowan Zellers, Eli Pincus, and Louis-Philippe Morency. 2016 · 2016
Earlier work this paper cites.
Multimodal sentiment analysis with word-level fusion and reinforcement learning. In Proceedings of the 2017 ACM International Conference on Multimodal Interaction . 163–171
Minghai Chen, Sen Wang, Paul Pu Liang, Tadas Baltrušaitis, Amir Zadeh, and Louis-Philippe Morency. 2017 · 2017
Earlier work this paper cites.
Online early-late fusion based on adaptive HMM for sign language recognition
Dan Guo, Wengang Zhou, Houqiang Li, and Meng Wang. 2017 · 2017
Earlier work this paper cites.
VIGAN: Missing view imputation with generative adversarial networks. In 2017 IEEE International Conference on Big Data . IEEE, 766–775
Chao Shang, Aaron Palmer, Jiangwen Sun, Ko-Shin Chen, Jin Lu, and Jinbo Bi. 2017 · 2017
Earlier work this paper cites.
Missing modalities imputation via cascaded residual autoencoder. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 1405–1414
Luan Tran, Xiaoming Liu, Jiayu Zhou, and Rong Jin. 2017 · 2017
Earlier work this paper cites.
Attention is all you need. In Advances in Neural Information Processing Systems . 5998–6008
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Openface 2.0: Facial behavior analysis toolkit. In 2018 13th IEEE International Conference on Automatic Face & Gesture Recognition . IEEE, 59–66
Tadas Baltrusaitis, Amir Zadeh, Yao Chong Lim, and Louis-Philippe Morency. 2018 · 2018
Cited alongside, same era.
Deep adversarial learning for multi-modality missing data completion. In Proceedings of the 24th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining . 1158–1166
Lei Cai, Zhengyang Wang, Hongyang Gao, Dinggang Shen, and Shuiwang Ji. 2018 · 2018
Cited alongside, same era.
Semi-supervised deep generative modelling of incomplete multi-modality emotional data. In Proceedings of the 26th ACM International Conference on Multimedia . 108–116
Changde Du, Changying Du, Hao Wang, Jinpeng Li, Wei-Long Zheng, Bao-Liang Lu, and Huiguang He. 2018 · 2018
Cited alongside, same era.
Multimodal sentiment analysis using hierarchical fusion with context modeling
Navonil Majumder, Devamanyu Hazarika, Alexander Gelbukh, Erik Cambria, and Soujanya Poria. 2018 · 2018
Cited alongside, same era.
Web table retrieval using multimodal deep learning. In Proceedings of the 43rd International ACM SIGIR Conference on Research and Development in Information Retrieval . 1399–1408
Roee Shraga, Haggai Roitman, Guy Feigenblat, and Mustafa Cannim. 2020 · 2020
Later among the works it cites.
Transmodality: An end2end fusion method with transformer for multimodal sentiment analysis. In Proceedings of The Web Conference 2020 . 2514–2520
Zilong Wang, Zhaohong Wan, and Xiaojun Wan. 2020 · 2020
Later among the works it cites.
Social image sentiment analysis by exploiting multimodal content and heterogeneous relations
Jie Xu, Zhoujun Li, Feiran Huang, Chaozhuo Li, and S Yu Philip. 2020a · 2020
Later among the works it cites.
Ch-sims: A chinese multimodal sentiment analysis dataset with fine-grained annotation of modality. In Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics . 3718–3727
Wenmeng Yu, Hua Xu, Fanyang Meng, Yilin Zhu, Yixiao Ma, Jiele Wu, Jiyun Zou, and Kaicheng Yang. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Multimodal sentiment analysis: Addressing key issues and setting up the baselines
Soujanya Poria, Navonil Majumder, Devamanyu Hazarika, Erik Cambria, Alexander Gelbukh, and Amir Hussain. 2018 · 2018
Cited alongside, same era.
A co-memory network for multimodal sentiment analysis. In Proceedings of the 2018 International ACM SIGIR Conference on Research and Development in Information Retrieval . 929–932
Nan Xu, Wenji Mao, and Guandan Chen. 2018 · 2018
Cited alongside, same era.
Multimodal language analysis in the wild: Cmu-mosei dataset and interpretable dynamic fusion graph. In Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics . 2236–2246
AmirAli Bagher Zadeh, Paul Pu Liang, Soujanya Poria, Erik Cambria, and Louis-Philippe Morency. 2018 · 2018
Cited alongside, same era.
Transformer-xl: Attentive language models beyond a fixed-length context. In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics . 2978–2988
Zihang Dai, Zhilin Yang, Yiming Yang, Jaime Carbonell, Quoc V Le, and Ruslan Salakhutdinov. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . 4171–4186
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Found in translation: Learning robust joint representations by cyclic translations between modalities. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 33. 6892–6899
Hai Pham, Paul Pu Liang, Thomas Manzini, Louis-Philippe Morency, and Barnabás Póczos. 2019 · 2019
Cited alongside, same era.
Investigating dynamic routing in tree-structured LSTM for sentiment analysis. In Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing . 3423–3428
Jin Wang, Liang-Chih Yu, K Robert Lai, and Xuejie Zhang. 2019 · 2019
Cited alongside, same era.
Entity-sensitive attention and fusion network for entity-level multimodal sentiment classification
Jianfei Yu, Jing Jiang, and Rui Xia. 2019 · 2019
Cited alongside, same era.
Knowledge guided capsule attention network for aspect-based sentiment analysis
Bowen Zhang, Xutao Li, Xiaofei Xu, Ka-Cheong Leung, Zhiyao Chen, and Yunming Ye. 2020b · 2020
Later among the works it cites.
Deep partial multi-view learning
Changqing Zhang, Yajie Cui, Zongbo Han, Joey Tianyi Zhou, Huazhu Fu, and Qinghua Hu. 2020a · 2020
Later among the works it cites.
Joint aspect-sentiment analysis with minimal user guidance. In Proceedings of the 43rd International ACM SIGIR Conference on Research and Development in Information Retrieval . 1241–1250
Honglei Zhuang, Fang Guo, Chao Zhang, Liyuan Liu, and Jiawei Han. 2020 · 2020
Later among the works it cites.
Vatt: Transformers for multimodal self-supervised learning from raw video, audio and text. In Proceedings of the 35th Annual Conference on Neural Information Processing Systems
Hassan Akbari, Linagzhe Yuan, Rui Qian, Wei-Hong Chuang, Shih-Fu Chang, Yin Cui, and Boqing Gong. 2021 · 2021
Later among the works it cites.
Twins: Revisiting the Design of Spatial Attention in Vision Transformers. In Advances in Neural Information Processing Systems , Vol. 34. Curran Associates, Inc., 9355–9366
Xiangxiang Chu, Zhi Tian, Yuqing Wang, Bo Zhang, Haibing Ren, Xiaolin Wei, Huaxia Xia, and Chunhua Shen. 2021 · 2021
Later among the works it cites.
Iterative network pruning with uncertainty regularization for lifelong sentiment classification. In Proceedings of the 44th International ACM SIGIR conference on Research and Development in Information Retrieval . 1229–1238
Binzong Geng, Min Yang, Fajie Yuan, Shupeng Wang, Xiang Ao, and Ruifeng Xu. 2021 · 2021
Later among the works it cites.
Analyzing multimodal sentiment via acoustic-and visual-LSTM with channel-aware temporal convolution network
Sijie Mai, Songlong Xing, and Haifeng Hu. 2021 · 2021
Later among the works it cites.
QuTI! Quantifying Text-Image Consistency in Multimodal Documents. In Proceedings of the 44th International ACM SIGIR Conference on Research and Development in Information Retrieval . 2575–2579
Matthias Springstein, Eric Müller-Budack, and Ralph Ewerth. 2021 · 2021
Later among the works it cites.
Sentiment analysis and topic recognition in video transcriptions
Lukas Stappen, Alice Baird, Erik Cambria, and Björn W Schuller. 2021 · 2021
Later among the works it cites.
Learning modality-specific representations with self-supervised multi-task learning for multimodal sentiment analysis. In Proceedings of the AAAI Conference on Artificial Intelligence . 10790–10797
Wenmeng Yu, Hua Xu, Ziqi Yuan, and Jiele Wu. 2021 · 2021
Later among the works it cites.
Transformer-based feature reconstruction network for robust multimodal sentiment analysis. In Proceedings of the 29th ACM International Conference on Multimedia . 4400–4407
Ziqi Yuan, Wei Li, Hua Xu, and Wenmeng Yu. 2021 · 2021
Later among the works it cites.
Missing Modality Imagination Network for Emotion Recognition with Uncertain Missing Modalities. In Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing . 2608–2618
Jinming Zhao, Ruichen Li, and Qin Jin. 2021 · 2021
Later among the works it cites.