Fetching the paper…
Reading the bibliography…
Multimodal models trained on complete modality data often exhibit a substantial decrease in performance when faced with imperfect data containing corruptions or missing modalities.
Pattern recognition and machine learning
Christopher M Bishop and Nasser M Nasrabadi · 2006
Earlier work this paper cites.
Iemocap: Interactive emotional dyadic motion capture database
Carlos Busso, Murtaza Bulut, Chi-Chun Lee, Abe Kazemzadeh, Emily Mower, Samuel Kim, Jeannette N Chang, Sungbok Lee, and Shrikanth S Narayanan · 2008
Earlier work this paper cites.
Opensmile: the munich versatile and fast open-source audio feature extractor
Florian Eyben, Martin Wöllmer, and Björn Schuller · 2010
Earlier work this paper cites.
Auto-encoding variational bayes
Diederik P Kingma and Max Welling · 2013
Earlier work this paper cites.
Food-101–mining discriminative components with random forests
Lukas Bossard, Matthieu Guillaumin, and Luc Van Gool · 2014
Earlier work this paper cites.
Covarep—a collaborative voice analysis repository for speech technologies
Gilles Degottex, John Kane, Thomas Drugman, Tuomo Raitio, and Stefan Scherer · 2014
Earlier work this paper cites.
Convolutional neural networks for sentence classification
Yoon Kim · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Haşim Sak, Andrew Senior, and Françoise Beaufays · 2014
Earlier work this paper cites.
Openface: an open source facial behavior analysis toolkit
Tadas Baltrušaitis, Peter Robinson, and Louis-Philippe Morency · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Learning with side information through modality hallucination
Judy Hoffman, Saurabh Gupta, and Trevor Darrell · 2016
Earlier work this paper cites.
Densely connected convolutional networks
Gao Huang, Zhuang Liu, Laurens Van Der Maaten, and Kilian Q Weinberger · 2017
Earlier work this paper cites.
Missing modalities imputation via cascaded residual autoencoder
Luan Tran, Xiaoming Liu, Jiayu Zhou, and Rong Jin · 2017
Earlier work this paper cites.
Multimodal generative models for scalable weakly-supervised learning
Mike Wu and Noah Goodman · 2018
Earlier work this paper cites.
Multimodal language analysis in the wild: Cmu-mosei dataset and interpretable dynamic fusion graph
AmirAli Bagher Zadeh, Paul Pu Liang, Soujanya Poria, Erik Cambria, and Louis-Philippe Morency · 2018
Earlier work this paper cites.
Multimodal fusion of brain networks with longitudinal couplings
Wen Zhang, Kai Shu, Suhang Wang, Huan Liu, and Yalin Wang · 2018
Earlier work this paper cites.
Robust multimodal brain tumor segmentation via feature disentanglement and gated fusion
Cheng Chen, Qi Dou, Yueming Jin, Hao Chen, Jing Qin, and Pheng-Ann Heng · 2019
Cited alongside, same era.
BERT: pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Cited alongside, same era.
Robust deep multi-modal learning based on gated information fusion network
Jaekyum Kim, Junho Koh, Yecheol Kim, Jaehyung Choi, Youngbae Hwang, and Jun Won Choi · 2019
Cited alongside, same era.
Benchmarking robustness in object detection: Autonomous driving when winter is coming
Claudio Michaelis, Benjamin Mitzkus, Robert Geirhos, Evgenia Rusak, Oliver Bringmann, Alexander S. Ecker, Matthias Bethge, and Wieland Brendel · 2019
Cited alongside, same era.
Variational mixture-of-experts autoencoders for multi-modal deep generative models
Yuge Shi, Brooks Paige, Philip Torr, et al · 2019
Cited alongside, same era.
N15news: A new dataset for multimodal news classification
Zhen Wang, Xu Shan, and Jie Yang · 2021
Later among the works it cites.
Learning modality-specific representations with self-supervised multi-task learning for multimodal sentiment analysis
Wenmeng Yu, Hua Xu, Ziqi Yuan, and Jiele Wu · 2021
Later among the works it cites.
Modality-aware mutual learning for multi-modal medical image segmentation
Yao Zhang, Jiawei Yang, Jiang Tian, Zhongchao Shi, Cheng Zhong, Yang Zhang, and Zhiqiang He · 2021
Later among the works it cites.
Missing modality imagination network for emotion recognition with uncertain missing modalities
Jinming Zhao, Ruichen Li, and Qin Jin · 2021
Later among the works it cites.
Multimodal dynamics: Dynamical fusion for trustworthy multimodal classification
Zongbo Han, Fan Yang, Junzhou Huang, Changqing Zhang, and Jianhua Yao · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Uncertainty-aware audiovisual activity recognition using deep bayesian variational inference
Mahesh Subedar, Ranganath Krishnan, Paulo Lopez Meyer, Omesh Tickoo, and Jonathan Huang · 2019
Cited alongside, same era.
On robustness of multi-modal fusion—robotics perspective
Michal Bednarek, Piotr Kicki, and Krzysztof Walas · 2020
Cited alongside, same era.
Knowledge distillation from multi-modal to mono-modal segmentation networks
Minhao Hu, Matthis Maillard, Ya Zhang, Tommaso Ciceri, Giammarco La Barbera, Isabelle Bloch, and Pietro Gori · 2020
Cited alongside, same era.
Semi-supervised multi-modal emotion recognition with cross-modal distribution matching
Jingjun Liang, Ruichen Li, and Qin Jin · 2020
Cited alongside, same era.
What makes training multi-modal classification networks hard?
Weiyao Wang, Du Tran, and Matt Feiszli · 2020
Cited alongside, same era.
Multimodal video sentiment analysis using deep learning approaches, a survey
Sarah A Abdu, Ahmed H Yousef, and Ashraf Salem · 2021
Cited alongside, same era.
Improving multimodal fusion with hierarchical mutual information maximization for multimodal sentiment analysis
Wei Han, Hui Chen, and Soujanya Poria · 2021
Cited alongside, same era.
Are multimodal transformers robust to missing modality?
Mengmeng Ma, Jian Ren, Long Zhao, Davide Testuggine, and Xi Peng · 2022
Later among the works it cites.
Geometric multimodal contrastive representation learning
Petra Poklukar, Miguel Vasco, Hang Yin, Francisco S Melo, Ana Paiva, and Danica Kragic · 2022
Later among the works it cites.
Tag-assisted multimodal sentiment analysis under uncertain missing modalities
Jiandian Zeng, Tianyi Liu, and Jiantao Zhou · 2022
Later among the works it cites.
Deep partial multi-view learning
Changqing Zhang, Yajie Cui, Zongbo Han, Joey Tianyi Zhou, Huazhu Fu, and Qinghua Hu · 2022
Later among the works it cites.
mmformer: Multimodal medical transformer for incomplete multimodal learning of brain tumor segmentation
Yao Zhang, Nanjun He, Jiawei Yang, Yuexiang Li, Dong Wei, Yawen Huang, Yang Zhang, Zhiqiang He, and Yefeng Zheng · 2022
Later among the works it cites.
Multimodal robustness for neural machine translation
Yuting Zhao and Ioan Calapodescu · 2022
Later among the works it cites.
Enhanced multimodal representation learning with cross-modal kd
Mengxi Chen, Linyu Xing, Yu Wang, and Ya Zhang · 2023
Closest in time.
Trusted multi-view classification with dynamic evidential fusion
Zongbo Han, Changqing Zhang, Huazhu Fu, and Joey Tianyi Zhou · 2023
Closest in time.
Mixgen: A new multi-modal data augmentation
Xiaoshuai Hao, Yi Zhu, Srikar Appalaraju, Aston Zhang, Wanqian Zhang, Bo Li, and Mu Li · 2023
Closest in time.
Decoupled multimodal distilling for emotion recognition
Yong Li, Yuanzhi Wang, and Zhen Cui · 2023
Closest in time.
Robust multimodal fusion for human activity recognition
Sanju Xaviar, Xin Yang, and Omid Ardakanian · 2023
Closest in time.