Fetching the paper…
Reading the bibliography…
During multimodal model training and testing, certain data modalities may be absent due to sensor limitations, cost constraints, privacy concerns, or data loss, negatively affecting performance.
Low to high dimensional modality hallucination using aggregated fields of view
Kausic Gunasekar, Qiang Qiu, and Yezhou Yang. 2020 · 1990
Earlier work this paper cites.
Autoencoders, minimum description length and Helmholtz free energy
Geoffrey E Hinton and Richard Zemel. 1993 · 1993
Earlier work this paper cites.
Analysis of a sleep-dependent neuronal feedback loop: the slow-wave microcontinuity of the EEG
Bob Kemp, Aeilko H Zwinderman, Bert Tuk, Hilbert AC Kamphuisen, and Josefien JL Oberye. 2000 · 2000
Earlier work this paper cites.
The impact of the MIT-BIH arrhythmia database
George B Moody and Roger G Mark. 2001 · 2001
Earlier work this paper cites.
Semisupervised learning of classifiers: Theory, algorithms, and their application to human-computer interaction
Ira Cohen, Fabio Gagliardi Cozman, Nicu Sebe, Marcelo Cesar Cirelo, and Thomas S Huang. 2004 · 2004
Earlier work this paper cites.
Multimodal approaches for emotion recognition: a survey. In Internet Imaging VI , Vol. 5670. SPIE, 56–67
Nicu Sebe, Ira Cohen, Theo Gevers, and Thomas S Huang. 2005 · 2005
Earlier work this paper cites.
The eNTERFACE’05 audio-visual emotion database. In 22nd international conference on data engineering workshops (ICDEW’06) . IEEE, 8–8
Olivier Martin, Irene Kotsia, Benoit Macq, and Ioannis Pitas. 2006 · 2006
Earlier work this paper cites.
IEMOCAP: Interactive emotional dyadic motion capture database
Carlos Busso, Murtaza Bulut, Chi-Chun Lee, Abe Kazemzadeh, Emily Mower, Samuel Kim, Jeannette N Chang, Sungbok Lee, and Shrikanth S Narayanan. 2008 · 2008
Earlier work this paper cites.
The Alzheimer’s disease neuroimaging initiative (ADNI): MRI methods
Clifford R Jack Jr, Matt A Bernstein, Nick C Fox, Paul Thompson, Gene Alexander, Danielle Harvey, Bret Borowski, Paula J Britson, Jennifer L. Whitwell, Chadwick Ward, et al · 2008
Earlier work this paper cites.
NUS-WIDE: A Real-World Web Image Database from National University of Singapore. In Proc. of ACM Conf. on Image and Video Retrieval (CIVR’09) . Santorini, Greece
Tat-Seng Chua, Jinhui Tang, Richang Hong, Haojie Li, Zhiping Luo, and Yan-Tao Zheng. July 8-10, 2009 · 2009
Earlier work this paper cites.
Learning to recognize objects from unseen modalities. In Computer Vision–ECCV 2010 . Springer, 677–691
C Mario Christoudias, Raquel Urtasun, Mathieu Salzmann, and Trevor Darrell. 2010 · 2010
Earlier work this paper cites.
A large-scale hierarchical multi-view rgb-d object dataset. In 2011 IEEE international conference on robotics and automation . IEEE, 1817–1824
Kevin Lai, Liefeng Bo, Xiaofeng Ren, and Dieter Fox. 2011 · 2011
Earlier work this paper cites.
Towards multimodal sentiment analysis: Harvesting opinions from the web. In Proceedings of the 13th international conference on multimodal interfaces . 169–176
Louis-Philippe Morency, Rada Mihalcea, and Payal Doshi. 2011 · 2011
Earlier work this paper cites.
Exploring fusion methods for multimodal emotion recognition with missing data
Johannes Wagner, Elisabeth Andre, Florian Lingenfelser, and Jonghwa Kim. 2011 · 2011
Earlier work this paper cites.
Indoor segmentation and support inference from rgbd images. In ECCV 2012: 12th European Conference on Computer Vision, Florence, Italy, October 7-13, 2012, Proceedings, Part V 12 . Springer, 746–760
Nathan Silberman, Derek Hoiem, Pushmeet Kohli, and Rob Fergus. 2012 · 2012
Earlier work this paper cites.
Multi-source feature learning for joint analysis of incomplete multiple heterogeneous neuroimaging data
Lei Yuan, Yalin Wang, Paul M Thompson, Vaibhav A Narayan, Jieping Ye, Alzheimer’s Disease Neuroimaging Initiative, et al · 2012
Earlier work this paper cites.
eBDtheque: a representative database of comics. In 2013 12th International Conference on Document Analysis and Recognition . IEEE, 1145–1149
Clément Guérin, Christophe Rigaud, Antoine Mercier, Farid Ammar-Boudjelal, Karell Bertet, Alain Bouju, Jean-Christophe Burie, Georges Louis, Jean-Marc Ogier, and Arnaud Revel. 2013 · 2013
Earlier work this paper cites.
What’s in a name? understanding the interplay between titles, content, and communities in social media. In Proceedings of the international AAAI conference on web and social media , Vol. 7. 311–320
Himabindu Lakkaraju, Julian McAuley, and Jure Leskovec. 2013 · 2013
Earlier work this paper cites.
A randomized trial of adenotonsillectomy for childhood sleep apnea
Carole L Marcus, Reneé H Moore, Carol L Rosen, Bruno Giordani, Susan L Garetz, H Gerry Taylor, Ron B Mitchell, Raouf Amin, Eliot S Katz, Raanan Arens, et al · 2013
Earlier work this paper cites.
Introducing the RECOLA multimodal corpus of remote collaborative and affective interactions. In 2013 10th IEEE international conference and workshops on automatic face and gesture recognition (FG) . IEEE, 1–8
Fabien Ringeval, Andreas Sonderegger, Juergen Sauer, and Denis Lalanne. 2013 · 2013
Earlier work this paper cites.
Youtube movie reviews: Sentiment analysis in an audio-visual context
Martin Wöllmer, Felix Weninger, Tobias Knaup, Björn Schuller, Congkai Sun, Kenji Sagae, and Louis-Philippe Morency. 2013 · 2013
Earlier work this paper cites.
Crema-d: Crowd-sourced emotional multimodal actors dataset
Houwei Cao, David G Cooper, Michael K Keutmann, Ruben C Gur, Ani Nenkova, and Ragini Verma. 2014 · 2014
Earlier work this paper cites.
Hyperspectral and LiDAR data fusion: Outcome of the 2013 GRSS data fusion contest
Christian Debes, Andreas Merentitis, Roel Heremans, Jürgen Hahn, Nikolaos Frangiadakis, Tim van Kasteren, Wenzhi Liao, Rik Bellens, Aleksandra Pižurica, Sidharta Gautama, et al · 2014
Earlier work this paper cites.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
Deep learning based imaging data completion for improved brain disease diagnosis. In Medical Image Computing and Computer-Assisted Intervention–MICCAI 2014: 17th International Conference, Boston, MA, USA, September 14-18, 2014, Proceedings, Part III 17 . Springer, 305–312
Rongjian Li, Wenlu Zhang, Heung-Il Suk, Li Wang, Jiang Li, Dinggang Shen, and Shuiwang Ji. 2014 · 2014
Earlier work this paper cites.
Improved multimodal deep learning with variation of information
Kihyuk Sohn, Wenling Shang, and Honglak Lee. 2014 · 2014
Earlier work this paper cites.
Real-time continuous pose recovery of human hands using convolutional networks
Jonathan Tompson, Murphy Stein, Yann Lecun, and Ken Perlin. 2014 · 2014
Earlier work this paper cites.
Cross-view action modeling, learning and recognition. In Proceedings of the IEEE conference on computer vision and pattern recognition . 2649–2656
Jiang Wang, Xiaohan Nie, Yin Xia, Ying Wu, and Song-Chun Zhu. 2014 · 2014
Earlier work this paper cites.
Evaluating imputation techniques for missing data in ADNI: a patient classification study. In Progress in Pattern Recognition, Image Analysis, Computer Vision, and Applications: 20th Iberoamerican Congress, CIARP 2015, Montevideo, Uruguay, November 9-12, 2015, Proceedings 20 . Springer, 3–10
Sergio Campos, Luis Pizarro, Carlos Valle, Katherine R Gray, Daniel Rueckert, and Héctor Allende. 2015 · 2015
Earlier work this paper cites.
Distilling the knowledge in a neural network
Geoffrey Hinton, Oriol Vinyals, and Jeff Dean. 2015 · 2015
Earlier work this paper cites.
Multispectral pedestrian detection: Benchmark dataset and baseline. In Proceedings of the IEEE conference on computer vision and pattern recognition . 1037–1045
Soonmin Hwang, Jaesik Park, Namil Kim, Yukyung Choi, and In So Kweon. 2015 · 2015
Earlier work this paper cites.
Image-based recommendations on styles and substitutes. In Proceedings of the 38th international ACM SIGIR conference on research and development in information retrieval . 43–52
Julian McAuley, Christopher Targett, Qinfeng Shi, and Anton Van Den Hengel. 2015 · 2015
Earlier work this paper cites.
Review The Cancer Genome Atlas (TCGA): an immeasurable source of knowledge
Katarzyna Tomczak, Patrycja Czerwińska, and Maciej Wiznerowicz. 2015 · 2015
Earlier work this paper cites.
Learning to hash on partial multi-modal data. In 24th International Joint Conference on Artificial Intelligence
Qifan Wang, Luo Si, and Bin Shen. 2015 · 2015
Earlier work this paper cites.
Investigating critical frequency bands and channels for EEG-based emotion recognition with deep neural networks
Wei-Long Zheng and Bao-Liang Lu. 2015 · 2015
Earlier work this paper cites.
Forecasting fine-grained air quality based on big data. In Proceedings of the 21th ACM SIGKDD international conference on knowledge discovery and data mining . 2267–2276
Yu Zheng, Xiuwen Yi, Ming Li, Ruiyuan Li, Zhangqing Shan, Eric Chang, and Tianrui Li. 2015 · 2015
Earlier work this paper cites.
MSP-IMPROV: An acted corpus of dyadic interactions to study emotion perception
Carlos Busso, Srinivas Parthasarathy, Alec Burmania, Mohammed AbdelWahab, Najmeh Sadoughi, and Emily Mower Provost. 2016 · 2016
Earlier work this paper cites.
Ups and downs: Modeling the visual evolution of fashion trends with one-class collaborative filtering. In proceedings of the 25th international conference on world wide web . 507–517
Ruining He and Julian McAuley. 2016 · 2016
Earlier work this paper cites.
Learning with side information through modality hallucination. In Proceedings of the IEEE conference on computer vision and pattern recognition . 826–834
Judy Hoffman, Saurabh Gupta, and Trevor Darrell. 2016 · 2016
Earlier work this paper cites.
MIMIC-III, a freely accessible critical care database
Alistair EW Johnson, Tom J Pollard, Lu Shen, Li-wei H Lehman, Mengling Feng, Mohammad Ghassemi, Benjamin Moody, Peter Szolovits, Leo Anthony Celi, and Roger G Mark. 2016 · 2016
Earlier work this paper cites.
An open access database for the evaluation of heart sound algorithms
Chengyu Liu, David Springer, Qiao Li, Benjamin Moody, Ricardo Abad Juan, Francisco J Chorro, Francisco Castells, José Millet Roig, Ikaro Silva, Alistair EW Johnson, et al · 2016
Earlier work this paper cites.
Histogram of oriented principal components for cross-view action recognition
Hossein Rahmani, Arif Mahmood, Du Huynh, and Ajmal Mian. 2016 · 2016
Earlier work this paper cites.
Ntu rgb+ d: A large scale dataset for 3d human activity analysis. In Proceedings of the IEEE conference on computer vision and pattern recognition . 1010–1019
Amir Shahroudy, Jun Liu, Tian-Tsong Ng, and Gang Wang. 2016 · 2016
Earlier work this paper cites.
Msr-vtt: A large video description dataset for bridging video and language. In Proceedings of the IEEE conference on computer vision and pattern recognition . 5288–5296
Jun Xu, Tao Mei, Ting Yao, and Yong Rui. 2016 · 2016
Earlier work this paper cites.
Mosi: multimodal corpus of sentiment intensity and subjectivity analysis in online opinion videos
Amir Zadeh, Rowan Zellers, Eli Pincus, and Louis-Philippe Morency. 2016 · 2016
Earlier work this paper cites.
Multimodal MR synthesis via modality-invariant latent representation
Agisilaos Chartsias, Thomas Joyce, Mario Valerio Giuffrida, and Sotirios A Tsaftaris. 2017 · 2017
Earlier work this paper cites.
Learning hand articulations by hallucinating heat distribution. In Proceedings of the IEEE International Conference on Computer Vision . 3104–3113
Chiho Choi, Sangpil Kim, and Karthik Ramani. 2017 · 2017
Earlier work this paper cites.
MFNet: Towards real-time semantic segmentation for autonomous vehicles with multi-spectral scenes. In 2017 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)
Qishen Ha, Kohei Watanabe, Takumi Karasawa, Yoshitaka Ushiku, and Tatsuya Harada. 2017 · 2017
Earlier work this paper cites.
Sketch-based manga retrieval using manga109 dataset
Yusuke Matsui, Kota Ito, Yuji Aramaki, Azuma Fujimoto, Toru Ogawa, Toshihiko Yamasaki, and Kiyoharu Aizawa. 2017 · 2017
Earlier work this paper cites.
Toy products on Amazon
PromptCloud. 2017 · 2017
Earlier work this paper cites.
Hyperspectral and LiDAR fusion using extinction profiles and total variation component analysis
Behnood Rasti, Pedram Ghamisi, and Richard Gloaguen. 2017 · 2017
Earlier work this paper cites.
Multispectral object detection for autonomous vehicles. In Proceedings of the on Thematic Workshops of ACM Multimedia 2017 . 35–43
Karasawa Takumi, Kohei Watanabe, Qishen Ha, Antonio Tejero-De-Pablos, Yoshitaka Ushiku, and Tatsuya Harada. 2017 · 2017
Earlier work this paper cites.
Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results
Antti Tarvainen and Harri Valpola. 2017 · 2017
Earlier work this paper cites.
Missing modalities imputation via cascaded residual autoencoder. In Proceedings of the IEEE conference on computer vision and pattern recognition . 1405–1414
Luan Tran, Xiaoming Liu, Jiayu Zhou, and Rong Jin. 2017 · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
RGB-infrared cross-modality person re-identification. In Proceedings of the IEEE international conference on computer vision . 5380–5389
Ancong Wu, Wei-Shi Zheng, Hong-Xing Yu, Shaogang Gong, and Jianhuang Lai. 2017 · 2017
Earlier work this paper cites.
The UNC-Wisconsin rhesus macaque neurodevelopment database: a structural MRI and DTI database of early postnatal development
Jeffrey T Young, Yundi Shi, Marc Niethammer, Michael Grauer, Christopher L Coe, Gabriele R Lubach, Bradley Davis, Francois Budin, Rebecca C Knickmeyer, Andrew L Alexander, et al · 2017
Earlier work this paper cites.
Multimodal machine learning: A survey and taxonomy
Tadas Baltrušaitis, Chaitanya Ahuja, and Louis-Philippe Morency. 2018 · 2018
Earlier work this paper cites.
Overcoming missing and incomplete modalities with generative adversarial networks for building footprint segmentation. In 2018 International Conference on Content-Based Multimedia Indexing (CBMI) . IEEE, 1–6
Benjamin Bischke, Patrick Helber, Florian Koenig, Damian Borth, and Andreas Dengel. 2018 · 2018
Earlier work this paper cites.
Urban land cover classification with missing data modalities using deep convolutional neural networks
Michael Kampffmeyer, Arnt-Børre Salberg, and Robert Jenssen. 2018 · 2018
Earlier work this paper cites.
A novel public MR image dataset of multiple sclerosis patients with lesion segmentations based on multi-rater consensus
Žiga Lesjak, Alfiia Galimzianova, Aleš Koren, Matej Lukin, Franjo Pernuš, Boštjan Likar, and Žiga Špiclin. 2018 · 2018
Earlier work this paper cites.
The Ryerson Audio-Visual Database of Emotional Speech and Song (RAVDESS): A dynamic, multimodal set of facial and vocal expressions in North American English
Steven R Livingstone and Frank A Russo. 2018 · 2018
Earlier work this paper cites.
Generalized bayesian canonical correlation analysis with missing modalities. In Proceedings of the European Conference on Computer Vision (ECCV) Workshops . 0–0
Toshihiko Matsuura, Kuniaki Saito, Yoshitaka Ushiku, and Tatsuya Harada. 2018 · 2018
Earlier work this paper cites.
Learning a text-video embedding from incomplete and heterogeneous data
Antoine Miech, Ivan Laptev, and Josef Sivic. 2018 · 2018
Earlier work this paper cites.
3D MRI brain tumor segmentation using autoencoder regularization. In Brainlesion: Glioma, Multiple Sclerosis, Stroke and Traumatic Brain Injuries: 4th International Workshop, BrainLes 2018, Held in Conjunction with MICCAI 2018, Granada, Spain, September 16, 2018, Revised Selected Papers, Part II 4 . Springer, 311–320
Andriy Myronenko. 2019 · 2018
Earlier work this paper cites.
Digital comics image indexing based on deep learning
Nhu-Van Nguyen, Christophe Rigaud, and Jean-Christophe Burie. 2018 · 2018
Earlier work this paper cites.
Meld: A multimodal multi-party dataset for emotion recognition in conversations
Soujanya Poria, Devamanyu Hazarika, Navonil Majumder, Gautam Naik, Erik Cambria, and Rada Mihalcea. 2018 · 2018
Earlier work this paper cites.
Introducing wesad, a multimodal dataset for wearable stress and affect detection. In Proceedings of the 20th ACM international conference on multimodal interaction . 400–408
Philip Schmidt, Attila Reiss, Robert Duerichen, Claus Marberger, and Kristof Van Laerhoven. 2018 · 2018
Earlier work this paper cites.
Learning factorized multimodal representations
Yao-Hung Hubert Tsai, Paul Pu Liang, Amir Zadeh, Louis-Philippe Morency, and Ruslan Salakhutdinov. 2018 · 2018
Earlier work this paper cites.
Centralnet: a multilayer approach for multimodal fusion. In Proceedings of the European Conference on Computer Vision (ECCV) Workshops . 0–0
Valentin Vielzeuf, Alexis Lechervy, Stéphane Pateux, and Frédéric Jurie. 2018 · 2018
Earlier work this paper cites.
Multi-attention recurrent network for human communication comprehension. In Thirty-Second AAAI Conference on Artificial Intelligence
Amir Zadeh, Paul Pu Liang, Soujanya Poria, Prateek Vij, Erik Cambria, and Louis-Philippe Morency. 2018 · 2018
Cited alongside, same era.
Graph-based multimodal fusion with metric learning for multimodal classification
Michalis Angelou, Vassilis Solachidis, Nicholas Vretos, and Petros Daras. 2019 · 2019
Cited alongside, same era.
Robust multimodal brain tumor segmentation via feature disentanglement and gated fusion. In Medical Image Computing and Computer Assisted Intervention–MICCAI 2019: 22nd International Conference, Shenzhen, China, October 13–17, 2019, Proceedings, Part III 22 . Springer, 447–456
Cheng Chen, Qi Dou, Yueming Jin, Hao Chen, Jing Qin, and Pheng-Ann Heng. 2019 · 2019
Cited alongside, same era.
MIMIC-CXR, a de-identified publicly available database of chest radiographs with free-text reports
Alistair EW Johnson, Tom J Pollard, Seth J Berkowitz, Nathaniel R Greenbaum, Matthew P Lungren, Chih-ying Deng, Roger G Mark, and Steven Horng. 2019 · 2019
Cited alongside, same era.
MMG-ego4D: multimodal generalization in egocentric action recognition. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 6481–6491
Xinyu Gong, Sreyas Mohan, Naina Dhingra, Jean-Charles Bazin, Yilei Li, Zhangyang Wang, and Rakesh Ranjan. 2023 · 2023
Later among the works it cites.
Multi-Modal Deep Learning for Multi-Temporal Urban Mapping with a Partly Missing Optical Modality. In IGARSS 2023-2023 IEEE International Geoscience and Remote Sensing Symposium . IEEE, 6843–6846
Sebastian Hafner and Yifang Ban. 2023 · 2023
Later among the works it cites.
Planning-oriented autonomous driving. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 17853–17862
Yihan Hu, Jiazhi Yang, Li Chen, Keyu Li, Chonghao Sima, Xizhou Zhu, Siqi Chai, Senyao Du, Tianwei Lin, Wenhai Wang, et al · 2023
Later among the works it cites.
Unimf: a unified multimodal framework for multimodal sentiment analysis in missing modalities and unaligned multimodal sequences
Ruohong Huan, Guowei Zhong, Peng Chen, and Ronghua Liang. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Paul Pu Liang, Zhun Liu, Yao-Hung Hubert Tsai, Qibin Zhao, Ruslan Salakhutdinov, and Louis-Philippe Morency. 2019 · 2019
Cited alongside, same era.
MMKG: multi-modal knowledge graphs. In The Semantic Web: 16th International Conference, ESWC 2019, Portorož, Slovenia, June 2–6, 2019, Proceedings 16 . Springer, 459–474
Ye Liu, Hui Li, Alberto Garcia-Duran, Mathias Niepert, Daniel Onoro-Rubio, and David S Rosenblum. 2019 · 2019
Cited alongside, same era.
Brain tumor segmentation on MRI with missing modalities. In Information Processing in Medical Imaging: 26th International Conference, IPMI 2019, Hong Kong, China, June 2–7, 2019, Proceedings 26 . Springer, 417–428
Yan Shen and Mingchen Gao. 2019 · 2019
Cited alongside, same era.
Be your own teacher: Improve the performance of convolutional neural networks via self distillation. In Proceedings of the IEEE/CVF international conference on computer vision . 3713–3722
Linfeng Zhang, Jiebo Song, Anni Gao, Jingwei Chen, Chenglong Bao, and Kaisheng Ma. 2019 · 2019
Cited alongside, same era.
Multi-modal classification for human breast cancer prognosis prediction: proposal of deep-learning based stacked ensemble model
Nikhilanand Arya and Sriparna Saha. 2020 · 2020
Cited alongside, same era.
nuscenes: A multimodal dataset for autonomous driving. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 11621–11631
Holger Caesar, Varun Bankiti, Alex H Lang, Sourabh Vora, Venice Erin Liong, Qiang Xu, Anush Krishnan, Yu Pan, Giancarlo Baldan, and Oscar Beijbom. 2020 · 2020
Cited alongside, same era.
Hgmf: heterogeneous graph-based fusion for multimodal data with incompleteness. In Proceedings of the 26th ACM SIGKDD international conference on knowledge discovery & data mining . 1295–1305
Jiayi Chen and Aidong Zhang. 2020 · 2020
Cited alongside, same era.
Multi-Modality Matters: A Performance Leap on VoxCeleb.. In INTERSPEECH . 2252–2256
Zhengyang Chen, Shuai Wang, and Yanmin Qian. 2020 · 2020
Cited alongside, same era.
Epic-sounds: A large-scale dataset of actions that sound. In ICASSP 2023-2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 1–5
Jaesung Huh, Jacob Chalk, Evangelos Kazakos, Dima Damen, and Andrew Zisserman. 2023 · 2023
Later among the works it cites.
Multispectral video semantic segmentation: A benchmark dataset and baseline. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 1094–1104
Wei Ji, Jingjing Li, Cheng Bian, Zongwei Zhou, Jiaying Zhao, Alan L Yuille, and Li Cheng. 2023 · 2023
Later among the works it cites.
MIMIC-IV, a freely accessible electronic health record dataset
Alistair EW Johnson, Lucas Bulgarelli, Lu Shen, Alvin Gayles, Ayad Shammout, Steven Horng, Tom J Pollard, Sicheng Hao, Benjamin Moody, Brian Gow, et al · 2023
Later among the works it cites.
Otterhd: A high-resolution multi-modality model
Bo Li, Peiyuan Zhang, Jingkang Yang, Yuanhan Zhang, Fanyi Pu, and Ziwei Liu. 2023c · 2023
Later among the works it cites.
What Makes for Robust Multi-Modal Models in the Face of Missing Modalities?
Siting Li, Chenzhuang Du, Yue Zhao, Yu Huang, and Hang Zhao. 2023a · 2023
Later among the works it cites.
Gcnet: Graph completion network for incomplete multimodal learning in conversation
Zheng Lian, Lan Chen, Licai Sun, Bin Liu, and Jianhua Tao. 2023 · 2023
Later among the works it cites.
MissModal: Increasing Robustness to Missing Modality in Multimodal Sentiment Analysis
Ronghao Lin and Haifeng Hu. 2023 · 2023
Later among the works it cites.
Contrastive Intra-and Inter-Modality Generation for Enhancing Incomplete Multimedia Recommendation. In Proceedings of the 31st ACM International Conference on Multimedia . 6234–6242
Zhenghong Lin, Yanchao Tan, Yunfei Zhan, Weiming Liu, Fan Wang, Chaochao Chen, Shiping Wang, and Carl Yang. 2023 · 2023
Later among the works it cites.
Llava-plus: Learning to use tools for creating multimodal agents
Shilong Liu, Hao Cheng, Haotian Liu, Hao Zhang, Feng Li, Tianhe Ren, Xueyan Zou, Jianwei Yang, Hang Su, Jun Zhu, et al · 2023
Later among the works it cites.
A Comprehensive and Versatile Multimodal Deep-Learning Approach for Predicting Diverse Properties of Advanced Materials
Shun Muroga, Yasuaki Miki, and Kenji Hata. 2023 · 2023
Later among the works it cites.
COM: Contrastive Masked-attention model for incomplete multimodal learning
Shuwei Qian and Chongjun Wang. 2023 · 2023
Later among the works it cites.
Modal-aware visual prompting for incomplete multi-modal brain tumor segmentation. In Proceedings of the 31st ACM International Conference on Multimedia . 3228–3239
Yansheng Qiu, Ziyuan Zhao, Hongdou Yao, Delin Chen, and Zheng Wang. 2023 · 2023
Later among the works it cites.
Robust Multimodal Learning with Missing Modalities via Parameter-Efficient Adaptation
Md Kaykobad Reza, Ashley Prater-Bennette, and M Salman Asif. 2023 · 2023
Later among the works it cites.
A Survey on Brain Tumor Segmentation with Missing MRI Modalities. In International Conference on Information Technology . Springer, 299–308
Deep Shah, Amit Barve, Brijesh Vala, and Jay Gandhi. 2023 · 2023
Later among the works it cites.
Aniruddh Sikdar, Jayant Teotia, and Suresh Sundaram. 2023 · 2023
Later among the works it cites.
Hoov: Hand out-of-view tracking for proprioceptive interaction using inertial sensing. In Proceedings of the 2023 CHI Conference on Human Factors in Computing Systems . 1–16
Paul Streli, Rayan Armani, Yi Fei Cheng, and Christian Holz. 2023 · 2023
Later among the works it cites.
Vipergpt: Visual inference via python execution for reasoning. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 11888–11898
Dídac Surís, Sachit Menon, and Carl Vondrick. 2023 · 2023
Later among the works it cites.
Prototype knowledge distillation for medical segmentation with missing modality. In ICASSP 2023-2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 1–5
Shuai Wang, Zipei Yan, Daoan Zhang, Haining Wei, Zhongsen Li, and Rui Li. 2023d · 2023
Later among the works it cites.
MSH-Net: Modality-shared hallucination with joint adaptation distillation for remote sensing image classification using missing modalities
Shicai Wei, Yang Luo, Xiaoguang Ma, Peng Ren, and Chunbo Luo. 2023 · 2023
Later among the works it cites.
Towards good practices for missing modality robust action recognition. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 37. 2776–2784
Sangmin Woo, Sumin Lee, Yeonju Park, Muhammad Adi Nugroho, and Changick Kim. 2023 · 2023
Later among the works it cites.
Visual chatgpt: Talking, drawing and editing with visual foundation models
Chenfei Wu, Shengming Yin, Weizhen Qi, Xiaodong Wang, Zecheng Tang, and Nan Duan. 2023b · 2023
Later among the works it cites.
Next-gpt: Any-to-any multimodal llm
Shengqiong Wu, Hao Fei, Leigang Qu, Wei Ji, and Tat-Seng Chua. 2023a · 2023
Later among the works it cites.
Dynamic multimodal fusion. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 2575–2584
Zihui Xue and Radu Marculescu. 2023 · 2023
Later among the works it cites.
Learning unified hyper-network for multi-modal mr image synthesis and tumor segmentation with missing modalities
Heran Yang, Jian Sun, and Zongben Xu. 2023b · 2023
Later among the works it cites.
Mm-react: Prompting chatgpt for multimodal reasoning and action
Zhengyuan Yang, Linjie Li, Jianfeng Wang, Kevin Lin, Ehsan Azarnasab, Faisal Ahmed, Zicheng Liu, Ce Liu, Michael Zeng, and Lijuan Wang. 2023a · 2023
Later among the works it cites.
Unified multi-modal image synthesis for missing modality imputation
Yue Zhang, Chengtao Peng, Qiuli Wang, Dan Song, Kaiyan Li, and S Kevin Zhou. 2023b · 2023
Later among the works it cites.
Distilling Missing Modality Knowledge from Ultrasound for Endometriosis Diagnosis with Magnetic Resonance Images. In 2023 IEEE 20th International Symposium on Biomedical Imaging
Yuan Zhang, Hu Wang, David Butler, Minh-Son To, Jodie Avery, M Louise Hull, and Gustavo Carneiro. 2023c · 2023
Later among the works it cites.
Feature fusion and latent feature learning guided brain tumor segmentation and missing modality recovery network
Tongxue Zhou. 2023 · 2023
Later among the works it cites.
A literature survey of MR-based brain tumor segmentation with missing modalities
Tongxue Zhou, Su Ruan, and Haigen Hu. 2023 · 2023
Later among the works it cites.
Bin Zhu, Bin Lin, Munan Ning, Yang Yan, Jiaxi Cui, HongFa Wang, Yatian Pang, Wenhao Jiang, Junwu Zhang, Zongwei Li, et al · 2023
Later among the works it cites.
Juhan Cha, Minseok Joo, Jihwan Park, Sanghyeok Lee, Injae Kim, and Hyunwoo J. Kim. 2024 · 2024
Closest in time.
StressID: a Multimodal Dataset for Stress Identification
Hava Chaptoukaev, Valeriya Strizhkova, Michele Panariello, Bianca Dalpaos, Aglind Reka, Valeria Manera, Susanne Thümmler, Esma Ismailova, Massimiliano Todisco, Maria A Zuluaga, et al · 2024
Closest in time.
Modality-Specific Information Disentanglement From Multi-Parametric MRI for Breast Tumor Segmentation and Computer-Aided Diagnosis
Qianqian Chen, Jiadong Zhang, Runqi Meng, Lei Zhou, Zhenhui Li, Qianjin Feng, and Dinggang Shen. 2024b · 2024
Closest in time.
Towards Multimodal Video Paragraph Captioning Models Robust to Missing Modality
Sishuo Chen, Lei Li, Shuhuai Ren, Rundong Gao, Yuanxin Liu, Xiaohan Bi, Xu Sun, and Lu Hou. 2024a · 2024
Closest in time.
After Three Years on Mars, NASA’s Ingenuity Helicopter Mission Ends - NASA — nasa.gov
Abbey A. Donaldson. 2024 · 2024
Closest in time.
Multimodal Prompt Learning with Missing Modalities for Sentiment Analysis and Emotion Recognition
Zirun Guo, Tao Jin, and Zhou Zhao. 2024 · 2024
Closest in time.
Mosaic integration and knowledge transfer of single-cell multimodal data with MIDAS
Zhen He, Shuofeng Hu, Yaowen Chen, Sijing An, Jiahao Zhou, Runyan Liu, Junfeng Shi, Jing Wang, Guohua Dong, Jinhui Shi, et al · 2024
Closest in time.
Towards Robust Multimodal Prompting with Missing Modalities. In ICASSP 2024-2024 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 8070–8074
Jaehyuk Jang, Yooseung Wang, and Changick Kim. 2024 · 2024
Closest in time.
Deformation-aware and reconstruction-driven multimodal representation learning for brain tumor segmentation with missing modalities
Zhiyuan Li, Yafei Zhang, Huafeng Li, Yi Chai, and Yushi Yang. 2024b · 2024
Closest in time.
Suppress and Rebalance: Towards Generalized Multi-Modal Face Anti-Spoofing
Xun Lin, Shuai Wang, Rizhao Cai, Yizhong Liu, Ying Fu, Zitong Yu, Wenzhong Tang, and Alex Kot. 2024 · 2024
Closest in time.
Visual instruction tuning
Haotian Liu, Chunyuan Li, Qingyang Wu, and Yong Jae Lee. 2024a · 2024
Closest in time.
Fourier Prompt Tuning for Modality-Incomplete Scene Segmentation
Ruiping Liu, Jiaming Zhang, Kunyu Peng, Yufan Chen, Ke Cao, Junwei Zheng, M Saquib Sarfraz, Kailun Yang, and Rainer Stiefelhagen. 2024b · 2024
Closest in time.
MC-DBN: A Deep Belief Network-Based Model for Modality Completion
Zihong Luo, Haochen Xue, Mingyu Jin, Chengzhi Liu, Zile Huang, Chong Zhang, and Shuliang Zhao. 2024 · 2024
Closest in time.
Dealing with Missing Modalities in Multimodal Recommendation: a Feature Propagation-based Approach
Daniele Malitesta, Emanuele Rossi, Claudio Pomo, Fragkiskos D Malliaros, and Tommaso Di Noia. 2024 · 2024
Closest in time.
The multimodal brain tumor image segmentation benchmark (BRATS)
Bjoern H Menze, Andras Jakab, Stefan Bauer, Jayashree Kalpathy-Cramer, Keyvan Farahani, Justin Kirby, Yuliya Burren, Nicole Porz, Johannes Slotboom, Roland Wiest, et al · 2024
Closest in time.
ADAPT: Multimodal Learning for Detecting Physiological Changes under Missing Modalities. In Medical Imaging with Deep Learning
Julie Mordacq, Leo Milecki, Maria Vakalopoulou, Steve Oudot, and Vicky Kalogeiton. 2024 · 2024
Closest in time.
Learning Trimodal Relation for Audio-Visual Question Answering with Missing Modality
Kyu Ri Park, Hong Joo Lee, and Jung Uk Kim. 2024 · 2024
Closest in time.
2024 IEEE GRSS Data Fusion Contest - Flood Rapid Mapping
Claudio Persello; Saurabh Prasad; Gemine Vivone; Vincent Lonjou ; Frédéric Bretar ; Raquel Rodriguez-Suquet ; Pauline Guntzburger ; Vincent Poulain ; Jacqueline Le Moigne; Benjamin Smith ; Sujay Kumar ; Thomas Huang ; Sophie Ricci ; Thanh Huy Nguyen ; Andrea Piacentini. 2023 · 2024
Closest in time.
Humanoid Locomotion as Next Token Prediction
Ilija Radosavovic, Bike Zhang, Baifeng Shi, Jathushan Rajasegaran, Sarthak Kamat, Trevor Darrell, Koushil Sreenath, and Jitendra Malik. 2024 · 2024
Closest in time.
Combating Missing Modalities in Egocentric Videos at Test Time
Merey Ramazanova, Alejandro Pardo, Bernard Ghanem, and Motasem Alfarra. 2024 · 2024
Closest in time.
Pramit Saha, Divyanshu Mishra, Felix Wagner, Konstantinos Kamnitsas, and J Alison Noble. 2024 · 2024
Closest in time.
Hugginggpt: Solving ai tasks with chatgpt and its friends in hugging face
Yongliang Shen, Kaitao Song, Xu Tan, Dongsheng Li, Weiming Lu, and Yueting Zhuang. 2024 · 2024
Closest in time.
Junjie Shi, Caozhi Shang, Zhaobin Sun, Li Yu, Xin Yang, and Zengqiang Yan. 2024 · 2024
Closest in time.
Similar modality completion-based multimodal sentiment analysis under uncertain missing modalities
Yuhang Sun, Zhizhong Liu, Quan Z Sheng, Dianhui Chu, Jian Yu, and Hongxiang Sun. 2024a · 2024
Closest in time.
Any-to-any generation via composable diffusion
Zineng Tang, Ziyi Yang, Chenguang Zhu, Michael Zeng, and Mohit Bansal. 2024 · 2024
Closest in time.
Hu Wang, Congbo Ma, Yuyuan Liu, Yuanhong Chen, Yu Tian, Jodie Avery, Louise Hull, and Gustavo Carneiro. 2024b · 2024
Closest in time.
Incomplete multimodality-diffused emotion recognition
Yuanzhi Wang, Yong Li, and Zhen Cui. 2024a · 2024
Closest in time.
Towards semantic consistency: Dirichlet energy driven robust multi-modal entity alignment
Yuanyi Wang, Haifeng Sun, Jiabo Wang, Jingyu Wang, Wei Tang, Qi Qi, Shaoling Sun, and Jianxin Liao. 2024c · 2024
Closest in time.
Multimodal patient representation learning with missing modalities and labels. In The Twelfth International Conference on Learning Representations
Zhenbang Wu, Anant Dadu, Nicholas Tustison, Brian Avants, Mike Nalls, Jimeng Sun, and Faraz Faghri. 2024 · 2024
Closest in time.
Incomplete learning of multi-modal connectome for brain disorder diagnosis via modal-mixup and deep supervision. In Medical Imaging With Deep Learning . PMLR, 1006–1018
Yanwu Yang, Hairui Chen, Zhikai Chang, Yang Xiang, Chenfei Ye, and Ting Ma. 2024 · 2024
Closest in time.
DrFuse: Learning Disentangled Representation for Clinical Multi-Modal Fusion with Missing Modality and Modal Inconsistency. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 38. 16416–16424
Wenfang Yao, Kejing Yin, William K Cheung, Jia Liu, and Jing Qin. 2024 · 2024
Closest in time.
Anygpt: Unified multimodal llm with discrete sequence modeling
Jun Zhan, Junqi Dai, Jiasheng Ye, Yunhua Zhou, Dong Zhang, Zhigeng Liu, Xin Zhang, Ruibin Yuan, Ge Zhang, Linyang Li, et al · 2024
Closest in time.
TMFormer: Token Merging Transformer for Brain Tumor Segmentation with Missing Modalities. In Proceedings of the AAAI Conference on Artificial Intelligence
Zheyu Zhang, Gang Yang, Yueyi Zhang, Huanjing Yue, Aiping Liu, Yunwei Ou, Jian Gong, and Xiaoyan Sun. 2024 · 2024
Closest in time.
Deep Multimodal Data Fusion
Fei Zhao, Chengcui Zhang, and Baocheng Geng. 2024 · 2024
Closest in time.
Learning Modality-agnostic Representation for Semantic Segmentation from Any Modalities
Xu Zheng, Yuanhuiyi Lyu, and Lin Wang. 2024 · 2024
Closest in time.
Zhuo Zhi, Ziquan Liu, Moe Elbadawi, Adam Daneshmend, Mine Orlu, Abdul Basit, Andreas Demosthenous, and Miguel Rodrigues. 2024 · 2024
Closest in time.