Multimodal Patient Representation Learning with Missing Modalities and Labels
Zhenbang Wu, Anant Dadu, Nicholas J. Tustison, Brian B. Avants, Mike A. Nalls, Jimeng Sun, Faraz Faghri
Abstract
Multimodal patient representation learning aims to integrate information from multiple modalities and generate comprehensive patient representations for subsequent clinical predictive tasks. However, many existing approaches either presuppose the availability of all modalities and labels for each patient or only deal with missing modalities. In reality, patient data often comes with both missing modalities and labels for various reasons (i.e., the missing modality and label issue). Moreover, multimodal models might over-rely on certain modalities, causing suboptimal performance when these modalities are absent (i.e., the modality collapse issue). To address these issues, we introduce MUSE: a mutual-consistent graph contrastive learning method. MUSE uses a flexible bipartite graph to represent the patient-modality relationship, which can adapt to various missing modality patterns. To tackle the modality collapse issue, MUSE learns to focus on modalitygeneral and label-decisive features via a mutual-consistent contrastive learning loss. Notably, the unsupervised component of the contrastive objective only requires self-supervision signals, thereby broadening the training scope to incorporate patients with missing labels. We evaluate MUSE on three publicly available datasets: MIMIC-IV, eICU, and ADNI. Results show that MUSE outperforms all baselines, and MUSE+ further elevates the absolute improvement to ∼4% by extending the training scope to patients with absent labels.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 534328ed-8f0f-4ef2-a861-b232daa72106Cited by top-tier papers16
- Adversarial Graph Fusion for Incomplete Multi-view Semi-supervised Learning with Tensorial ImputationZhangqi Jiang, Tingjin Luo, Xu Yang, Xinyan LiangNeurIPS 2025 · 6 citations
- To Predict or Not to Predict? Proportionally Masked Autoencoders for Tabular Data ImputationJungkyu Kim, Kibok Lee, Taeyoung ParkAAAI 2025 · 4 citations
- Inference-Time Dynamic Modality Selection for Incomplete Multimodal ClassificationSiyi Du, Xinzhe Luo, Declan O'regan, Chen QinICLR 2026 · 4 citations
- NTSFormer: A Self-Teaching Graph Transformer for Multimodal Isolated Cold-Start Node ClassificationJun Hu, Yufei He, Yuan Li, Bryan Hooi et al.AAAI 2026 · 3 citations
- Unified Insights: Harnessing Multi-modal Data for Phenotype Imputation via View DecouplingQiannan Zhang, Weishen Pan, Zilong Bai, Chang Su et al.NeurIPS 2024 · 1 citation
Builds on8
- ViLT: Vision-and-Language Transformer Without Convolution or Region SupervisionWonjae Kim, Bokyung Son, Ildoo KimICML 2021 · 2,258 citations
- SMIL: Multimodal Learning with Severely Missing ModalityMengmeng Ma, Jian Ren, Long Zhao, Sergey Tulyakov et al.AAAI 2021 · 393 citations
- Handling Missing Data with Graph Representation LearningJiaxuan You, Xiaobai Ma, Daisy Yi Ding, Mykel J. Kochenderfer et al.NeurIPS 2020 · 274 citations
- Are Multimodal Transformers Robust to Missing Modality?Mengmeng Ma, Jian Ren, Long Zhao, Davide Testuggine et al.CVPR 2022 · 153 citations
- HGMF: Heterogeneous Graph-based Fusion for Multimodal Data with IncompletenessJiayi Chen, Aidong ZhangKDD 2020 · 89 citations
Related papers
- Causal Representation Learning from Multimodal Clinical Records under Non-Random Modality MissingnessZihan Liang, Ziwen Pan, Ruoxuan XiongEMNLP 2025
- Contrasting with Symile: Simple Model-Agnostic Representation Learning for Unlimited ModalitiesAdriel Saporta, Aahlad Manas Puli, Mark Goldstein, Rajesh RanganathNeurIPS 2024 · 28 citations
- Multimodal Emotion Recognition with Missing Modality via a Unified Multi-task Pre-training FrameworkZiyi Li, Wei-Long Zheng, Bao-Liang LuACM MM 2025 · 2 citations
- Geometric Multimodal Contrastive Representation LearningPetra Poklukar, Miguel Vasco, Hang Yin, Francisco S. Melo et al.ICML 2022 · 68 citations
- Multi-view Graph Contrastive Representation Learning for Drug-Drug Interaction PredictionYingheng Wang, Yaosen Min, Xin Chen, Ji WuWWW 2021 · 186 citations
