Multimodal Patient Representation Learning with Missing Modalities and Labels
Zhenbang Wu, Anant Dadu, Nicholas J. Tustison, Brian B. Avants, Mike A. Nalls, Jimeng Sun, Faraz Faghri
摘要
Multimodal patient representation learning aims to integrate information from multiple modalities and generate comprehensive patient representations for subsequent clinical predictive tasks. However, many existing approaches either presuppose the availability of all modalities and labels for each patient or only deal with missing modalities. In reality, patient data often comes with both missing modalities and labels for various reasons (i.e., the missing modality and label issue). Moreover, multimodal models might over-rely on certain modalities, causing suboptimal performance when these modalities are absent (i.e., the modality collapse issue). To address these issues, we introduce MUSE: a mutual-consistent graph contrastive learning method. MUSE uses a flexible bipartite graph to represent the patient-modality relationship, which can adapt to various missing modality patterns. To tackle the modality collapse issue, MUSE learns to focus on modalitygeneral and label-decisive features via a mutual-consistent contrastive learning loss. Notably, the unsupervised component of the contrastive objective only requires self-supervision signals, thereby broadening the training scope to incorporate patients with missing labels. We evaluate MUSE on three publicly available datasets: MIMIC-IV, eICU, and ADNI. Results show that MUSE outperforms all baselines, and MUSE+ further elevates the absolute improvement to ∼4% by extending the training scope to patients with absent labels.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- Adversarial Graph Fusion for Incomplete Multi-view Semi-supervised Learning with Tensorial ImputationZhangqi Jiang, Tingjin Luo, Xu Yang, Xinyan LiangNeurIPS 2025 · 被引用 6 次
- To Predict or Not to Predict? Proportionally Masked Autoencoders for Tabular Data ImputationJungkyu Kim, Kibok Lee, Taeyoung ParkAAAI 2025 · 被引用 4 次
- Inference-Time Dynamic Modality Selection for Incomplete Multimodal ClassificationSiyi Du, Xinzhe Luo, Declan O'regan, Chen QinICLR 2026 · 被引用 4 次
- NTSFormer: A Self-Teaching Graph Transformer for Multimodal Isolated Cold-Start Node ClassificationJun Hu, Yufei He, Yuan Li, Bryan Hooi 等AAAI 2026 · 被引用 3 次
- Unified Insights: Harnessing Multi-modal Data for Phenotype Imputation via View DecouplingQiannan Zhang, Weishen Pan, Zilong Bai, Chang Su 等NeurIPS 2024 · 被引用 1 次
它引用的顶会 Paper8
- ViLT: Vision-and-Language Transformer Without Convolution or Region SupervisionWonjae Kim, Bokyung Son, Ildoo KimICML 2021 · 被引用 2,258 次
- SMIL: Multimodal Learning with Severely Missing ModalityMengmeng Ma, Jian Ren, Long Zhao, Sergey Tulyakov 等AAAI 2021 · 被引用 393 次
- Handling Missing Data with Graph Representation LearningJiaxuan You, Xiaobai Ma, Daisy Yi Ding, Mykel J. Kochenderfer 等NeurIPS 2020 · 被引用 274 次
- Are Multimodal Transformers Robust to Missing Modality?Mengmeng Ma, Jian Ren, Long Zhao, Davide Testuggine 等CVPR 2022 · 被引用 153 次
- HGMF: Heterogeneous Graph-based Fusion for Multimodal Data with IncompletenessJiayi Chen, Aidong ZhangKDD 2020 · 被引用 89 次
相关 Paper
- Causal Representation Learning from Multimodal Clinical Records under Non-Random Modality MissingnessZihan Liang, Ziwen Pan, Ruoxuan XiongEMNLP 2025
- Contrasting with Symile: Simple Model-Agnostic Representation Learning for Unlimited ModalitiesAdriel Saporta, Aahlad Manas Puli, Mark Goldstein, Rajesh RanganathNeurIPS 2024 · 被引用 28 次
- Multimodal Emotion Recognition with Missing Modality via a Unified Multi-task Pre-training FrameworkZiyi Li, Wei-Long Zheng, Bao-Liang LuACM MM 2025 · 被引用 2 次
- Geometric Multimodal Contrastive Representation LearningPetra Poklukar, Miguel Vasco, Hang Yin, Francisco S. Melo 等ICML 2022 · 被引用 68 次
- Multi-view Graph Contrastive Representation Learning for Drug-Drug Interaction PredictionYingheng Wang, Yaosen Min, Xin Chen, Ji WuWWW 2021 · 被引用 186 次
