RedCore: Relative Advantage Aware Cross-Modal Representation Learning for Missing Modalities with Imbalanced Missing Rates
Jun Sun, Xinxin Zhang, Shoukang Han, Yu-Ping Ruan, Taihao Li
摘要
Multimodal learning is susceptible to modality missing, which poses a major obstacle for its practical applications and, thus, invigorates increasing research interest. In this paper, we investigate two challenging problems: 1) when modality missing exists in the training data, how to exploit the incomplete samples while guaranteeing that they are properly supervised? 2) when the missing rates of different modalities vary, causing or exacerbating the imbalance among modalities, how to address the imbalance and ensure all modalities are well-trained? To tackle these two challenges, we first introduce the variational information bottleneck (VIB) method for the cross-modal representation learning of missing modalities, which capitalizes on the available modalities and the labels as supervision. Then, accounting for the imbalanced missing rates, we define relative advantage to quantify the advantage of each modality over others. Accordingly, a bilevel optimization problem is formulated to adaptively regulate the supervision of all modalities during training. As a whole, the proposed approach features Relative advantage aware Cross-modal representation learning (abbreviated as RedCore) for missing modalities with imbalanced missing rates. Extensive empirical results demonstrate that RedCore outperforms competing models in that it exhibits superior robustness against either large or imbalanced missing rates.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- PASSION: Towards Effective Incomplete Multi-Modal Medical Image Segmentation with Imbalanced Missing RatesJunjie Shi, Caozhi Shang, Zhaobin Sun, Li Yu 等ACM MM 2024 · 被引用 20 次
- Incomplete Modality Disentangled Representation for Ophthalmic Disease Grading and DiagnosisChengzhi Liu, Zile Huang, Zhe Chen, Feilong Tang 等AAAI 2025 · 被引用 10 次
- BALM: A Model-Agnostic Framework for Balanced Multimodal Learning under Imbalanced Missing RatesPhuong-Anh Nguyen, Tien Anh Pham, Duc-Trong Le, Cam-Van Thi NguyenCVPR 2026 · 被引用 1 次
- Boomda: Balanced Multi-objective Optimization for Multimodal Domain AdaptationJun Sun, Xinxin Zhang, Simin Hong, Jian Zhu 等AAAI 2026 · 被引用 1 次
- Diffmv: A Unified Diffusion Framework for Healthcare Predictions with Random Missing Views and View LazinessChuang Zhao, Hui Tang, Hongke Zhao, Xiaomeng LiKDD 2025
它引用的顶会 Paper8
- SMIL: Multimodal Learning with Severely Missing ModalityMengmeng Ma, Jian Ren, Long Zhao, Sergey Tulyakov 等AAAI 2021 · 被引用 393 次
- Balanced Multimodal Learning via On-the-fly Gradient ModulationXiaokang Peng, Yake Wei, Andong Deng, Dong Wang 等CVPR 2022 · 被引用 264 次
- M3AE: Multimodal Representation Learning for Brain Tumor Segmentation with Missing ModalitiesHong Liu, Dong Wei, Donghuan Lu, Jinghan Sun 等AAAI 2023 · 被引用 101 次
- Semi-supervised Multi-modal Emotion Recognition with Cross-Modal Distribution MatchingJingjun Liang, Ruichen Li, Qin JinACM MM 2020 · 被引用 67 次
- Fairness without Imputation: A Decision Tree Approach for Fair Prediction with Missing ValuesHaewon Jeong, Hao Wang, Flávio P. CalmonAAAI 2022 · 被引用 48 次
相关 Paper
- Learning Optimal Multimodal Information Bottleneck RepresentationsQilong Wu, Yiyang Shao, Jun Wang, Xiaobo SunICML 2025
- I3-MRec: Invariant Learning with Information Bottleneck for Incomplete Modality RecommendationHuilin Chen, Miaomiao Cai, Fan Liu, Zhiyong Cheng 等ACM MM 2025 · 被引用 1 次
- IBMA: Information Bottleneck-Based Multimodal AlignmentYancheng Wang, Zeyu Dong, Dongfang Sun, Alvin Silva 等ICML 2026
- Disentangled Cross-Modal Representation Learning with Enhanced Mutual SupervisionLu Gao, Wenlan Chen, Daoyuan Wang, Fei Guo 等NeurIPS 2025 · 被引用 5 次
- Permutation-Consistent Variational Encoding for Incomplete Multi-View Multi-Label ClassificationChengliang Liu, Bo Li, Bob Zhang, Xiaoling Luo 等ICLR 2026
