MoEA-Net: Modality-Incremental Expert Aggregation Network for Retinal Prognostic Prediction
Hua Wang, Xiaodan Zhang, Yanzhao Shi, Chengxin Zheng, Wanyu Zhang, Zhen Wang, Jianing Wang, Xiaobing Yu
Abstract
Automated analysis of temporal changes in multimodal retinal images is critical for the prognostic assessment of ophthalmic diseases. Unlike traditional single-timepoint diagnosis, tracking longitudinal changes across multiple imaging modalities introduces significant data bias challenges:
(1) Imbalanced modality samples compromise the integration of knowledge within minority modalities; (2) Heterogeneous visual patterns across modalities undermine the perception of disease-relevant biomarkers. To tackle these issues, we propose a Modality-Incremental Expert Aggregation Network (MoEA-Net), which unifies the inter-modal integration and intra-modal perception for enhanced retinal prognostic prediction. Specifically, we employ the large language model (LLM) with incremental LoRA layers for specific modalities to effectively integrate knowledge from imbalanced data. Besides, we introduce a Spatiotemporal-aware Expert (SAE) module to better perceive both the anatomical structures and longitudinal changes within modalities. By progressively combining the SAE module with incremental LoRA, MoEA-Net supports continual knowledge accumulation and improves accurate reasoning. Experimental results show that MoEA-Net achieves state-of-the-art performance on subretinal fluid change and visual recovery classification tasks, validating its effectiveness.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f0ad62d8-7ff6-4b6e-97a3-8ad866a4ff9eBuilds on4
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- MEPNet: Medical Entity-Balanced Prompting Network for Brain CT Report GenerationXiaodan Zhang, Yanzhao Shi, Junzhong Ji, Chengxin Zheng et al.AAAI 2025 · 5 citations
- Granularity Matters: Pathological Graph-driven Cross-modal Alignment for Brain CT Report GenerationYanzhao Shi, Junzhong Ji, Xiaodan Zhang, Liangqiong Qu et al.EMNLP 2023 · 5 citations
- Selective Aggregation for Low-Rank Adaptation in Federated LearningPengxin Guo, Shuang Zeng, Yanran Wang, Huijie Fan et al.ICLR 2025
Related papers
- X-PCR: A Benchmark for Cross-modality Progressive Clinical Reasoning in Ophthalmic DiagnosisGui Wang, Zehao Zhong, YongSong Zhou, Yudong Li et al.CVPR 2026
- Contrastive Regularization over LoRA for Multimodal Biomedical Image Incremental LearningHaojie Zhang, Yixiong Liang, Hulin Kuang, Lihui Cen et al.ACM MM 2025 · 2 citations
- R2-T2: Re-Routing in Test-Time for Multimodal Mixture-of-ExpertsZhongyang Li, Ziyue Li, Tianyi ZhouICML 2025
- CL-MoE: Enhancing Multimodal Large Language Model with Dual Momentum Mixture-of-Experts for Continual Visual Question AnsweringTianyu Huai, Jie Zhou, Xingjiao Wu, Qin Chen et al.CVPR 2025
- Multi-Modal Multi-Instance Learning for Retinal Disease RecognitionXirong Li, Yang Zhou, Jie Wang, Hailan Lin et al.ACM MM 2021 · 52 citations
