Like an Ophthalmologist: Dynamic Selection Driven Multi-View Learning for Diabetic Retinopathy Grading
Xiaoling Luo, Qihao Xu, Huisi Wu, Chengliang Liu, Zhihui Lai, Linlin Shen
Abstract
Diabetic retinopathy (DR), with its large patient population, has become a formidable threat to human visual health. In the clinical diagnosis of DR, multi-view fundus images are considered to be more suitable for DR diagnosis because of the wide coverage of the field of view. Therefore, different from most of the previous single-view DR grading methods, we design a dynamic selection-driven multi-view DR grading method to fit clinical scenarios better. Since lesion information plays a key role in DR diagnosis, previous methods usually boost the model performance by enhancing the lesion feature. However, during the actual diagnosis, ophthalmologists not only focus on the crucial parts, but also exclude irrelevant features to ensure the accuracy of judgment. To this end, we introduce the idea of dynamic selection and design a series of selection mechanisms from fine granularity to coarse granularity. In this work, we first introduce an Ophthalmic Image Reader (OIR) agent to provide the model with pixel-level prompts of suspected lesion areas. Moreover, a Multi-View Token Selection Module (MVTSM) is designed to prune redundant feature tokens and realize dynamic selection of key information. In the final decision stage, we dynamically fuse multi-view features through the novel Multi-View Mixture of Experts Module (MVMoEM), to enhance key views and reduce the impact of conflicting views. Extensive experiments on a large multi-view fundus image dataset with 34,452 images demonstrate that our method performs favorably against state-of-the-art models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ca818ca0-665a-4522-8bc2-a535d5a6a964Cited by top-tier papers9
- Hierarchical Information Aggregation for Incomplete Multimodal Alzheimer's Disease DiagnosisChengliang Liu, Yuanxi Que, Qihao Xu, Yabo Liu et al.NeurIPS 2025 · 2 citations
- ProConMV: Provenance-Enabled Conceptual Framework for Interpretable Multi-View Diabetic Retinopathy DiagnosisXiaoling Luo, Shuo Yang, Qihao Xu, Jiansong Zhang et al.ICML 2026
- Neural Collapse Priors Driven Trust Semi-Supervised Multi-View ClassificationTaotao Guo, Honglin Yuan, Xujian Zhao, Yuan Sun et al.AAAI 2026
- Cross-View Distillation and Adaptive Masking for Incomplete Multi-View Multi-Label ClassificationYadong Liu, Qiaoqi Li, Yueying Wang, Lunke Fei et al.CVPR 2026
- Quality-aware and Soft Consistency Driven Representation Fusion for Incomplete Multi-view Multi-label ClassificationYadong Liu, Waikeung Wong, Yulong Chen, Jie WenAAAI 2026
Builds on12
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu et al.ICCV 2021 · 31,683 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space ModelLianghui Zhu, Bencheng Liao, Qian Zhang, Xinlong Wang et al.ICML 2024 · 1,725 citations
- Scaling Vision with Sparse Mixture of ExpertsCarlos Riquelme, Joan Puigcerver, Basil Mustafa, Maxim Neumann et al.NeurIPS 2021 · 1,213 citations
- Mixture-of-Experts with Expert Choice RoutingYanqi Zhou, Tao Lei, Hanxiao Liu, Nan Du et al.NeurIPS 2022 · 933 citations
Related papers
- Towards Zero-Shot Diabetic Retinopathy Grading: Learning Generalized Knowledge via Prompt-Driven Matching and EmulatingHuan Wang, Haoran Li, Yuxin Lin, Huaming Chen et al.AAAI 2026
- Deep Multi-Task Learning for Diabetic Retinopathy Grading in Fundus ImagesXiaofei Wang, Mai Xu, Jicong Zhang, Lai Jiang et al.AAAI 2021 · 48 citations
- MVCINN: Multi-View Diabetic Retinopathy Detection Using a Deep Cross-Interaction Neural NetworkXiaoling Luo, Chengliang Liu, Waikeung Wong, Jie Wen et al.AAAI 2023 · 14 citations
- Incomplete Multi-view Diabetic Retinopathy Grading via Self-Supervised Inter- and Intra-View RestorationZhihao Wu, Yuxin Lin, Jie Wen, Wuzhen Shi et al.AAAI 2026
- Vision-Language Models Guided Graph Concept Reasoning for Interpretable Diabetic Retinopathy DiagnosisQihao Xu, Xiaoling Luo, Yuxin Lin, Chengliang Liu et al.AAAI 2026
