FACT: Fuzzy Alignment with Comorbidity Topology for Reliable Multi-Label Medical Image Diagnosis
Yingyu Chen, Yongqiang Huang, Yang Qin, Ziyuan Yang, Lang Yuan, Maosong Ran, Yi Zhang
Abstract
In clinical practice, patients often present with multiple co-occurring diseases, yet most existing Multi-Label-Diagnosis (MLD) methods treat diagnosis as a rigid discriminative partitioning task, implicitly assuming that overlapping pathologies are separable. This assumption is problematic in medical images, where identical or highly similar visual observations may simultaneously support multiple disease labels, and disease concepts are inherently correlated rather than independent. Enforcing hard decision boundaries under such overlap suppresses shared evidence, biases feature representations, and ultimately undermines model reliability. To address this limitation, we propose Fuzzy Alignment with Comorbidity Topology (FACT), a novel paradigm that reformulates MLD as a fuzzy alignment problem between atomic visual evidence and disease semantic anchors. FACT is characterized by three key features: (1) modeling visual polysemy through shared and reusable atomic visual evidence; (2) encoding disease correlation via semantic anchors structured by comorbidity topology; and
(3) employing a metric-based fuzzy membership function for non-discriminative visual-semantic alignment. Extensive experiments on three public clinical benchmarks, together with additional evaluation under long-tailed and noisy settings, demonstrate that FACT consistently improves diagnostic performance while delivering clinically plausible predictions. The code is available at https://github.com/yyuChen9/FACT.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bcf4c4f3-7382-44ad-8d70-d30ed622fc14Builds on10
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Gradient Surgery for Multi-Task LearningTianhe Yu, Saurabh Kumar, Abhishek Gupta, Sergey Levine et al.NeurIPS 2020 · 2,261 citations
- Asymmetric Loss For Multi-Label ClassificationTal Ridnik, Emanuel Ben Baruch, Nadav Zamir, Asaf Noy et al.ICCV 2021 · 778 citations
- Multi-Label Supervised Contrastive LearningPingyue Zhang, Mengyue WuAAAI 2024 · 42 citations
Related papers
- Scalable Medical Multimodal Fusion via Symmetric Consistency ModelingXiaowen Sun, Hui Liu, Gongguan Chen, Ning MaoICML 2026
- Rethinking Federated Prompt Learning for Medical Images: From Textual Tuning to Visual Manifold AnchoringYipan Wei, Wenke Huang, Yapeng Li, He Li et al.ICML 2026
- Multi-Granularity Cross-modal Alignment for Generalized Medical Visual Representation LearningFuying Wang, Yuyin Zhou, Shujun Wang, Varut Vardhanabhuti et al.NeurIPS 2022 · 302 citations
- Enhance Multi-View Classification Through Multi-Scale Alignment and Expanded BoundaryYuena Lin, Yiyuan Wang, Gengyu Lyu, Yongjian Deng et al.ICLR 2025
- Intra-view and Inter-view Correlation Guided Multi-view Novel Class DiscoveryXinhang Wan, Jiyuan Liu, Qian Qu, Suyuan Liu et al.ICCV 2025
