Label Decoupling and Reconstruction: A Two-Stage Training Framework for Long-tailed Multi-label Medical Image Recognition
Jie Huang, Zhao-Min Chen, Xiaoqin Zhang, Yisu Ge, Lusi Ye, Guodao Zhang, Huiling Chen
摘要
Deep learning has made significant advancements and breakthroughs in medical image recognition. However, the clinical reality is complex and multifaceted, with patients often suffering from multiple intertwined diseases, not all of which are equally common, leading to medical datasets that are frequently characterized by multi-labels and a long-tailed distribution. In this paper, we propose a method involving label decoupling and reconstruction (LDRNet) to address these two specific challenges. The label decoupling utilizes the fusion of semantic information from both categories and images to capture the class-aware features across different labels. This process not only integrates semantic information from labels and images to improve the model's ability to recognize diseases, but also captures comprehensive features across various labels to facilitate a deeper understanding of disease characteristics within the dataset. Following this, our label reconstruction method uses the class-aware features to reconstruct the label distribution. This step generates a diverse array of virtual features for tail categories, promoting unbiased learning for the classifier and significantly enhancing the model's generalization ability and robustness. Extensive experiments conducted on three multi-label long-tailed medical image datasets, including the Axial Spondyloarthritis Dataset, NIH Chest X-ray 14 Dataset, and ODIR-5K Dataset, have demonstrated that our approach achieves state-of-the-art performance, showcasing its effectiveness in handling the complexities associated with multi-label and long-tailed distributions in medical image recognition.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper5
- UniMedVL: Unifying Medical Multimodal Understanding and Generation through Observation-Knowledge-AnalysisJunzhi Ning, Wei Li, Cheng Tang, Jiashi Lin 等ICML 2026 · 被引用 13 次
- Overcoming Dual Drift for Continual Long-Tailed Visual Question AnsweringFeifei Zhang, Zhihao Wang, Xi Zhang, Changsheng XuICCV 2025 · 被引用 3 次
- Prototype-based Causal Intervention for Multi-Label Image ClassificationYanmin Li, Zhilong Mao, Mao Wang, Lihua Liu 等CVPR 2026
- FACT: Fuzzy Alignment with Comorbidity Topology for Reliable Multi-Label Medical Image DiagnosisYingyu Chen, Yongqiang Huang, Yang Qin, Ziyuan Yang 等ICML 2026
- Beyond Traditional Diagnostics: Transforming Patient-Side Information Into Predictive Insights with Knowledge Graphs and PrototypesYibowen Zhao, Yinan Zhang, Zhixiang Su, Li-Zhen Cui 等ICDE 2026
相关 Paper
- Long-tailed Recognition with Model RebalancingJiaan Luo, Feng Hong, Qiang Hu, Xiaofeng Cao 等NeurIPS 2025 · 被引用 12 次
- Long-Tailed Anomaly Detection with Learnable Class NamesChih-Hui Ho, Kuan-Chuan Peng, Nuno VasconcelosCVPR 2024
- Self Supervision to Distillation for Long-Tailed Visual RecognitionTianhao Li, Limin Wang, Gangshan WuICCV 2021 · 被引用 122 次
- Category-Specific Selective Feature Enhancement for Long-Tailed Multi-Label Image ClassificationRuiqi Du, Xu Tang, Xiangrong Zhang, Jingjing MaICCV 2025 · 被引用 1 次
- Improving Calibration for Long-Tailed RecognitionZhisheng Zhong, Jiequan Cui, Shu Liu, Jiaya JiaCVPR 2021
