Towards Trustable Skin Cancer Diagnosis via Rewriting Model's Decision
Siyuan Yan, Zhen Yu, Xuelin Zhang, Dwarikanath Mahapatra, Shekhar S. Chandra, Monika Janda, H. Peter Soyer, Zongyuan Ge
Abstract
Deep neural networks have demonstrated promising performance on image recognition tasks. However, they may heavily rely on confounding factors, using irrelevant artifacts or bias within the dataset as the cue to improve performance. When a model performs decision-making based on these spurious correlations, it can become untrustable and lead to catastrophic outcomes when deployed in the real-world scene. In this paper, we explore and try to solve this problem in the context of skin cancer diagnosis. We introduce a human-in-the-loop framework in the model training process such that users can observe and correct the model's decision logic when confounding behaviors happen. Specifically, our method can automatically discover confounding factors by analyzing the co-occurrence behavior of the samples. It is capable of learning confounding concepts using easily obtained concept exemplars. By mapping the blackbox model's feature representation onto an explainable concept space, human users can interpret the concept and intervene via first order-logic instruction. We systematically evaluate our method on our newly crafted, well-controlled skin lesion dataset and several public skin lesion datasets. Experiments show that our method can effectively detect and remove confounding factors from datasets without any prior knowledge about the category distribution and does not require fully annotated concept labels. We also show that our method enables the model to focus on clinicalrelated concepts, improving the model's performance and trustworthiness during model inference.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext cff8fbb7-8cd9-4ba2-b64b-381fc34e7afaCited by top-tier papers7
- MICA: Towards Explainable Skin Lesion Diagnosis via Multi-Level Image-Concept AlignmentYequan Bie, Luyang Luo, Hao ChenAAAI 2024 · 28 citations
- Derm1M: A Million-Scale Vision-Language Dataset Aligned with Clinical Ontology Knowledge for DermatologySiyuan Yan, Ming Hu, Yiwen Jiang, Xieji Li et al.ICCV 2025 · 7 citations
- Vision-Language Models Guided Graph Concept Reasoning for Interpretable Diabetic Retinopathy DiagnosisQihao Xu, Xiaoling Luo, Yuxin Lin, Chengliang Liu et al.AAAI 2026
- Enhancing Interpretable Image Classification Through LLM Agents and Conditional Concept Bottleneck ModelsYiwen Jiang, Deval Mehta, Wei Feng, Zongyuan GeACL 2025
- WISE: Weak-Supervision-Guided Step-by-Step Explanations for Multimodal LLMs in Image ClassificationYiwen Jiang, Deval Mehta, Siyuan Yan, Yaling Shen et al.EMNLP 2025
Builds on9
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Concept Bottleneck ModelsPang Wei Koh, Thao Nguyen, Yew Siang Tang, Stephen Mussmann et al.ICML 2020 · 1,233 citations
- Interpretations are Useful: Penalizing Explanations to Align Neural Networks with Prior KnowledgeLaura Rieger, Chandan Singh, W. James Murdoch, Bin YuICML 2020 · 249 citations
- Entropy-Based Logic Explanations of Neural NetworksPietro Barbiero, Gabriele Ciravegna, Francesco Giannini, Pietro Lió et al.AAAI 2022 · 97 citations
- Attention-based Interpretability with Concept TransformersMattia Rigotti, Christoph Miksovic, Ioana Giurgiu, Thomas Gschwind et al.ICLR 2022 · 77 citations
Related papers
- Discover and Cure: Concept-aware Mitigation of Spurious CorrelationShirley Wu, Mert Yüksekgönül, Linjun Zhang, James ZouICML 2023 · 97 citations
- DARC: Dual Adjustment Reasoning with Counterfactuals for Trustworthy Chest X-ray ClassificationZhifang Liao, Junhao Li, Haokang Ding, Yucheng SongCVPR 2026
- Towards Ultrasound-based Reliable Disease Diagnosis Using Causal InferenceBolei Chen, Jiaxu Kang, Haonan Yang, Ping Zhong et al.AAAI 2026
- Reliable and Trustworthy Machine Learning for Health Using Dataset Shift DetectionChunjong Park, Anas Awadalla, Tadayoshi Kohno, Shwetak N. PatelNeurIPS 2021 · 50 citations
- Skin Deep Unlearning: Artefact and Instrument Debiasing in the Context of Melanoma ClassificationPeter J. Bevan, Amir Atapour-AbarghoueiICML 2022 · 24 citations
