ParseCaps: An Interpretable Parsing Capsule Network for Medical Image Diagnosis
Xinyu Geng, Jiaming Wang, Xiaolin Huang, Fanglin Chen, Jun Xu
Abstract
Deep learning has excelled in medical image classification, but its clinical application is limited by poor interpretability. Capsule networks, known for encoding hierarchical relationships and spatial features, show potential in addressing this issue. Nevertheless, traditional capsule networks often underperform due to their shallow structures, and deeper variants lack hierarchical architectures, thereby compromising interpretability. This paper introduces a novel capsule network, ParseCaps, which utilizes the sparse axial attention routing and parse convolutional capsule layer to form a parsetree-like structure, enhancing both depth and interpretability. Firstly, sparse axial attention routing optimizes connections between child and parent capsules, as well as emphasizes the weight distribution across instantiation parameters of parent capsules. Secondly, the parse convolutional capsule layer generates capsule predictions aligning with the parse tree. Finally, based on the loss design that is effective whether concept ground truth exists or not, ParseCaps advances interpretability by associating each dimension of the global capsule with a comprehensible concept, thereby facilitating clinician trust and understanding of the model's classification results. Experimental results on CE-MRI, PH 2 , and Derm7pt datasets show that ParseCaps not only outperforms other capsule network variants in classification accuracy, redundancy reduction and robustness, but also provides interpretable explanations, regardless of the availability of concept labels.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c2133465-01f9-49f1-bbf9-b874891a8ca5Cited by top-tier papers1
Ask how each one uses itBuilds on5
- TreeCaps: Tree-Based Capsule Networks for Source Code ProcessingNghi D. Q. Bui, Yijun Yu, Lingxiao JiangAAAI 2021 · 44 citations
- A Framework for Learning Ante-hoc Explainable Models via ConceptsAnirban Sarkar, Deepak Vijaykeerthy, Anindya Sarkar, Vineeth N. BalasubramanianCVPR 2022 · 40 citations
- HP-Capsule: Unsupervised Face Part Discovery by Hierarchical Parsing Capsule NetworkChang Yu, Xiangyu Zhu, Xiaomei Zhang, Zidu Wang et al.CVPR 2022 · 18 citations
- OrthCaps: An Orthogonal CapsNet with Sparse Attention Routing and PruningXinyu Geng, Jiaming Wang, Jiawei Gong, Yuerong Xue et al.CVPR 2024
- XProtoNet: Diagnosis in Chest Radiography With Global and Local ExplanationsEunji Kim, Siwon Kim, Minji Seo, Sungroh YoonCVPR 2021
Related papers
- Why Capsule Neural Networks Do Not Scale: Challenging the Dynamic Parse-Tree AssumptionMatthias Mitterreiter, Marcel Koch, Joachim Giesen, Sören LaueAAAI 2023 · 17 citations
- Adaptive Activation Thresholding: Dynamic Routing Type Behavior for Interpretability in Convolutional Neural NetworksYiyou Sun, Sathya N. Ravi, Vikas SinghICCV 2019 · 17 citations
- Interpretable Graph Capsule Networks for Object RecognitionJindong GuAAAI 2021 · 42 citations
- Linguistically Routing Capsule Network for Out-of-distribution Visual Question AnsweringQingxing Cao, Wentao Wan, Keze Wang, Xiaodan Liang et al.ICCV 2021 · 16 citations
- Towards Multi-dimensional Explanation Alignment for Medical ClassificationLijie Hu, Songning Lai, Wenshuo Chen, Hongru Xiao et al.NeurIPS 2024 · 8 citations
