An Unsupervised Approach to Achieve Supervised-Level Explainability in Healthcare Records
Joakim Edin, Maria Maistro, Lars Maaløe, Lasse Borgholt, Jakob D. Havtorn, Tuukka Ruotsalo
摘要
Electronic healthcare records are vital for patient safety as they document conditions, plans, and procedures in both free text and medical codes. Language models have significantly enhanced the processing of such records, streamlining workflows and reducing manual data entry, thereby saving healthcare providers significant resources. However, the black-box nature of these models often leaves healthcare professionals hesitant to trust them. State-ofthe-art explainability methods increase model transparency but rely on human-annotated evidence spans, which are costly. In this study, we propose an approach to produce plausible and faithful explanations without needing such annotations. We demonstrate on the automated medical coding task that adversarial robustness training improves explanation plausibility and introduce AttInGrad, a new explanation method superior to previous ones. By combining both contributions in a fully unsupervised setup, we produce explanations of comparable quality, or better, to that of a supervised approach. We release our code and model weights. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Towards Explainable Diagnosis: A Self-learned Explanatory Knowledge Base ApproachDongqi Huang, Tong Zhou, Zhuoran Jin, Shenghui Shi 等ACL 2026
- Less is More: Explainable and Efficient ICD Code Prediction with Clinical EntitiesJames C. Douglas, Yidong Gan, Ben Hachey, Jonathan K. KummerfeldACL 2025
- ICDAGENT: Empowering Agentic Large Language Models for Explainable Medical CodingZiyi Yin, Yuanpu Cao, Ting Wang, Jinghui Chen 等ACL 2026
它引用的顶会 Paper6
- ERASER: A Benchmark to Evaluate Rationalized NLP ModelsJay DeYoung, Sarthak Jain, Nazneen Fatema Rajani, Eric P. Lehman 等ACL 2020 · 被引用 36 次
- Incorporating Residual and Normalization Layers into Analysis of Masked Language ModelsGoro Kobayashi, Tatsuki Kuribayashi, Sho Yokoi, Kentaro InuiEMNLP 2021 · 被引用 28 次
- Discriminative Feature Attributions: Bridging Post Hoc Explainability and Inherent InterpretabilityUsha Bhalla, Suraj Srinivas, Himabindu LakkarajuNeurIPS 2023 · 被引用 18 次
- Ignorance is Bliss: Robust Control via Information GatingManan Tomar, Riashat Islam, Matthew E. Taylor, Sergey Levine 等NeurIPS 2023 · 被引用 14 次
- MDACE: MIMIC Documents Annotated with Code EvidenceHua Cheng, Rana Jafari, April Russell, Russell Klopfer 等ACL 2023 · 被引用 11 次
相关 Paper
- Beyond Label Attention: Transparency in Language Models for Automated Medical Coding via Dictionary LearningJohn Wu, David Wu, Jimeng SunEMNLP 2024 · 被引用 4 次
- Faithful Serum: Mitigating the Faithfulness Gap in Textual Explanations of LLM Decisions via Attribution GuidanceBar Alon, Itamar Zimerman, Lior WolfACL 2026
- Explaining Black-Box Language Models: Learning to Optimize Linguistically-Structured Word SubsetsMinyoung Hwang, Seokhyun Lee, Changhee LeeKDD 2026
- Exploring Accurate and Transparent Domain Adaptation in Predictive Healthcare via Concept-Grounded Orthogonal InferencePengfei Hu, Chang Lu, Feifan Liu, Yue NingICML 2026 · 被引用 1 次
- Robust Explanation for Free or At the Cost of FaithfulnessZeren Tan, Yang TianICML 2023 · 被引用 12 次
