Looking at Radiology Report Generation through a Causal Lens: A Survey
Satyam Kumar, Kaustubh Shivshankar Shejole, Pushpak Bhattacharyya
Abstract
Automatic radiology report generation (RRG) has emerged as a promising approach to reduce clinicians' workload, yet existing systems are vulnerable to biases induced by spurious correlations across data, models, and evaluation pipelines. Such biases raise serious fairness concerns and may adversely affect patient care, making their mitigation critical in clinical settings. Leveraging causal inference to identify true cause-effect relationships can mitigate many biases and yield fair, reliable systems with clinically meaningful outputs. Existing surveys on RRG primarily emphasize deep learning approaches while overlooking the critical role of causality. This survey addresses this gap by analyzing bias across the RRG pipeline, formalizing RRG as a causal modeling problem, and reviewing representative causal techniques from the literature. Based on the level of intervention, we organize existing mitigation strategies into a three-tier taxonomy. We further examine commonly used public medical imaging datasets and evaluation metrics through a causal lens, revealing their biases and limitations in capturing causal alignment and clinical fidelity. To address these limitations, we advocate broader demographic coverage and causal-aware evaluation metrics to improve fairness and reliability, and identify important directions for future work.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5250c2fc-def8-48e0-a47c-8141f7dee1bdBuilds on16
- In Search of Lost Domain GeneralizationIshaan Gulrajani, David Lopez-PazICLR 2021 · 1,416 citations
- Generating Radiology Reports via Memory-driven TransformerZhihong Chen, Yan Song, Tsung-Hui Chang, Xiang WanEMNLP 2020 · 552 citations
- KG-BART: Knowledge Graph-Augmented BART for Generative Commonsense ReasoningYe Liu, Yao Wan, Lifang He, Hao Peng et al.AAAI 2021 · 220 citations
- The Power of Scale for Parameter-Efficient Prompt TuningBrian Lester, Rami Al-Rfou, Noah ConstantEMNLP 2021 · 94 citations
- Counterfactual Contrastive Learning for Weakly-Supervised Vision-Language GroundingZhu Zhang, Zhou Zhao, Zhijie Lin, Jieming Zhu et al.NeurIPS 2020 · 74 citations
Related papers
- Rethinking Fair Representation Learning for Performance-Sensitive TasksCharles Jones, Fabio De Sousa Ribeiro, Mélanie Roschewitz, Daniel C. Castro et al.ICLR 2025
- MEDFAIR: Benchmarking Fairness for Medical ImagingYongshuo Zong, Yongxin Yang, Timothy M. HospedalesICLR 2023 · 15 citations
- The Boundaries of Fair AI in Medical Image Prognosis: A Causal PerspectiveThai-Hoang Pham, Jiayuan Chen, Seungyeon Lee, Yuanlong Wang et al.NeurIPS 2025 · 3 citations
- Towards Robust Classification Model by Counterfactual and Invariant Data GenerationChun-Hao Chang, George-Alexandru Adam, Anna GoldenbergCVPR 2021
- Phrase-grounded APO for Improving Chest X-ray Report GenerationRazi Mahmood, Tanveer F. Syeda-MahmoodCVPR 2026
