Automated Lay Language Summarization of Biomedical Scientific Reviews
Yue Guo, Wei Qiu, Yizhong Wang, Trevor Cohen
Abstract
Health literacy has emerged as a crucial factor in making appropriate health decisions and ensuring treatment outcomes. However, medical jargon and the complex structure of professional language in this domain make health information especially hard to interpret. Thus, there is an urgent unmet need for automated methods to enhance the accessibility of the biomedical literature to the general population. This problem can be framed as a type of translation problem between the language of healthcare professionals, and that of the general public. In this paper, we introduce the novel task of automated generation of lay language summaries of biomedical scientific reviews, and construct a dataset to support the development and evaluation of automated methods through which to enhance the accessibility of the biomedical literature. We conduct analyses of the various challenges in performing this task, including not only summarization of the key points but also explanation of background knowledge and simplification of professional language. We experiment with state-of-the-art summarization models as well as several data augmentation techniques, and evaluate their performance using both automated metrics and human assessment. Results indicate that automatically generated summaries produced using contemporary neural architectures can achieve promising quality and readability as compared with reference summaries developed for the lay public by experts (best ROUGE-L of 50.24 and Flesch-Kincaid readability score of 13.30). We also discuss the limitations of the current effort, providing insights and directions for future work.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b2cd2cfa-ae38-4ab7-9a89-079aca5cfa42Cited by top-tier papers10
- Making Science Simple: Corpora for the Lay Summarisation of Scientific LiteratureTomas Goldsack, Zhihao Zhang, Chenghua Lin, Carolina ScartonEMNLP 2022 · 38 citations
- Multilingual Simplification of Medical TextsSebastian Joseph, Kathryn Kazanas, Keziah Reina, Vishnesh J. Ramanathan et al.EMNLP 2023 · 16 citations
- Know Your Audience: The benefits and pitfalls of generating plain language summaries beyond the "general" audienceTal August, Kyle Lo, Noah A. Smith, Katharina ReineckeCHI 2024 · 11 citations
- APPLS: Evaluating Evaluation Metrics for Plain Language SummarizationYue Guo, Tal August, Gondy Leroy, Trevor Cohen et al.EMNLP 2024 · 7 citations
- Generating Summaries with Controllable Readability LevelsLeonardo F. R. Ribeiro, Mohit Bansal, Markus DreyerEMNLP 2023 · 6 citations
Builds on3
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Don't Stop Pretraining: Adapt Language Models to Domains and TasksSuchin Gururangan, Ana Marasovic, Swabha Swayamdipta, Kyle Lo et al.ACL 2020 · 93 citations
- Expertise Style Transfer: A New Task Towards Better Communication between Experts and LaymenYixin Cao, Ruihao Shui, Liangming Pan, Min-Yen Kan et al.ACL 2020 · 50 citations
Related papers
- We Can Explain Your Research in Layman's Terms: Towards Automating Science Journalism at ScaleRumen Dangovski, Michelle Shen, Dawson Byrd, Li Jing et al.AAAI 2021 · 12 citations
- Evaluating Factuality in Text SimplificationAshwin Devaraj, William Sheffield, Byron C. Wallace, Junyi Jessy LiACL 2022
- Evaluating the Evaluators: Are readability metrics good measures of readability?Isabel Cachola, Daniel Khashabi, Mark DredzeEMNLP 2025 · 1 citation
- MedReadMe: A Systematic Study for Fine-grained Sentence Readability in Medical DomainChao Jiang, Wei XuEMNLP 2024 · 3 citations
- Enhancing Biomedical Lay Summarisation with External Knowledge GraphsTomas Goldsack, Zhihao Zhang, Chen Tang, Carolina Scarton et al.EMNLP 2023 · 2 citations
