Evaluating LLMs for Targeted Concept Simplification for Domain-Specific Texts
Sumit Asthana, Hannah Rashkin, Elizabeth Clark, Fantine Huot, Mirella Lapata
Abstract
One useful application of NLP models is to support people in reading complex text from unfamiliar domains (e.g., scientific articles). Simplifying the entire text makes it understandable but sometimes removes important details. On the contrary, helping adult readers understand difficult concepts in context can enhance their vocabulary and knowledge. In a preliminary human study, we first identify that lack of context and unfamiliarity with difficult concepts is a major reason for adult readers' difficulty with domain-specific text. We then introduce targeted concept simplification, a simplification task for rewriting text to help readers comprehend text containing unfamiliar concepts. We also introduce WIKIDOMAINS 1 , a new dataset of 22k definitions from 13 academic domains paired with a difficult concept within each definition. We benchmark the performance of opensource and commercial LLMs, and a simple dictionary baseline on this task across human judgments of ease of understanding and meaning preservation. Interestingly, our human judges preferred explanations about the difficult concept more than simplification of the concept phrase. Further, no single model achieved superior performance across all quality dimensions, and automated metrics also show low correlations with human evaluations of concept simplification (∼ 0.2), opening up rich avenues for research on personalized human reading comprehension support. * Work done as student researcher at Google DeepMind. 1 https://github.com/google-deepmind/wikidomains Domain:CS Definition: In computer science, arbitrary-precision arithmetic indicates that calculations are performed on numbers whose digits of precision are limited only by the available memory of the host system.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f6ea28ac-4934-4cee-9b19-8ec0e7a3c188Cited by top-tier papers1
Ask how each one uses itBuilds on8
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger et al.ICLR 2020 · 8,443 citations
- Why Johnny Can't Prompt: How Non-AI Experts Try (and Fail) to Design LLM PromptsJ. D. Zamfirescu-Pereira, Richmond Y. Wong, Bjoern Hartmann, Qian YangCHI 2023 · 892 citations
- Multilingual Simplification of Medical TextsSebastian Joseph, Kathryn Kazanas, Keziah Reina, Vishnesh J. Ramanathan et al.EMNLP 2023 · 16 citations
- BLESS: Benchmarking Large Language Models on Sentence SimplificationTannon Kew, Alison Chi, Laura Vásquez-Rodríguez, Sweta Agrawal et al.EMNLP 2023 · 15 citations
- To Test Machine Comprehension, Start by Defining ComprehensionJesse Dunietz, Gregory Burnham, Akash Bharadwaj, Owen Rambow et al.ACL 2020 · 7 citations
Related papers
- Generating Scientific Definitions with Controllable ComplexityTal August, Katharina Reinecke, Noah A. SmithACL 2022
- Expertise Style Transfer: A New Task Towards Better Communication between Experts and LaymenYixin Cao, Ruihao Shui, Liangming Pan, Min-Yen Kan et al.ACL 2020 · 50 citations
- On the Automatic Generation and Simplification of Children's StoriesMaria R. Valentini, Jennifer Weber, Jesus Salcido, Téa Wright et al.EMNLP 2023 · 7 citations
- Know Your Audience: The benefits and pitfalls of generating plain language summaries beyond the "general" audienceTal August, Kyle Lo, Noah A. Smith, Katharina ReineckeCHI 2024 · 11 citations
- Explainable Prediction of Text Complexity: The Missing Preliminaries for Text SimplificationCristina Garbacea, Mengtian Guo, Samuel Carton, Qiaozhu MeiACL 2021
