InfoLossQA: Characterizing and Recovering Information Loss in Text Simplification
Jan Trienes, Sebastian Joseph, Jörg Schlötterer, Christin Seifert, Kyle Lo, Wei Xu, Byron C. Wallace, Junyi Jessy Li
摘要
Text simplification aims to make technical texts more accessible to laypeople but often results in deletion of information and vagueness. This work proposes INFOLOSSQA, a framework to characterize and recover simplificationinduced information loss in form of questionand-answer (QA) pairs. Building on the theory of Questions Under Discussion, the QA pairs are designed to help readers deepen their knowledge of a text. First, we collect a dataset of 1,000 linguist-curated QA pairs derived from 104 LLM simplifications of English medical study abstracts. Our analyses of this data reveal that information loss occurs frequently, and that the QA pairs give a high-level overview of what information was lost. Second, we devise two methods for this task: end-to-end prompting of open-source and commercial language models, and a natural language inference pipeline. With a novel evaluation framework considering the correctness of QA pairs and their linguistic suitability, our expert evaluation reveals that models struggle to reliably identify information loss and applying similar standards as humans at what constitutes information loss. 1 * Work done while visiting UT Austin. 1 Code, dataset and an interactive data viewer is available at https://jantrienes.github.io/ts-info-loss/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- FactPICO: Factuality Evaluation for Plain Language Summarization of Medical EvidenceSebastian Joseph, Lily Chen, Jan Trienes, Hannah Louisa Göke 等ACL 2024 · 被引用 12 次
- Which questions should I answer? Salience Prediction of Inquisitive QuestionsYating Wu, Ritika Mangla, Alex Dimakis, Greg Durrett 等EMNLP 2024 · 被引用 1 次
它引用的顶会 Paper16
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma 等NeurIPS 2022 · 被引用 22,562 次
- Asking and Answering Questions to Evaluate the Factual Consistency of SummariesAlex Wang, Kyunghyun Cho, Mike LewisACL 2020 · 被引用 317 次
- FEQA: A Question Answering Evaluation Framework for Faithfulness Assessment in Abstractive SummarizationEsin Durmus, He He, Mona T. DiabACL 2020 · 被引用 90 次
- Discourse Level Factors for Sentence Deletion in Text SimplificationYang Zhong, Chao Jiang, Wei Xu, Junyi Jessy LiAAAI 2020 · 被引用 57 次
- ERASER: A Benchmark to Evaluate Rationalized NLP ModelsJay DeYoung, Sarthak Jain, Nazneen Fatema Rajani, Eric P. Lehman 等ACL 2020 · 被引用 36 次
相关 Paper
- Elaborative Simplification as Implicit Questions Under DiscussionYating Wu, William Sheffield, Kyle Mahowald, Junyi Jessy LiEMNLP 2023 · 被引用 4 次
- Evaluating Factuality in Text SimplificationAshwin Devaraj, William Sheffield, Byron C. Wallace, Junyi Jessy LiACL 2022
- Evaluating LLMs for Portuguese Sentence Simplification with Linguistic InsightsArthur Mariano Rocha De Azevedo Scalercio, Elvis A. de Souza, Maria José Bocorny Finatto, Aline PaesACL 2025 · 被引用 2 次
- Explainable Prediction of Text Complexity: The Missing Preliminaries for Text SimplificationCristina Garbacea, Mengtian Guo, Samuel Carton, Qiaozhu MeiACL 2021
- PaperTrail: A Claim-Evidence Interface for Grounding Provenance in LLM-based Scholarly Q&AAnna Martin-Boyle, Cara A. C. Leckey, Martha Brown, Harmanpreet KaurCHI 2026 · 被引用 3 次
