Semi-Supervised Exaggeration Detection of Health Science Press Releases
Dustin Wright, Isabelle Augenstein
摘要
Public trust in science depends on honest and factual communication of scientific papers. However, recent studies have demonstrated a tendency of news media to misrepresent scientific papers by exaggerating their findings. Given this, we present a formalization of and study into the problem of exaggeration detection in science communication. While there are an abundance of scientific papers and popular media articles written about them, very rarely do the articles include a direct link to the original paper, making data collection challenging. We address this by curating a set of labeled press release/abstract pairs from existing expert annotated studies on exaggeration in press releases of scientific papers suitable for benchmarking the performance of machine learning models on the task. Using limited data from this and previous studies on exaggeration detection in science, we introduce MT-PET, a multi-task version of Pattern Exploiting Training (PET), which leverages knowledge from complementary clozestyle QA tasks to improve few-shot learning. We demonstrate that MT-PET outperforms PET and supervised learning both when data is limited, as well as when there is an abundance of data for the main task. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper4
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Inducing Relational Knowledge from BERTZied Bouraoui, José Camacho-Collados, Steven SchockaertAAAI 2020 · 被引用 183 次
- We Can Explain Your Research in Layman's Terms: Towards Automating Science Journalism at ScaleRumen Dangovski, Michelle Shen, Dawson Byrd, Li Jing 等AAAI 2021 · 被引用 12 次
- Fact or Fiction: Verifying Scientific ClaimsDavid Wadden, Shanchuan Lin, Kyle Lo, Lucy Lu Wang 等EMNLP 2020 · 被引用 6 次
相关 Paper
- Few-Shot Text Generation with Natural Language InstructionsTimo Schick, Hinrich SchützeEMNLP 2021 · 被引用 101 次
- 'Don't Get Too Technical with Me': A Discourse Structure-Based Framework for Automatic Science JournalismRonald Cardenas, Bingsheng Yao, Dakuo Wang, Yufang HouEMNLP 2023 · 被引用 2 次
- MetaAdapt: Domain Adaptive Few-Shot Misinformation Detection via Meta LearningZhenrui Yue, Huimin Zeng, Yang Zhang, Lanyu Shang 等ACL 2023 · 被引用 23 次
- Don't Stop Pretraining: Adapt Language Models to Domains and TasksSuchin Gururangan, Ana Marasovic, Swabha Swayamdipta, Kyle Lo 等ACL 2020 · 被引用 93 次
- PASTA: Table-Operations Aware Fact Verification via Sentence-Table Cloze Pre-trainingZihui Gu, Ju Fan, Nan Tang, Preslav Nakov 等EMNLP 2022 · 被引用 19 次
