Do LLMs Plan Like Human Writers? Comparing Journalist Coverage of Press Releases with LLMs
Alexander Spangher, Nanyun Peng, Sebastian Gehrmann, Mark Dredze
摘要
Journalists engage in multiple steps in the news writing process that depend on human creativity, like exploring different "angles" (i.e. the specific perspectives a reporter takes). These can potentially be aided by large language models (LLMs). By affecting planning decisions, such interventions can have an outsize impact on creative output. We advocate a careful approach to evaluating these interventions to ensure alignment with human values. In a case study of journalistic coverage of press releases, we assemble a large dataset of 250k press releases 1 and 650k articles covering them. 2 We develop methods to identify news articles that challenge and contextualize press releases. Finally, we evaluate suggestions made by LLMs for these articles and compare these with decisions made by human journalists. Our findings are three-fold: (1) Human-written news articles that challenge and contextualize press releases more take more creative angles and use more informational sources. (2) LLMs align better with humans when recommending angles, compared with informational sources. (3) Both the angles and sources LLMs suggest are significantly less creative than humans.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Critical Confabulation: Can LLMs Hallucinate for Social Good?Peiqi Sui, Eamon Duede, Hoyt Long, Richard Jean SoICLR 2026 · 被引用 2 次
- AI as Humanity's Salieri: Quantifying Linguistic Creativity of Language Models via Systematic Attribution of Machine Text against Web TextXiming Lu, Melanie Sclar, Skyler Hallinan, Niloofar Mireshghallah 等ICLR 2025
- Measuring Psychological Depth in Language ModelsFabrice Harel-Canada, Hanyu Zhou, Sreya Muppalla, Zeynep Yildiz 等EMNLP 2024
- CoT is Not the Chain of Truth: An Empirical Internal Analysis of Reasoning LLMs for Fake News GenerationZhao Tong, Chunlin Gong, Yiping Zhang, Haichao Shi 等ICML 2026
它引用的顶会 Paper13
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma 等NeurIPS 2022 · 被引用 22,562 次
- Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksPatrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni 等NeurIPS 2020 · 被引用 19,162 次
- Toolformer: Language Models Can Teach Themselves to Use ToolsTimo Schick, Jane Dwivedi-Yu, Roberto Dessì, Roberta Raileanu 等NeurIPS 2023 · 被引用 5,989 次
- Adversarial NLI: A New Benchmark for Natural Language UnderstandingYixin Nie, Adina Williams, Emily Dinan, Mohit Bansal 等ACL 2020 · 被引用 602 次
- Co-Writing Screenplays and Theatre Scripts with Language Models: Evaluation by Industry ProfessionalsPiotr Mirowski, Kory W. Mathewson, Jaylen Pittman, Richard EvansCHI 2023 · 被引用 235 次
相关 Paper
- AngleKindling: Supporting Journalistic Angle Ideation with Large Language ModelsSavvas Petridis, Nicholas Diakopoulos, Kevin Crowston, Mark Hansen 等CHI 2023 · 被引用 94 次
- Beyond Accuracy: Experts See AI Fact-Checks as Accurate but Less UsefulChenyan Jia, Apoorva Gondimalla, Angie Zhang, David Joseph Mullings 等CHI 2026 · 被引用 1 次
- NewsInterview: a Dataset and a Playground to Evaluate LLMs' Grounding Gap via Informational InterviewsAlexander Spangher, Michael Lu, Sriya Kalyan, Hyundong Justin Cho 等ACL 2025 · 被引用 2 次
- LLM or Human? Perceptions of Trust and Quality in Research SummariesNil-Jana Akpinar, Sandeep Avula, Chia-Jung Lee, Brandon Dang 等CHI 2026 · 被引用 2 次
- Towards Designing a Question-Answering Chatbot for Online News: Understanding Questions and PerspectivesMd. Naimul Hoque, Ayman A. Mahfuz, Mayukha Sridhatri Kindi, Naeemul HassanCHI 2024 · 被引用 8 次
