Are Large Language Models Capable of Generating Human-Level Narratives?
Yufei Tian, Tenghao Huang, Miri Liu, Derek Jiang, Alexander Spangher, Muhao Chen, Jonathan May, Nanyun Peng
摘要
This paper investigates the capability of LLMs in storytelling, focusing on narrative development and plot progression. We introduce a novel computational framework to analyze narratives through three discourse-level aspects: i) story arcs, ii) turning points, and iii) affective dimensions, including arousal and valence. By leveraging expert and automatic annotations, we uncover significant discrepancies between the LLM-and human-written stories. While human-written stories are suspenseful, arousing, and diverse in narrative structures, LLM stories are homogeneously positive and lack tension. Next, we measure narrative reasoning skills as a precursor to generative capacities, concluding that most LLMs fall short of human abilities in discourse understanding. Finally, we show that explicit integration of aforementioned discourse features can enhance storytelling, as is demonstrated by over 40% improvement in neural storytelling in terms of diversity, suspense, and arousal.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Can AI writing be salvaged? Mitigating Idiosyncrasies and Improving Human-AI Alignment in the Writing Process through EditsTuhin Chakrabarty, Philippe Laban, Chien-Sheng WuCHI 2025 · 被引用 14 次
- Towards Enhanced Immersion and Agency for LLM-based Interactive DramaHongqiu Wu, Weiqi Wu, Tianyang Xu, Jiameng Zhang 等ACL 2025 · 被引用 7 次
- Do LLMs Plan Like Human Writers? Comparing Journalist Coverage of Press Releases with LLMsAlexander Spangher, Nanyun Peng, Sebastian Gehrmann, Mark DredzeEMNLP 2024 · 被引用 4 次
- StoryAlign: Evaluating and Training Reward Models for Story GenerationHaotian Xia, Hao Peng, Yunjia Qi, Bin Xu 等ICLR 2026 · 被引用 2 次
- Narrix: Remixing Narrative Strategies from Examples for Story WritingChao Zhang, Shunan Guo, Abe Davis, Eunyee KohCHI 2026 · 被引用 1 次
它引用的顶会 Paper5
- Art or Artifice? Large Language Models and the False Promise of CreativityTuhin Chakrabarty, Philippe Laban, Divyansh Agarwal, Smaranda Muresan 等CHI 2024 · 被引用 122 次
- Re3: Generating Longer Stories With Recursive Reprompting and RevisionKevin Yang, Yuandong Tian, Nanyun Peng, Dan KleinEMNLP 2022 · 被引用 77 次
- Discourse as a Function of Event: Profiling Discourse Structure in News Articles around the Main EventPrafulla Kumar Choubey, Aaron Lee, Ruihong Huang, Lu WangACL 2020 · 被引用 54 次
- Multitask Semi-Supervised Learning for Class-Imbalanced Discourse ClassificationAlexander Spangher, Jonathan May, Sz-Rung Shiang, Lingjia DengEMNLP 2021 · 被引用 16 次
- Measuring Psychological Depth in Language ModelsFabrice Harel-Canada, Hanyu Zhou, Sreya Muppalla, Zeynep Yildiz 等EMNLP 2024
相关 Paper
- LitVISTA: A Benchmark for Narrative Orchestration in Literary TextMingzhe Lu, Yiwen Wang, Yanbing Liu, Qi You 等ACL 2026
- RolePlot: A Systematic Framework for Evaluating and Enhancing the Plot-Progression Capabilities of Role-Playing AgentsPinyi Zhang, Siyu An, Lingfeng Qiao, Yifei Yu 等ACL 2025 · 被引用 4 次
- StoryER: Automatic Story Evaluation via Ranking, Rating and ReasoningHong Chen, Duc Minh Vo, Hiroya Takamura, Yusuke Miyao 等EMNLP 2022 · 被引用 3 次
- Improving Large Language Models in Event Relation Logical PredictionMeiqi Chen, Yubo Ma, Kaitao Song, Yixin Cao 等ACL 2024 · 被引用 7 次
- Rule or Story, Which is a Better Commonsense Expression for Talking with Large Language Models?Ning Bian, Xianpei Han, Hongyu Lin, Yaojie Lu 等ACL 2024 · 被引用 1 次
