Are Large Language Models Capable of Generating Human-Level Narratives?
Yufei Tian, Tenghao Huang, Miri Liu, Derek Jiang, Alexander Spangher, Muhao Chen, Jonathan May, Nanyun Peng
Abstract
This paper investigates the capability of LLMs in storytelling, focusing on narrative development and plot progression. We introduce a novel computational framework to analyze narratives through three discourse-level aspects: i) story arcs, ii) turning points, and iii) affective dimensions, including arousal and valence. By leveraging expert and automatic annotations, we uncover significant discrepancies between the LLM-and human-written stories. While human-written stories are suspenseful, arousing, and diverse in narrative structures, LLM stories are homogeneously positive and lack tension. Next, we measure narrative reasoning skills as a precursor to generative capacities, concluding that most LLMs fall short of human abilities in discourse understanding. Finally, we show that explicit integration of aforementioned discourse features can enhance storytelling, as is demonstrated by over 40% improvement in neural storytelling in terms of diversity, suspense, and arousal.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 17768bb6-c4e2-4320-a6b5-32821be1a8f3Cited by top-tier papers10
- Can AI writing be salvaged? Mitigating Idiosyncrasies and Improving Human-AI Alignment in the Writing Process through EditsTuhin Chakrabarty, Philippe Laban, Chien-Sheng WuCHI 2025 · 14 citations
- Towards Enhanced Immersion and Agency for LLM-based Interactive DramaHongqiu Wu, Weiqi Wu, Tianyang Xu, Jiameng Zhang et al.ACL 2025 · 7 citations
- Do LLMs Plan Like Human Writers? Comparing Journalist Coverage of Press Releases with LLMsAlexander Spangher, Nanyun Peng, Sebastian Gehrmann, Mark DredzeEMNLP 2024 · 4 citations
- StoryAlign: Evaluating and Training Reward Models for Story GenerationHaotian Xia, Hao Peng, Yunjia Qi, Bin Xu et al.ICLR 2026 · 2 citations
- Narrix: Remixing Narrative Strategies from Examples for Story WritingChao Zhang, Shunan Guo, Abe Davis, Eunyee KohCHI 2026 · 1 citation
Builds on5
- Art or Artifice? Large Language Models and the False Promise of CreativityTuhin Chakrabarty, Philippe Laban, Divyansh Agarwal, Smaranda Muresan et al.CHI 2024 · 122 citations
- Re3: Generating Longer Stories With Recursive Reprompting and RevisionKevin Yang, Yuandong Tian, Nanyun Peng, Dan KleinEMNLP 2022 · 77 citations
- Discourse as a Function of Event: Profiling Discourse Structure in News Articles around the Main EventPrafulla Kumar Choubey, Aaron Lee, Ruihong Huang, Lu WangACL 2020 · 54 citations
- Multitask Semi-Supervised Learning for Class-Imbalanced Discourse ClassificationAlexander Spangher, Jonathan May, Sz-Rung Shiang, Lingjia DengEMNLP 2021 · 16 citations
- Measuring Psychological Depth in Language ModelsFabrice Harel-Canada, Hanyu Zhou, Sreya Muppalla, Zeynep Yildiz et al.EMNLP 2024
Related papers
- LitVISTA: A Benchmark for Narrative Orchestration in Literary TextMingzhe Lu, Yiwen Wang, Yanbing Liu, Qi You et al.ACL 2026
- RolePlot: A Systematic Framework for Evaluating and Enhancing the Plot-Progression Capabilities of Role-Playing AgentsPinyi Zhang, Siyu An, Lingfeng Qiao, Yifei Yu et al.ACL 2025 · 4 citations
- StoryER: Automatic Story Evaluation via Ranking, Rating and ReasoningHong Chen, Duc Minh Vo, Hiroya Takamura, Yusuke Miyao et al.EMNLP 2022 · 3 citations
- Improving Large Language Models in Event Relation Logical PredictionMeiqi Chen, Yubo Ma, Kaitao Song, Yixin Cao et al.ACL 2024 · 7 citations
- Rule or Story, Which is a Better Commonsense Expression for Talking with Large Language Models?Ning Bian, Xianpei Han, Hongyu Lin, Yaojie Lu et al.ACL 2024 · 1 citation
