Critical Confabulation: Can LLMs Hallucinate for Social Good?
Peiqi Sui, Eamon Duede, Hoyt Long, Richard Jean So
摘要
LLMs hallucinate, yet some confabulations can have social affordances if carefully bounded. We propose critical confabulation (inspired by critical fabulation from literary and social theory), the use of LLM hallucinations to "fill-in-the-gap" for omissions in archives due to social and political inequality, and reconstruct divergent yet evidence-bound narratives for history's "hidden figures". We simulate these gaps with an open-ended narrative cloze task: asking LLMs to generate a masked event in a character-centric timeline sourced from a novel corpus of unpublished texts. We evaluate audited (for data contamination), fully-open models (the OLMO-2 family) and unaudited open-weight and proprietary baselines under a range of prompts designed to elicit controlled and useful hallucinations. Our findings validate LLMs' foundational narrative understanding capabilities to perform critical confabulation, and show how controlled and well-specified hallucinations can support LLM applications for knowledge production without collapsing speculation into a lack of historical accuracy and fidelity. "As an emblematic figure of the enslaved woman in the Atlantic world, Venus makes plain the convergence of terror and pleasure in the libidinal economy of slavery. . . [critical fabulation] attempts to redress it by describing as fully as possible the conditions that determine the appearance of Venus and that dictate her silence." -Saidiya Hartman, (2008) 1 BACKGROUND Large language models (LLMs) are prone to hallucinate, generating "plausible yet nonfactual" outputs (Huang et al., 2025). While typically treated as a failure mode, recent studies show that some hallucinations could in fact be valuable (Jiang et al., 2024; Hu et al., 2024; Taveekitworachai et al., 2024) , especially a subset termed confabulations: a narrative-driven tendency to "fill in" missing information with self-consistent stories that bear close verisimilitude with reality (Sui et al., 2024) . Confabulated texts are more narrative-rich and better match the patterns of human storytelling, a vital communicative and cognitive resource for sense-making (Herman, 2013). Thus, a principled
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper21
- The Curious Case of Neural Text DegenerationAri Holtzman, Jan Buys, Li Du, Maxwell Forbes 等ICLR 2020 · 被引用 4,112 次
- Detecting Pretraining Data from Large Language ModelsWeijia Shi, Anirudh Ajith, Mengzhou Xia, Yangsibo Huang 等ICLR 2024 · 被引用 365 次
- HaluEval: A Large-Scale Hallucination Evaluation Benchmark for Large Language ModelsJunyi Li, Xiaoxue Cheng, Xin Zhao, Jian-Yun Nie 等EMNLP 2023 · 被引用 224 次
- Proving Test Set Contamination in Black-Box Language ModelsYonatan Oren, Nicole Meister, Niladri S. Chatterji, Faisal Ladhak 等ICLR 2024 · 被引用 220 次
- Time Travel in LLMs: Tracing Data Contamination in Large Language ModelsShahriar Golchin, Mihai SurdeanuICLR 2024 · 被引用 165 次
相关 Paper
- Confabulation: The Surprising Value of Large Language Model HallucinationsPeiqi Sui, Eamon Duede, Sophie Wu, Richard Jean SoACL 2024
- Where Confabulation Lives: Latent Feature Discovery in LLMsThibaud Ardoin, Yi Cai, Gerhard WunderEMNLP 2025 · 被引用 1 次
- Beyond Noise: Characterizing Creative Potential in Unverifiable LLM HallucinationsYu Yan, Chunhong Zhang, Haiyu Zhao, Ziyang Zeng 等ACL 2026
- Past Meets Present: Creating Historical Analogy with Large Language ModelsNianqi Li, Siyu Yuan, Jiangjie Chen, Jiaqing Liang 等ACL 2025
- Addressing Pitfalls in the Evaluation of Uncertainty Estimation Methods for Natural Language GenerationMykyta Ielanskyi, Kajetan Schweighofer, Lukas Aichberger, Sepp HochreiterICLR 2026 · 被引用 10 次
