LoGU: Long-form Generation with Uncertainty Expressions
Ruihan Yang, Caiqi Zhang, Zhisong Zhang, Xinting Huang, Sen Yang, Nigel Collier, Dong Yu, Deqing Yang
摘要
While Large Language Models (LLMs) demonstrate impressive capabilities, they still struggle with hallucinations. A promising approach to mitigate hallucinations is enabling models to express uncertainty when unsure. Previous research on uncertainty estimation has primarily focused on short-form QA, but real-world applications often require much longer responses. In this work, we introduce the task of Longform Generation with Uncertainty (LoGU), which requires the models to explicitly express uncertainty during the generation. We identify two key challenges: Uncertainty Suppression, where models hesitate to express uncertainty, and Uncertainty Misalignment, where models convey uncertainty inaccurately. To tackle these challenges, we propose a novel decomposition-based data collection framework and a two-stage training pipeline. Specifically, we use supervised fine-tuning (SFT) for uncertainty suppression problem and direct preference optimization (DPO) for uncertainty misalignment. Experiments on three long-form datasets demonstrate the effectiveness of our approach, showing improvements in factual accuracy, reduction of incorrect statements, and preservation of the overall comprehensiveness of the generated responses. Further analysis reveals that baseline methods tend to express uncertainty in vague and broad terms, while our method generates more specific and targeted uncertainty expressions. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- LoVeC: Reinforcement Learning for Better Verbalized Confidence in Long-Form GenerationCaiqi Zhang, Xiaochen Zhu, Chengzu Li, Nigel Collier 等ACL 2026 · 被引用 16 次
- EpiCaR: Knowing What You Don't Know Matters for Better Reasoning in LLMsJe Won Yeom, Jaewon Sok, Seonghyeon Park, Jeongjae Park 等ACL 2026 · 被引用 1 次
- Embedding Trust: Semantic Isotropy Predicts Nonfactuality in Long-Form Text GenerationDhrupad Bhardwaj, Julia Kempe, Tim G. J. RudnerICML 2026
- UNCLE: Benchmarking Uncertainty Expressions in Long-Form GenerationRuihan Yang, Caiqi Zhang, Zhisong Zhang, Xinting Huang 等EMNLP 2025
- GrACE: A Generative Approach to Better Confidence Elicitation and Efficient Test-Time Scaling in Large Language ModelsZhaohan Zhang, Ziquan Liu, Ioannis PatrasACL 2026
它引用的顶会 Paper16
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning 等NeurIPS 2023 · 被引用 10,924 次
- Self-Refine: Iterative Refinement with Self-FeedbackAman Madaan, Niket Tandon, Prakhar Gupta, Skyler Hallinan 等NeurIPS 2023 · 被引用 4,972 次
- Efficient Memory Management for Large Language Model Serving with PagedAttentionWoosuk Kwon, Zhuohan Li, Siyuan Zhuang, Ying Sheng 等SOSP 2023 · 被引用 1,016 次
- Fine-Tuning Language Models for FactualityKatherine Tian, Eric Mitchell, Huaxiu Yao, Christopher D. Manning 等ICLR 2024 · 被引用 270 次
相关 Paper
- I Don't Know: Explicit Modeling of Uncertainty with an [IDK] TokenRoi Cohen, Konstantin Dobler, Eden Biran, Gerard de MeloNeurIPS 2024 · 被引用 35 次
- IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model GenerationHaozhi Fan, Jinhao Duan, Kaidi XuACL 2026 · 被引用 1 次
- SCOPE: A Self-supervised Framework for Improving Faithfulness in Conditional Text GenerationSong Duong, Florian Le Bronnec, Alexandre Allauzen, Vincent Guigue 等ICLR 2025
- A Head to Predict and a Head to Question: Pre-trained Uncertainty Quantification Heads for Hallucination Detection in LLM OutputsArtem Shelmanov, Ekaterina Fadeeva, Akim Tsvigun, Ivan Tsvigun 等EMNLP 2025 · 被引用 3 次
- Probabilities Are All You Need: A Probability-Only Approach to Uncertainty Estimation in Large Language ModelsManh Nguyen, Sunil Gupta, Hung LeAAAI 2026 · 被引用 4 次
