Aligning Language Models to Explicitly Handle Ambiguity
Hyuhng Joon Kim, Youna Kim, Cheonbok Park, Junyeob Kim, Choonghyun Park, Kang Min Yoo, Sang-goo Lee, Taeuk Kim
摘要
In interactions between users and language model agents, user utterances frequently exhibit ellipsis (omission of words or phrases) or imprecision (lack of exactness) to prioritize efficiency. This can lead to varying interpretations of the same input based on different assumptions or background knowledge. It is thus crucial for agents to adeptly handle the inherent ambiguity in queries to ensure reliability. However, even state-of-the-art large language models (LLMs) still face challenges in such scenarios, primarily due to the following hurdles: (1) LLMs are not explicitly trained to deal with ambiguous utterances; (2) the degree of ambiguity perceived by the LLMs may vary depending on the possessed knowledge. To address these issues, we propose Alignment with Perceived Ambiguity (APA), a novel pipeline that aligns LLMs to manage ambiguous queries by leveraging their own assessment of ambiguity (i.e., perceived ambiguity). Experimental results on question-answering datasets demonstrate that APA empowers LLMs to explicitly detect and manage ambiguous queries while retaining the ability to answer clear questions. Furthermore, our finding proves that APA excels beyond training with gold-standard labels, especially in out-of-distribution scenarios. The data and code are available at https://github.com/heyjoonkim/APA .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- Ambig-SWE: Interactive Agents to Overcome Underspecificity in Software EngineeringSanidhya Vijayvargiya, Xuhui Zhou, Akhila Yerukola, Maarten Sap 等ICLR 2026 · 被引用 35 次
- Knowledge Boundary of Large Language Models: A SurveyMoxin Li, Yong Zhao, Wenxuan Zhang, Shuaiyi Li 等ACL 2025 · 被引用 33 次
- HonestLLM: Toward an Honest and Helpful Large Language ModelChujie Gao, Siyuan Wu, Yue Huang, Dongping Chen 等NeurIPS 2024 · 被引用 30 次
- A Comparative Analysis of Information Gathering by Chatbots, Questionnaires, and Humans in Clinical Pre-ConsultationBrenna Li, Saba Tauseef, Khai N. Truong, Alex MariakakisCHI 2025 · 被引用 16 次
- Asking What Matters: Reward-Driven Clarification for Software Engineering TasksSanidhya Vijayvargiya, Vijay Viswanathan, Graham NeubigICML 2026 · 被引用 3 次
它引用的顶会 Paper14
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning 等NeurIPS 2023 · 被引用 10,924 次
- QLoRA: Efficient Finetuning of Quantized LLMsTim Dettmers, Artidoro Pagnoni, Ari Holtzman, Luke ZettlemoyerNeurIPS 2023 · 被引用 5,863 次
- LIMA: Less Is More for AlignmentChunting Zhou, Pengfei Liu, Puxin Xu, Srinivasan Iyer 等NeurIPS 2023 · 被引用 1,486 次
- WizardLM: Empowering Large Pre-Trained Language Models to Follow Complex InstructionsCan Xu, Qingfeng Sun, Kai Zheng, Xiubo Geng 等ICLR 2024 · 被引用 1,206 次
- LESS: Selecting Influential Data for Targeted Instruction TuningMengzhou Xia, Sadhika Malladi, Suchin Gururangan, Sanjeev Arora 等ICML 2024 · 被引用 460 次
相关 Paper
- UAlign: Leveraging Uncertainty Estimations for Factuality Alignment on Large Language ModelsBoyang Xue, Fei Mi, Qi Zhu, Hongru Wang 等ACL 2025 · 被引用 10 次
- Disambiguation in Conversational Question Answering in the Era of LLMs and Agents: A SurveyMd. Mehrab Tanjim, Yeonjun In, Xiang Chen, Victor S. Bursztyn 等EMNLP 2025 · 被引用 2 次
- Question Answering as Programming for Solving Time-Sensitive QuestionsXinyu Zhu, Cheng Yang, Bei Chen, Siheng Li 等EMNLP 2023 · 被引用 6 次
- Disentangling Memory and Reasoning Ability in Large Language ModelsMingyu Jin, Weidi Luo, Sitao Cheng, Xinyi Wang 等ACL 2025
- Beyond "I Don't Know": Evaluating LLM Self-Awareness in Discriminating Data and Model UncertaintyJingyi Ren, Ante Wang, Yunghwei Lai, Xiaolong Wang 等ACL 2026 · 被引用 1 次
