Generating Clarifying Questions for Information Retrieval
Hamed Zamani, Susan T. Dumais, Nick Craswell, Paul N. Bennett, Gord Lueck
Abstract
Search queries are often short, and the underlying user intent may be ambiguous. This makes it challenging for search engines to predict possible intents, only one of which may pertain to the current user. To address this issue, search engines often diversify the result list and present documents relevant to multiple intents of the query. An alternative approach is to ask the user a question to clarify her information need. Asking clarifying questions is particularly important for scenarios with “limited bandwidth” interfaces, such as speech-only and small-screen devices. In addition, our user studies and large-scale online experiments show that asking clarifying questions is also useful in web search. Although some recent studies have pointed out the importance of asking clarifying questions, generating them for open-domain search tasks remains unstudied and is the focus of this paper. Lack of training data even within major search engines for this task makes it challenging. To mitigate this issue, we first identify a taxonomy of clarification for open-domain search queries by analyzing large-scale query reformulation data sampled from Bing search logs. This taxonomy leads us to a set of question templates and a simple yet effective slot filling algorithm. We further use this model as a source of weak supervision to automatically generate clarifying questions for training. Furthermore, we propose supervised and reinforcement learning models for generating clarifying questions learned from weak supervision data. We also investigate methods for generating candidate answers for each clarifying question, so users can select from a set of pre-defined answers. Human evaluation of the clarifying questions and candidate answers for hundreds of search queries demonstrates the effectiveness of the proposed solutions.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d03f875f-c870-4a8e-a2ad-4096af89e97dCited by top-tier papers31
- Analyzing and Learning from User Interactions for Search ClarificationHamed Zamani, Bhaskar Mitra, Everest Chen, Gord Lueck et al.SIGIR 2020 · 84 citations
- Few-Shot Conversational Dense RetrievalShi Yu, Zhenghao Liu, Chenyan Xiong, Tao Feng et al.SIGIR 2021 · 75 citations
- Building and Evaluating Open-Domain Dialogue Corpora with Clarifying QuestionsMohammad Aliannejadi, Julia Kiseleva, Aleksandr Chuklin, Jeff Dalton et al.EMNLP 2021 · 61 citations
- Guided Transformer: Leveraging Multiple External Sources for Representation Learning in Conversational SearchHelia Hashemi, Hamed Zamani, W. Bruce CroftSIGIR 2020 · 61 citations
- Zero-shot Clarifying Question Generation for Conversational SearchZhenduo Wang, Yuancheng Tu, Corby Rosset, Nick Craswell et al.WWW 2023 · 33 citations
Related papers
- Generating Multi-turn Clarification for Web Information SeekingZiliang Zhao, Zhicheng DouWWW 2024 · 15 citations
- Generating Clarifying Questions with Web Search ResultsZiliang Zhao, Zhicheng Dou, Jiaxin Mao, Ji-Rong WenSIGIR 2022 · 18 citations
- Controlling the Risk of Conversational Search via Reinforcement LearningZhenduo Wang, Qingyao AiWWW 2021 · 28 citations
- Towards a Better Understanding of Query Reformulation Behavior in Web SearchJia Chen, Jiaxin Mao, Yiqun Liu, Fan Zhang et al.WWW 2021 · 67 citations
- Retrieving Intent-covering Demonstrations for Clarification Generation in Conversational Search SystemsZiliang Zhao, Changle Qu, Zhicheng Dou, Haonan Chen et al.KDD 2025
