QUEST: A Retrieval Dataset of Entity-Seeking Queries with Implicit Set Operations
Chaitanya Malaviya, Peter Shaw, Ming-Wei Chang, Kenton Lee, Kristina Toutanova
摘要
Formulating selective information needs results in queries that implicitly specify set operations, such as intersection, union, and difference. For instance, one might search for "shorebirds that are not sandpipers" or "science-fiction films shot in England". To study the ability of retrieval systems to meet such information needs, we construct QUEST, a dataset of 3357 natural language queries with implicit set operations, that map to a set of entities corresponding to Wikipedia documents. The dataset challenges models to match multiple constraints mentioned in queries with corresponding evidence in documents and correctly perform various set operations. The dataset is constructed semi-automatically using Wikipedia category names. Queries are automatically composed from individual categories, then paraphrased and further validated for naturalness and fluency by crowdworkers. Crowdworkers also assess the relevance of entities based on their documents and highlight attribution of query constraints to spans of document text. We analyze several modern retrieval systems, finding that they often struggle on such queries. Queries involving negation and conjunction are particularly challenging and systems are further challenged with combinations of these operations. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- On the Theoretical Limitations of Embedding-Based RetrievalOrion Weller, Michael Boratko, Iftekhar Naim, Jinhyuk LeeICLR 2026 · 被引用 138 次
- Efficient Long Context Language Model Retrieval with CompressionMinju Seo, Jinheon Baek, Seongyun Lee, Sung Ju HwangACL 2025 · 被引用 2 次
- Atomic Self-Consistency for Better Long Form GenerationsRaghuveer Thirukovalluru, Yukun Huang, Bhuwan DhingraEMNLP 2024 · 被引用 2 次
- Chatgpt Inaccuracy Mitigation During Technical Report Understanding: Are we There Yet?Salma Begum Tamanna, Gias Uddin, Song Wang, Lan Xia 等ICSE 2025 · 被引用 2 次
- LogiCoL: Logically-Informed Contrastive Learning for Set-based Dense RetrievalYanzhen Shen, Sihao Chen, Xueqiang Xu, Yunyi Zhang 等EMNLP 2025 · 被引用 1 次
它引用的顶会 Paper4
- Measuring Compositional Generalization: A Comprehensive Method on Realistic DataDaniel Keysers, Nathanael Schärli, Nathan Scales, Hylke Buisman 等ICLR 2020 · 被引用 401 次
- Beyond I.I.D.: Three Levels of Generalization for Question Answering on Knowledge BasesYu Gu, Sue Kase, Michelle Vanni, Brian M. Sadler 等WWW 2021 · 被引用 304 次
- Large Dual Encoders Are Generalizable RetrieversJianmo Ni, Chen Qu, Jing Lu, Zhuyun Dai 等EMNLP 2022 · 被引用 145 次
- Joint Passage Ranking for Diverse Multi-Answer RetrievalSewon Min, Kenton Lee, Ming-Wei Chang, Kristina Toutanova 等EMNLP 2021 · 被引用 1 次
相关 Paper
- ComLQ: Benchmarking Complex Logical Queries in Information RetrievalGanlin Xu, Zhitao Yin, Linghao Zhang, Jiaqing Liang 等AAAI 2026
- SetCSE: Set Operations using Contrastive Learning of Sentence EmbeddingsKang LiuICLR 2024 · 被引用 5 次
- You Make me Feel like a Natural Question: Training QA Systems on Transformed Trivia QuestionsTasnim Kabir, Yoo Yeon Sung, Saptarashmi Bandyopadhyay, Hao Zou 等EMNLP 2024
- IIRC: A Dataset of Incomplete Information Reading Comprehension QuestionsJames Ferguson, Matt Gardner, Hannaneh Hajishirzi, Tushar Khot 等EMNLP 2020 · 被引用 42 次
- Classifying Term Variants in Query FormulationNuha Abu Onq, Mark Sanderson, Falk ScholerSIGIR 2025 · 被引用 1 次
