A Non-Factoid Question-Answering Taxonomy
Valeria Bolotova, Vladislav Blinov, Falk Scholer, W. Bruce Croft, Mark Sanderson
摘要
Non-factoid question answering (NFQA) is a challenging and underresearched task that requires constructing long-form answers, such as explanations or opinions, to open-ended non-factoid questions -NFQs. There is still little understanding of the categories of NFQs that people tend to ask, what form of answers they expect to see in return, and what the key research challenges of each category are.
This work presents the first comprehensive taxonomy of NFQ categories and the expected structure of answers. The taxonomy was constructed with a transparent methodology and extensively evaluated via crowdsourcing. The most challenging categories were identified through an editorial user study. We also release a dataset of categorised NFQs and a question category classifier 1 .
Finally, we conduct a quantitative analysis of the distribution of question categories using major NFQA datasets, showing that the NFQ categories that are the most challenging for current NFQA systems are poorly represented in these datasets. This imbalance may lead to insufficient system performance for challenging categories. The new taxonomy, along with the category classifier, will aid research in the area, helping to create more balanced benchmarks and to focus models on addressing specific categories.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- MemoryBench: A Benchmark for Memory and Continual Learning in LLM SystemsQingyao Ai, Yichen Tang, Changyue Wang, Jianming Long 等ICML 2026 · 被引用 47 次
- WikiHowQA: A Comprehensive Benchmark for Multi-Document Non-Factoid Question AnsweringValeria Bolotova-Baranova, Vladislav Blinov, Sofya Filippova, Falk Scholer 等ACL 2023 · 被引用 12 次
- A User-Centric Multi-Intent Benchmark for Evaluating Large Language ModelsJiayin Wang, Fengran Mo, Weizhi Ma, Peijie Sun 等EMNLP 2024 · 被引用 10 次
- Gesture and Audio-Haptic Guidance Techniques to Direct Conversations with Intelligent Voice InterfacesShwetha Rajaram, Hemant Bhaskar Surale, Codie McConkey, Carine Rognon 等CHI 2025 · 被引用 5 次
- Effective Contrastive Weighting for Dense Query ExpansionXiao Wang, Sean MacAvaney, Craig Macdonald, Iadh OunisACL 2023 · 被引用 2 次
它引用的顶会 Paper2
相关 Paper
- ASQA: Factoid Questions Meet Long-Form AnswersIvan Stelmakh, Yi Luan, Bhuwan Dhingra, Ming-Wei ChangEMNLP 2022 · 被引用 51 次
- An Empirical Study of Evaluating Long-form Question AnsweringNing Xian, Yixing Fan, Ruqing Zhang, Maarten de Rijke 等SIGIR 2025 · 被引用 2 次
- A Critical Evaluation of Evaluations for Long-form Question AnsweringFangyuan Xu, Yixiao Song, Mohit Iyyer, Eunsol ChoiACL 2023 · 被引用 25 次
- A Taxonomy of Empathetic Questions in Social DialogsEkaterina Svikhnushina, Iuliana Voinea, Anuradha Welivita, Pearl PuACL 2022 · 被引用 17 次
- How Do We Answer Complex Questions: Discourse Structure of Long-form AnswersFangyuan Xu, Junyi Jessy Li, Eunsol ChoiACL 2022 · 被引用 25 次
