Beyond I.I.D.: Three Levels of Generalization for Question Answering on Knowledge Bases
Yu Gu, Sue Kase, Michelle Vanni, Brian M. Sadler, Percy Liang, Xifeng Yan, Yu Su
摘要
Who is the producer of Spamalot? (AND Theater_Producer (JOIN (R producer) Spamalot)) -How many plays has Bob Boyett produced? (COUNT (AND Theater_Production (JOIN producer Bob_Boyett)) -Find plays that were staged in large theaters that could hold at least 20,000 people. (AND Theater_Production (JOIN (R staged_here) (JOIN (GE capacity 20000))) -How many theater productions has Oprah produced? (COUNT (AND Theater_Production (JOIN producer Oprah_Winfrey)) -Bob Boyett's production was housed in what theater capable of holding at least 10,000 people? (AND Theater (AND (GE capacity 10000) (JOIN staged_here (JOIN producer Bob_Boyett)))) -How many TV programs has Bob Boyett created? (COUNT (AND TV_Program (JOIN (R program_created) Bob_Boyett)) Knowledge Base KBQA Model Training Data I.I.D. Generalization Compositional Generalization Zero-Shot Generalization Figure 1: On large-scale KBs, collecting sufficient training data for KBQA to ensure i.i.d. distribution at test time is very difficult, if possible at all. We argue that practical KBQA models should have three levels of built-in generalization rather than solely relying on training data: (1) i.i.d. generalization to questions following the training distribution, (2) compositional generalization to novel compositions of schema items seen in training (marked blue), and (3) zero-shot generalization to unseen schema items or even domains (marked red). Our definition of generalization is based on the underlying logical forms (shown as S-expressions). Orthogonally, as illustrated by the examples, KBQA models should also have strong generalization to linguistic variation. Figure best viewed in color.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper70
- AgentBench: Evaluating LLMs as AgentsXiao Liu, Hao Yu, Hanchen Zhang, Yifan Xu 等ICLR 2024 · 被引用 748 次
- Think-on-Graph: Deep and Responsible Reasoning of Large Language Model on Knowledge GraphJiashuo Sun, Chengjin Xu, Lumingyuan Tang, Saizhuo Wang 等ICLR 2024 · 被引用 247 次
- UnifiedSKG: Unifying and Multi-Tasking Structured Knowledge Grounding with Text-to-Text Language ModelsTianbao Xie, Chen Henry Wu, Peng Shi, Ruiqi Zhong 等EMNLP 2022 · 被引用 222 次
- Plan-on-Graph: Self-Correcting Adaptive Planning of Large Language Model on Knowledge GraphsLiyi Chen, Panrong Tong, Zhongming Jin, Ying Sun 等NeurIPS 2024 · 被引用 160 次
- FlexKBQA: A Flexible LLM-Powered Framework for Few-Shot Knowledge Base Question AnsweringZhenyu Li, Sunqi Fan, Yu Gu, Xiuxing Li 等AAAI 2024 · 被引用 143 次
它引用的顶会 Paper2
- Measuring Compositional Generalization: A Comprehensive Method on Realistic DataDaniel Keysers, Nathanael Schärli, Nathan Scales, Hylke Buisman 等ICLR 2020 · 被引用 401 次
- SPARQA: Skeleton-Based Semantic Parsing for Complex Questions over Knowledge BasesYawei Sun, Lingling Zhang, Gong Cheng, Yuzhong QuAAAI 2020 · 被引用 143 次
相关 Paper
- Beyond Seen Data: Improving KBQA Generalization Through Schema-Guided Logical Form GenerationShengxiang Gao, Jey Han Lau, Jianzhong QiEMNLP 2025
- Interactive-KBQA: Multi-Turn Interactions for Knowledge Base Question Answering with Large Language ModelsGuanming Xiong, Junwei Bao, Wen ZhaoACL 2024
- FC-KBQA: A Fine-to-Coarse Composition Framework for Knowledge Base Question AnsweringLingxi Zhang, Jing Zhang, Yanling Wang, Shulin Cao 等ACL 2023 · 被引用 21 次
- KQA Pro: A Dataset with Explicit Compositional Programs for Complex Question Answering over Knowledge BaseShulin Cao, Jiaxin Shi, Liangming Pan, Lunyiu Nie 等ACL 2022
- A Foundation Model for Zero-shot Logical Query ReasoningMichael Galkin, Jincheng Zhou, Bruno Ribeiro, Jian Tang 等NeurIPS 2024 · 被引用 20 次
