Cost-efficient Knowledge-based Question Answering with Large Language Models
Junnan Dong, Qinggang Zhang, Chuang Zhou, Hao Chen, Daochen Zha, Xiao Huang
摘要
Knowledge-based question answering (KBQA) is widely used in many scenarios that necessitate domain knowledge. Large language models (LLMs) bring opportunities to KBQA, while their costs are significantly higher and absence of domain-specific knowledge during pre-training. We are motivated to combine LLMs and prior small models on knowledge graphs (KGMs) for both inferential accuracy and cost saving. However, it remains challenging since accuracy and cost are not readily combined in the optimization as two distinct metrics. It is also laborious for model selection since different models excel in diverse knowledge. To this end, we propose Coke, a novel cost-efficient strategy for KBQA with LLMs, modeled as a tailored multi-armed bandit problem to minimize calls to LLMs within limited budgets. We first formulate the accuracy expectation with a cluster-level Thompson Sampling for either KGMs or LLMs. A context-aware policy is optimized to further distinguish the expert model subject to the question semantics. The overall decision is bounded by the cost regret according to historical expenditure on failures. Extensive experiments showcase the superior performance of Coke, which moves the Pareto frontier with up to 20.89% saving of GPT-4 fees while achieving a 2.74% higher accuracy on the benchmark datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Structure-Guided Large Language Models for Text-to-SQL GenerationQinggang Zhang, Hao Chen, Junnan Dong, Shengyuan Chen 等ICML 2025
- Enhancing Explainable Rating Prediction through Annotated Macro ConceptsHuachi Zhou, Shuang Zhou, Hao Chen, Ninghao Liu 等ACL 2024
- Taming Language Models for Text-attributed Graph Learning with Decoupled AggregationChuang Zhou, Zhu Wang, Shengyuan Chen, Jiahe Du 等ACL 2025
它引用的顶会 Paper9
- Improving Multi-hop Question Answering over Knowledge Graphs using Knowledge Base EmbeddingsApoorv Saxena, Aditay Tripathi, Partha P. TalukdarACL 2020 · 被引用 488 次
- Scalable Multi-Hop Relational Reasoning for Knowledge-Aware Question AnsweringYanlin Feng, Xinyue Chen, Bill Yuchen Lin, Peifeng Wang 等EMNLP 2020 · 被引用 207 次
- Aligning Distillation For Cold-start Item RecommendationFeiran Huang, Zefan Wang, Xiao Huang, Yufeng Qian 等SIGIR 2023 · 被引用 100 次
- Macro Graph Neural Networks for Online Billion-Scale Recommender SystemsHao Chen, Yuanchen Bei, Qijie Shen, Yue Xu 等WWW 2024 · 被引用 96 次
- Differentiable Neuro-Symbolic Reasoning on Large-Scale Knowledge GraphsShengyuan Chen, Yunfeng Cai, Huang Fang, Xiao Huang 等NeurIPS 2023 · 被引用 56 次
相关 Paper
- RJE: A Retrieval-Judgment-Exploration Framework for Efficient Knowledge Graph Question Answering with LLMsCan Lin, Zhengwang Jiang, Ling Zheng, Qi Zhao 等EMNLP 2025
- KAM-CoT: Knowledge Augmented Multimodal Chain-of-Thoughts ReasoningDebjyoti Mondal, Suraj Modi, Subhadarshi Panda, Rituraj Singh 等AAAI 2024 · 被引用 96 次
- LightPROF: A Lightweight Reasoning Framework for Large Language Model on Knowledge GraphTu Ao, Yanhua Yu, Yuling Wang, Yang Deng 等AAAI 2025 · 被引用 28 次
- Paths-over-Graph: Knowledge Graph Empowered Large Language Model ReasoningXingyu Tan, Xiaoyang Wang, Qing Liu, Xiwei Xu 等WWW 2025 · 被引用 86 次
- Decoding on Graphs: Faithful and Sound Reasoning on Knowledge Graphs through Generation of Well-Formed ChainsKun Li, Tianhua Zhang, Xixin Wu, Hongyin Luo 等ACL 2025
