Cost-efficient Knowledge-based Question Answering with Large Language Models
Junnan Dong, Qinggang Zhang, Chuang Zhou, Hao Chen, Daochen Zha, Xiao Huang
Abstract
Knowledge-based question answering (KBQA) is widely used in many scenarios that necessitate domain knowledge. Large language models (LLMs) bring opportunities to KBQA, while their costs are significantly higher and absence of domain-specific knowledge during pre-training. We are motivated to combine LLMs and prior small models on knowledge graphs (KGMs) for both inferential accuracy and cost saving. However, it remains challenging since accuracy and cost are not readily combined in the optimization as two distinct metrics. It is also laborious for model selection since different models excel in diverse knowledge. To this end, we propose Coke, a novel cost-efficient strategy for KBQA with LLMs, modeled as a tailored multi-armed bandit problem to minimize calls to LLMs within limited budgets. We first formulate the accuracy expectation with a cluster-level Thompson Sampling for either KGMs or LLMs. A context-aware policy is optimized to further distinguish the expert model subject to the question semantics. The overall decision is bounded by the cost regret according to historical expenditure on failures. Extensive experiments showcase the superior performance of Coke, which moves the Pareto frontier with up to 20.89% saving of GPT-4 fees while achieving a 2.74% higher accuracy on the benchmark datasets.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 74c55129-a3d4-46aa-901d-35fe6c2de7c7Cited by top-tier papers3
- Structure-Guided Large Language Models for Text-to-SQL GenerationQinggang Zhang, Hao Chen, Junnan Dong, Shengyuan Chen et al.ICML 2025
- Enhancing Explainable Rating Prediction through Annotated Macro ConceptsHuachi Zhou, Shuang Zhou, Hao Chen, Ninghao Liu et al.ACL 2024
- Taming Language Models for Text-attributed Graph Learning with Decoupled AggregationChuang Zhou, Zhu Wang, Shengyuan Chen, Jiahe Du et al.ACL 2025
Builds on9
- Improving Multi-hop Question Answering over Knowledge Graphs using Knowledge Base EmbeddingsApoorv Saxena, Aditay Tripathi, Partha P. TalukdarACL 2020 · 488 citations
- Scalable Multi-Hop Relational Reasoning for Knowledge-Aware Question AnsweringYanlin Feng, Xinyue Chen, Bill Yuchen Lin, Peifeng Wang et al.EMNLP 2020 · 207 citations
- Aligning Distillation For Cold-start Item RecommendationFeiran Huang, Zefan Wang, Xiao Huang, Yufeng Qian et al.SIGIR 2023 · 100 citations
- Macro Graph Neural Networks for Online Billion-Scale Recommender SystemsHao Chen, Yuanchen Bei, Qijie Shen, Yue Xu et al.WWW 2024 · 96 citations
- Differentiable Neuro-Symbolic Reasoning on Large-Scale Knowledge GraphsShengyuan Chen, Yunfeng Cai, Huang Fang, Xiao Huang et al.NeurIPS 2023 · 56 citations
Related papers
- RJE: A Retrieval-Judgment-Exploration Framework for Efficient Knowledge Graph Question Answering with LLMsCan Lin, Zhengwang Jiang, Ling Zheng, Qi Zhao et al.EMNLP 2025
- KAM-CoT: Knowledge Augmented Multimodal Chain-of-Thoughts ReasoningDebjyoti Mondal, Suraj Modi, Subhadarshi Panda, Rituraj Singh et al.AAAI 2024 · 96 citations
- LightPROF: A Lightweight Reasoning Framework for Large Language Model on Knowledge GraphTu Ao, Yanhua Yu, Yuling Wang, Yang Deng et al.AAAI 2025 · 28 citations
- Paths-over-Graph: Knowledge Graph Empowered Large Language Model ReasoningXingyu Tan, Xiaoyang Wang, Qing Liu, Xiwei Xu et al.WWW 2025 · 86 citations
- Decoding on Graphs: Faithful and Sound Reasoning on Knowledge Graphs through Generation of Well-Formed ChainsKun Li, Tianhua Zhang, Xixin Wu, Hongyin Luo et al.ACL 2025
