SEER : A Knapsack approach to Exemplar Selection for In-Context HybridQA
Jonathan Tonglet, Manon Reusens, Philipp Borchert, Bart Baesens
Abstract
<p>Question answering over hybrid contexts is a complex task, which requires the combina tion of information extracted from unstructured texts and structured tables in various ways. Re cently, In-Context Learning demonstrated sig nificant performance advances for reasoning tasks. In this paradigm, a large language model performs predictions based on a small set of supporting exemplars. The performance of In-Context Learning depends heavily on the selection procedure of the supporting exem plars, particularly in the case of HybridQA where considering the diversity of reasoning chains and the large size of the hybrid con texts becomes crucial. In this work, we present Selection of ExEmplars for hybrid Reasoning (SEER), a novel method for selecting a set of exemplars that is both representative and di verse. The key novelty of SEER is that it for mulates exemplar selection as a Knapsack Inte ger Linear Program. The Knapsack framework provides the flexibility to incorporate diversity constraints that prioritize exemplars with desir able attributes, and capacity constraints that en sure that the prompt size respects the provided capacity budgets. The effectiveness of SEER is demonstrated on FinQA and TAT-QA, two real-world benchmarks for HybridQA, where it outperforms previous exemplar selection meth ods.</p>
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- Unraveling the Mechanics of Learning-Based Demonstration Selection for In-Context LearningHui Liu, Wenya Wang, Hao Sun, Chris Xing Tian et al.ACL 2025 · 13 citations
- Sample Efficient Demonstration Selection for In-Context LearningKiran Purohit, Venktesh V, Sourangshu Bhattacharya, Avishek AnandICML 2025
- EXPLORA: Efficient Exemplar Subset Selection for Complex ReasoningKiran Purohit, Venktesh V, Raghuram Devalla, Krishna Yerragorla et al.EMNLP 2024
Builds on17
- Calibrate Before Use: Improving Few-shot Performance of Language ModelsZihao Zhao, Eric Wallace, Shi Feng, Dan Klein et al.ICML 2021 · 1,843 citations
- Fantastically Ordered Prompts and Where to Find Them: Overcoming Few-Shot Prompt Order SensitivityYao Lu, Max Bartolo, Alastair Moore, Sebastian Riedel et al.ACL 2022 · 1,494 citations
- PAL: Program-aided Language ModelsLuyu Gao, Aman Madaan, Shuyan Zhou, Uri Alon et al.ICML 2023 · 700 citations
- LEVER: Learning to Verify Language-to-Code Generation with ExecutionAnsong Ni, Srini Iyer, Dragomir Radev, Veselin Stoyanov et al.ICML 2023 · 318 citations
- Automatic Chain of Thought Prompting in Large Language ModelsZhuosheng Zhang, Aston Zhang, Mu Li, Alex SmolaICLR 2023 · 234 citations
Related papers
- DQ-LoRe: Dual Queries with Low Rank Approximation Re-ranking for In-Context LearningJing Xiong, Zixuan Li, Chuanyang Zheng, Zhijiang Guo et al.ICLR 2024
- Prompt Optimization with EASE? Efficient Ordering-aware Automated Selection of ExemplarsZhaoxuan Wu, Xiaoqiang Lin, Zhongxiang Dai, Wenyang Hu et al.NeurIPS 2024 · 44 citations
- Exploring Hybrid Question Answering via Program-based PromptingQi Shi, Han Cui, Haofeng Wang, Qingfu Zhu et al.ACL 2024 · 3 citations
- Structured Semantic Information Helps Retrieve Better Examples for In-Context Learning Applied to Few-Shot Relation ExtractionAunabil Chakma, Mihai Surdeanu, Eduardo BlancoACL 2026 · 1 citation
- Combining Distantly Supervised Models with In Context Learning for Monolingual and Cross-Lingual Relation ExtractionVipul Kumar Rathore, Malik Hammad Faisal, Parag Singla, MausamACL 2026
