KnowProxy: Adapting Large Language Models by Knowledge-guided Proxy
Gukhyeon Lee, Yeachan Kim, Sangkeun Lee
Abstract
Adapting large language models (LLMs) using smaller proxy models has been shown to improve training efficiency, where the LLMs remain frozen while the proxies are tuned on top. However, this approach typically requires access to the output probability distributions of LLMs, which are often inaccessible or unstable. To address this limitation, we propose KNOWPROXY, a knowledge-guided proxy framework in which the proxy is trained with textual knowledge rather than probability distributions. Specifically, we first elicit textual knowledge and reasoning from frozen LLMs through prompting, and then the proxy model learns to adapt this reasoning to target task distributions. We evaluate KNOWPROXY on diverse reasoning benchmarks with different fine-tuning scenarios. Comprehensive results show that KNOWPROXY achieves competitive or even better performance without direct access to probability distributions, thereby providing a scalable and versatile alternative to traditional fine-tuning. 1 * Equal contribution. 1 Our code and data are available at https://github.com/2gukhyeon/KnowProxy.git .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b4e83046-17a6-4dac-b859-faa0813e91f6Builds on26
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- Large Language Models are Zero-Shot ReasonersTakeshi Kojima, Shixiang Shane Gu, Machel Reid, Yutaka Matsuo et al.NeurIPS 2022 · 8,168 citations
- QLoRA: Efficient Finetuning of Quantized LLMsTim Dettmers, Artidoro Pagnoni, Ari Holtzman, Luke ZettlemoyerNeurIPS 2023 · 5,863 citations
- TruthfulQA: Measuring How Models Mimic Human FalsehoodsStephanie Lin, Jacob Hilton, Owain EvansACL 2022 · 3,228 citations
Related papers
- SLQ: Bridging Modalities via Shared Latent Queries for Retrieval with Frozen MLLMsHaoran Lou, Ziyan Liu, Chunxiao Fan, Yuexin Wu et al.ICML 2026
- Task-Aware Data Selection via Proxy-Label Enhanced Distribution Matching for LLM FinetuningHao Cheng, Rui Zhang, Ling Li, Na Di et al.ICLR 2026
- LightPROF: A Lightweight Reasoning Framework for Large Language Model on Knowledge GraphTu Ao, Yanhua Yu, Yuling Wang, Yang Deng et al.AAAI 2025 · 28 citations
- Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMsJaemin Kim, Hangeol Chang, Hyunmin Hwang, Choonghan Kim et al.ICML 2026 · 1 citation
- A Language-Guided Bayesian Optimization for Efficient LoRA Hyperparameter SearchBaek Seong-Eun, Lee Jung-Mok, Kim Sung-Bin, Tae-Hyun OhICML 2026 · 2 citations
