Instructive Decoding: Instruction-Tuned Large Language Models are Self-Refiner from Noisy Instructions
Taehyeon Kim, Joonkee Kim, Gihun Lee, Se-Young Yun
Abstract
While instruction-tuned language models have demonstrated impressive zero-shot generalization, these models often struggle to generate accurate responses when faced with instructions that fall outside their training set. This paper presents Instructive Decoding (ID), a simple yet effective approach that augments the efficacy of instruction-tuned models. Specifically, ID adjusts the logits for next-token prediction in a contrastive manner, utilizing predictions generated from a manipulated version of the original instruction, referred to as a noisy instruction. This noisy instruction aims to elicit responses that could diverge from the intended instruction yet remain plausible. We conduct experiments across a spectrum of such noisy instructions, ranging from those that insert semantic noise via random words to others like 'opposite' that elicit the deviated responses. Our approach achieves considerable performance gains across various instruction-tuned models and tasks without necessitating any additional parameter updates. Notably, utilizing 'opposite' as the noisy instruction in ID, which exhibits the maximum divergence from the original instruction, consistently produces the most significant performance gains across multiple models and tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9bfb0107-4e46-4a34-8cd0-81408b0781a9Cited by top-tier papers12
- T-REG: Preference Optimization with Token-Level Reward RegularizationWenxuan Zhou, Shujian Zhang, Lingxiao Zhao, Tao MengACL 2025 · 11 citations
- Accelerating Blockwise Parallel Language Models with Draft RefinementTaehyeon Kim, Ananda Theertha Suresh, Kishore Papineni, Michael D. Riley et al.NeurIPS 2024 · 8 citations
- CoCoLex: Confidence-guided Copy-based Decoding for Grounded Legal Text GenerationT. Y. S. S. Santosh, Youssef Tarek Elkhayat, Oana Ichim, Pranav Shetty et al.ACL 2025 · 6 citations
- GNN-as-Judge: Unleashing the Power of LLMs for Graph Learning with GNN FeedbackRuiyao Xu, Kaize DingICLR 2026 · 5 citations
- Safety Alignment of Large Language Models via Contrasting Safe and Harmful DistributionsXiaoyun Zhang, Zhengyue Zhao, Wenxuan Shi, Kaidi Xu et al.AAAI 2026 · 4 citations
Builds on19
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- Measuring Massive Multitask Language UnderstandingDan Hendrycks, Collin Burns, Steven Basart, Andy Zou et al.ICLR 2021 · 7,905 citations
- The Curious Case of Neural Text DegenerationAri Holtzman, Jan Buys, Li Du, Maxwell Forbes et al.ICLR 2020 · 4,112 citations
- Multitask Prompted Training Enables Zero-Shot Task GeneralizationVictor Sanh, Albert Webson, Colin Raffel, Stephen H. Bach et al.ICLR 2022 · 1,976 citations
Related papers
- Evaluating the Zero-shot Robustness of Instruction-tuned Language ModelsJiuding Sun, Chantal Shaib, Byron C. WallaceICLR 2024 · 75 citations
- Learning Instructions with Unlabeled Data for Zero-Shot Cross-Task GeneralizationYuxian Gu, Pei Ke, Xiaoyan Zhu, Minlie HuangEMNLP 2022 · 3 citations
- Finetuned Language Models are Zero-Shot LearnersJason Wei, Maarten Bosma, Vincent Y. Zhao, Kelvin Guu et al.ICLR 2022 · 4,966 citations
- Enhancing NLU in Large Language Models Using Adversarial Noisy Instruction TuningShengyuan Bai, Qibin Li, Zhe Wang, Nai Zhou et al.AAAI 2025 · 6 citations
- Active Instruction Tuning: Improving Cross-Task Generalization by Training on Prompt Sensitive TasksPo-Nien Kung, Fan Yin, Di Wu, Kai-Wei Chang et al.EMNLP 2023 · 7 citations
