PTP: Boosting Stability and Performance of Prompt Tuning with Perturbation-Based Regularizer
Lichang Chen, Jiuhai Chen, Heng Huang, Minhao Cheng
摘要
Recent studies show that prompt tuning can better leverage the power of large language models than fine-tuning on downstream natural language understanding tasks. Nonetheless, current prompt tuning methods encounter instability during training, marked by a high variance in scores given different random seeds. In addressing this crucial issue, we uncover that the loss landscape of standard prompt tuning, when visualized, is remarkably steep, i.e., minor alterations in the input data can trigger substantial fluctuations in the loss landscape, which is an essential factor that leads to the training instability. In light of this finding, we incorporate perturbation-based regularizers to temper the loss landscape within the prompt tuning process. We thus present a novel algorithm, called Prompt Tuning with Perturbation-based regularizer (PTP), that can significantly reduce training instability and concurrently enhance the performance of prompt tuning. Specifically, we design two variants of perturbation-based regularizers: one that employs random noise, and another that uses an adversarial approach. Importantly, our proposed perturbations display flexibility in both the text and embedding spaces. Extensive experiments show the effectiveness of our proposed methods in stabilizing the training. Our new algorithms improve the state-of-the-art prompt tuning methods by 1.94% and 2.34% on SuperGLUE and FewGLUE benchmarks, respectively.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- AlpaGasus: Training a Better Alpaca with Fewer DataLichang Chen, Shiyang Li, Jun Yan, Hai Wang 等ICLR 2024 · 被引用 295 次
- MAGIS: LLM-Based Multi-Agent Framework for GitHub Issue ResolutionWei Tao, Yucheng Zhou, Yanlin Wang, Wenqiang Zhang 等NeurIPS 2024 · 被引用 210 次
- InstructZero: Efficient Instruction Optimization for Black-Box Large Language ModelsLichang Chen, Jiuhai Chen, Tom Goldstein, Heng Huang 等ICML 2024 · 被引用 64 次
- xLSTM-Mixer: Multivariate Time Series Forecasting by Mixing via Scalar MemoriesMaurice Kraus, Felix Divo, Devendra Singh Dhami, Kristian KerstingNeurIPS 2025 · 被引用 27 次
- Fighting Fire with Fire: The Dual Role of LLMs in Crafting and Detecting Elusive DisinformationJason Samuel Lucas, Adaku Uchendu, Michiharu Yamashita, Jooyoung Lee 等EMNLP 2023 · 被引用 23 次
它引用的顶会 Paper13
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel 等ICLR 2020 · 被引用 7,418 次
- Deberta: decoding-Enhanced Bert with Disentangled AttentionPengcheng He, Xiaodong Liu, Jianfeng Gao, Weizhu ChenICLR 2021 · 被引用 3,729 次
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad 等ACL 2020 · 被引用 1,224 次
- Large-Scale Adversarial Training for Vision-and-Language Representation LearningZhe Gan, Yen-Chun Chen, Linjie Li, Chen Zhu 等NeurIPS 2020 · 被引用 561 次
- BERT-ATTACK: Adversarial Attack Against BERT Using BERTLinyang Li, Ruotian Ma, Qipeng Guo, Xiangyang Xue 等EMNLP 2020 · 被引用 529 次
相关 Paper
- APrompt: Attention Prompt Tuning for Efficient Adaptation of Pre-trained Language ModelsQifan Wang, Yuning Mao, Jingang Wang, Hanchao Yu 等EMNLP 2023 · 被引用 25 次
- Improving Calibration in Test-Time Prompt Tuning for Vision-Language Models via Data-Free Flatness-Aware Prompt PretrainingHyeonseo Jang, Jaebyeong Jeon, Joong-Won Hwang, Kibok LeeCVPR 2026 · 被引用 2 次
- StablePrompt : Automatic Prompt Tuning using Reinforcement Learning for Large Language ModelMinchan Kwon, Gaeun Kim, Jongsuk Kim, Haeil Lee 等EMNLP 2024 · 被引用 11 次
- Vector-Quantized Input-Contextualized Soft Prompts for Natural Language UnderstandingRishabh Bhardwaj, Amrita Saha, Steven C. H. Hoi, Soujanya PoriaEMNLP 2022 · 被引用 5 次
- On the Stability of Fine-tuning BERT: Misconceptions, Explanations, and Strong BaselinesMarius Mosbach, Maksym Andriushchenko, Dietrich KlakowICLR 2021 · 被引用 448 次
