Curriculum Debiasing: Toward Robust Parameter-Efficient Fine-Tuning Against Dataset Biases
Mingyu Lee, Yeachan Kim, Wing-Lam Mok, SangKeun Lee
摘要
Parameter-efficient fine-tuning (PEFT) addresses the memory footprint issue of full finetuning by modifying only a subset of model parameters. However, on datasets exhibiting spurious correlations, we observed that PEFT slows down the model's convergence on unbiased examples, while the convergence on biased examples remains fast. This leads to the model's overfitting on biased examples, causing significant performance degradation in outof-distribution (OOD) scenarios. Traditional debiasing methods mitigate this issue by emphasizing unbiased examples during training but often come at the cost of in-distribution (ID) performance drops. To address this trade-off issue, we propose a CURRICULUM DEBIASING framework that presents examples in a biasedto-unbiased order. Our framework initially limits the model's exposure to unbiased examples, which are more difficult to learn, allowing it to first establish a foundation on easy-to-converge biased examples. As training progresses, we gradually increase the proportion of unbiased examples in the training set, guiding the model away from reliance on spurious correlations. Compared to the original PEFT methods, our method accelerates convergence on unbiased examples by approximately twofold and improves ID and OOD performance by 1.2% and 8.0%, respectively. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper24
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- QLoRA: Efficient Finetuning of Quantized LLMsTim Dettmers, Artidoro Pagnoni, Ari Holtzman, Luke ZettlemoyerNeurIPS 2023 · 被引用 5,863 次
- Compacter: Efficient Low-Rank Hypercomplex Adapter LayersRabeeh Karimi Mahabadi, James Henderson, Sebastian RuderNeurIPS 2021 · 被引用 700 次
- Just Train Twice: Improving Group Robustness without Training Group InformationEvan Zheran Liu, Behzad Haghgoo, Annie S. Chen, Aditi Raghunathan 等ICML 2021 · 被引用 683 次
- DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding SharingPengcheng He, Jianfeng Gao, Weizhu ChenICLR 2023 · 被引用 394 次
相关 Paper
- Less Is More: Rethinking Parameter-Efficient Fine-Tuning from a Subtractive PerspectiveTianqi Jiang, Liu Yang, Xi-Le Zhao, Zixuan Qin 等AAAI 2026
- GIST: Improving Parameter Efficient Fine-Tuning via Knowledge InteractionJiacheng Ruan, Jingsheng Gao, Mingye Xie, Suncheng Xiang 等ACM MM 2024 · 被引用 6 次
- Interweaving Memories of a Siamese Large Language ModelXin Song, Zhikai Xue, Guoxiu He, Jiawei Liu 等AAAI 2025
- Towards Robust and Generalized Parameter-Efficient Fine-Tuning for Noisy Label LearningYeachan Kim, Junho Kim, SangKeun LeeACL 2024 · 被引用 4 次
- LT-Soups: Bridging Head and Tail Classes via Subsampled Model SoupsMasih Aminbeidokhti, Subhankar Roy, Eric Granger, Elisa Ricci 等NeurIPS 2025 · 被引用 1 次
