Lune

ACL2026Top-tier venue

WeightLoRA: Keep Only Necessary Adapters

Andrey Veprikov, Vladimir Solodkin, Alexander Zyl, Andrey V. Savchenko, Aleksandr Beznosikov

2026Year
2Citations

Abstract

The widespread utilization of language models in modern applications is inconceivable without Parameter-Efficient Fine-Tuning techniques, such as low-rank adaptation (LoRA\texttt{LoRA}), which adds trainable adapters to selected layers. Although LoRA\texttt{LoRA} may obtain accurate solutions, it requires significant memory to train large models and intuition on which layers to add adapters. In this paper, we propose a novel method, WeightLoRA\texttt{WeightLoRA}, which overcomes this issue by adaptive selection of the most critical LoRA\texttt{LoRA} heads throughout the optimization process. As a result, we can significantly reduce the number of trainable parameters while maintaining the capability to obtain consistent or even superior metric values. We conduct experiments for a series of competitive benchmarks and DeBERTa, BART, and Llama models, comparing our method with different adaptive approaches. The experimental results demonstrate the efficacy of WeightLoRA\texttt{WeightLoRA} and the superior performance of WeightLoRA+\texttt{WeightLoRA+} in almost all cases.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext 24e6c911-7c01-4f8d-b4b5-c853492788a6

Builds on6

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines