Parameter Efficient Fine-tuning via Explained Variance Adaptation
Fabian Paischer, Lukas Hauzenberger, Thomas Schmied, Benedikt Alkin, Marc Peter Deisenroth, Sepp Hochreiter
Abstract
Foundation models (FMs) are pre-trained on large-scale datasets and then fine-tuned for a specific downstream task. The most common fine-tuning method is to update pretrained weights via low-rank adaptation (LoRA). Existing initialization strategies for LoRA often rely on singular value decompositions (SVD) of gradients or weight matrices. However, they do not provably maximize the expected gradient signal, which is critical for fast adaptation. To this end, we introduce Explained Variance Adaptation (EVA), an initialization scheme that uses the directions capturing the most activation variance, provably maximizing the expected gradient signal and accelerating fine-tuning. EVA performs incremental SVD on minibatches of activation vectors and selects the right-singular vectors for initialization once they converged. Further, by selecting the directions that capture the most activation-variance for a given rank budget, EVA accommodates adaptive ranks that reduce the number of trainable parameters. We apply EVA to a variety of fine-tuning tasks as language generation and understanding, image classification, and reinforcement learning. EVA exhibits faster convergence than competitors and achieves the highest average score across a multitude of tasks per domain while reducing the number of trainable parameters through rank redistribution. In summary, EVA establishes a new Pareto frontier compared to existing LoRA initialization schemes in both accuracy and efficiency.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers7
- GoRA: Gradient-driven Adaptive Low Rank AdaptationHaonan He, Peng Ye, Yuchen Ren, Yuan Yuan et al.NeurIPS 2025 · 17 citations
- GraLoRA: Granular Low-Rank Adaptation for Parameter-Efficient Fine-TuningYeonjoon Jung, Daehyun Ahn, Hyungjun Kim, Taesu Kim et al.NeurIPS 2025 · 11 citations
- The Primacy of Magnitude in Low-Rank AdaptationZicheng Zhang, Haoran Li, Yifeng Zhang, Guoqiang Gong et al.NeurIPS 2025 · 7 citations
- ConsNoTrainLoRA: Data-driven Weight Initialization of Low-Rank Adapters Using ConstraintsDebasmit Das, Hyoungwoo Park, Munawar Hayat, Seokeon Choi et al.ICCV 2025 · 1 citation
- The Geometry of Narrow Fine-Tuning Degradation: Trajectory Lock-in and Spectral BifurcationJia Liu, Jiaxin Luo, Xinhao Qiu, Yixue Hao et al.ICML 2026
Builds on17
- WinoGrande: An Adversarial Winograd Schema Challenge at ScaleKeisuke Sakaguchi, Ronan Le Bras, Chandra Bhagavatula, Yejin ChoiAAAI 2020 · 3,037 citations
- Is Your Code Generated by ChatGPT Really Correct? Rigorous Evaluation of Large Language Models for Code GenerationJiawei Liu, Chunqiu Steven Xia, Yuyao Wang, Lingming ZhangNeurIPS 2023 · 2,317 citations
- Few-Shot Parameter-Efficient Fine-Tuning is Better and Cheaper than In-Context LearningHaokun Liu, Derek Tam, Mohammed Muqeeth, Jay Mohta et al.NeurIPS 2022 · 1,483 citations
- ZeRO: memory optimizations toward training trillion parameter modelsSamyam Rajbhandari, Jeff Rasley, Olatunji Ruwase, Yuxiong HeSC 2020 · 852 citations
- DoRA: Weight-Decomposed Low-Rank AdaptationShih-Yang Liu, Chien-Yi Wang, Hongxu Yin, Pavlo Molchanov et al.ICML 2024 · 820 citations
Related papers
- AIRA: Activation-Informed Low-Rank Adaptation for Large ModelsLujun Li, Dezhi Li, Cheng Lin, Wei Li et al.ICCV 2025
- AdaRankGrad: Adaptive Gradient Rank and Moments for Memory-Efficient LLMs Training and Fine-TuningYehonathan Refael, Jonathan Svirsky, Boris Shustin, Wasim Huleihel et al.ICLR 2025
- Efficient Fine-Tuning of Large Models Via Nested Low-Rank AdaptationLujun Li, Cheng Lin, Dezhi Li, You-Liang Huang et al.ICCV 2025 · 1 citation
- TLoRA: Task-aware Low Rank Adaptation of Large Language ModelsWeicheng Lin, Yi Zhang, Jiawei Dang, Liang-Jie ZhangACL 2026
- FedEx-LoRA: Exact Aggregation for Federated and Efficient Fine-Tuning of Large Language ModelsRaghav Singhal, Kaustubh Ponkshe, Praneeth VepakommaACL 2025 · 10 citations
