Spectral Bridge Variational Inference: Dynamic LoRA via Bures-Wasserstein Gradient Flows
Yuhang Xi, Yu-Feng Yu, Chuan-Xian Ren, Zhao-Rong Lai
摘要
Parameter-Efficient Fine-Tuning (PEFT) is essential for adapting Large Language Models, yet existing methods struggle to balance capacity with computational efficiency. Standard approaches enforce rigid low-rank constraints, while dynamic alternatives incur significant memory overheads. To resolve this, we propose Spectral Bridge Variational Inference (SBVI), a geometric framework reformulating LoRA as a continuous Wasserstein gradient flow on the manifold of Gaussian measures. Instead of fixing ranks at initialization, SBVI governs singular value evolution via a stochastic differential equation driven by thermodynamic competition between task gradients and adaptive entropic friction. This induces a spectral bifurcation that automatically prunes noise modes while amplifying signal-rich components, discovering an optimal layer-wise rank distribution. We derive a scalable algorithm with linear complexity using factorized Riemannian retractions and Empirical Bayes friction updates. Experiments on reasoning and coding benchmarks show SBVI achieves state-of-the-art performance, offering superior accuracy and memory efficiency over existing static and dynamic methods. Our code is publicly available at: https://github.com/ xiyuhang2003/SBVI-LoRA
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper20
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- Measuring Massive Multitask Language UnderstandingDan Hendrycks, Collin Burns, Steven Basart, Andy Zou 等ICLR 2021 · 被引用 7,905 次
- PIQA: Reasoning about Physical Commonsense in Natural LanguageYonatan Bisk, Rowan Zellers, Ronan Le Bras, Jianfeng Gao 等AAAI 2020 · 被引用 2,916 次
- GaLore: Memory-Efficient LLM Training by Gradient Low-Rank ProjectionJiawei Zhao, Zhenyu Zhang, Beidi Chen, Zhangyang Wang 等ICML 2024 · 被引用 433 次
- LoRA+: Efficient Low Rank Adaptation of Large ModelsSoufiane Hayou, Nikhil Ghosh, Bin YuICML 2024 · 被引用 388 次
相关 Paper
- DoRA: Enhancing Parameter-Efficient Fine-Tuning with Dynamic Rank DistributionYulong Mao, Kaiyu Huang, Changhao Guan, Ganglin Bao 等ACL 2024 · 被引用 15 次
- FlexLoRA: Entropy-Guided Flexible Low-Rank AdaptationMuqing Liu, Chongjie Si, Yuheng JiaICLR 2026 · 被引用 3 次
- Not All Directions Matter: Towards Structured and Task-Aware Low-Rank Model AdaptationXi Xiao, Chenrui Ma, Yunbei Zhang, Chen Liu 等ACL 2026 · 被引用 6 次
- COBRA: Contribution-Based Bayesian Rank Allocation for Parameter-Efficient Fine-TuningHongcheng Ding, Xuanze Zhao, LIU XUANHUANG, Jing Jin 等ICML 2026
- Flat-LoRA: Low-Rank Adaptation over a Flat Loss LandscapeTao Li, Zhengbao He, Yujun Li, Yasheng Wang 等ICML 2025
