Lune

ICML2026顶会

Bayesian-LoRA: Probabilistic Low-Rank Adaptation of Large Language Models

Moule Lin, Shuhao Guan, Andrea Patane, David Gregg, Goetz Botterweck

2026年份

摘要

Large language models are typically optimized for accuracy and, therefore, will guess even when uncertain about their predictions. This problem becomes especially pronounced when the model is fine-tuned on small datasets, which often causes overfitting and results in a tendency toward miscalibration. In this work, we introduce Bayesian-LoRA, which reformulates the deterministic LoRA update as a probabilistic low-rank representation inspired by Sparse Gaussian Processes (SGP). We identify a structural isomorphism between LoRA's factorization and Kronecker-factored SGP posteriors, and show that LoRA emerges as a limiting case when posterior uncertainty collapses. We conduct extensive experiments on various LLM architectures across commonsense reasoning, language modeling, and mathematical reasoning benchmarks. With only approximately 0.42M additional parameters and ≈1.2×{\approx}1.2{\times} training cost relative to standard LoRA, Bayesian-LoRA significantly improves calibration across models from 7B up to 30B, achieving up to 84% Expected Calibration Error (ECE) and 76% Negative Log-Likelihood (NLL) reduction while maintaining competitive accuracy for both in-distribution and out-of-distribution (OoD) evaluations.

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

它引用的顶会 Paper20

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖