Lune

ICLR2026Top-tier venue

MiSS: Revisiting the Trade-off in LoRA with an Efficient Shard-Sharing Structure

Jiale Kang, Qingyu Yin

2026Year
3Citations
1Top-tier citations

Abstract

Low-Rank Adaptation (LoRA) is a widely adopted technique for parameter-efficient fine-tuning, but its slow convergence has spurred the development of numerous variants. Nevertheless, current approaches struggle to achieve simultaneous improvements in performance, memory footprint, and computational efficiency. To address this challenge, we revisit the causes of LoRA’s slow convergence and, based on these insights, propose Matrix Shard Sharing (MiSS) that shards the original weight matrix and updates by sharing a single trainable matrix D\boldsymbol{D} initialized to zero. To simultaneously ensure computational efficiency, low memory footprint, and scalable serving, we introduce MiSSe^e. Through theoretical analyses and empirical results, our method reduces optimization complexity while maintaining strong performance, striking a favorable balance between performance, memory, and efficiency. Furthermore, we provide a comprehensive analysis of different PEFT methods with respect to memory usage, initialization time, and computational efficiency. By mapping the Pareto frontier, we show that MiSS achieves a favorable balance across these dimensions, integrating the strengths of prior approaches.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext edfcec7c-6909-47b2-b182-761ee9f75297

Cited by top-tier papers1

Ask how each one uses it

Builds on9

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines