RoCoFT: Efficient Finetuning of Large Language Models with Row-Column Updates
Md. Kowsher, Tara Esmaeilbeig, Chun-Nam Yu, Chen Chen, Mojtaba Soltanalian, Niloofar Yousefi
摘要
We propose Row-Column Fine-Tuning (Ro-CoFT), a parameter-efficient finetuning method for large language models based on updating only a few rows and columns of the weight matrices in transformers. Through extensive experiments with medium size language models like RoBERTa and DeBERTa, and large language models (LLMs) liken Bloom-7B, Llama2-7B and Llama2-13B, we show that our method gives comparable accuracies to the state-of-the-art Parameter-Efficient Finetuning methods while also being more memory and computation-efficient. We also study the reason behind the effectiveness of our method with tools from Neural Tangent Kernel (NTK) theory. We empirically demonstrate that our kernel, constructed using a restricted set of row and column parameters, is numerically close to the full-parameter kernel and gives comparable classification performance. Ablation studies are conducted to investigate the impact of different algorithmic choices, including the robustness of RoCoFT to any selection of rows and columns, as well as the optimal rank for the effective implementation of our method. * This work was done during an internship at Nokia Bell Labs.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- LiME: Lightweight Mixture of Experts for Efficient Multimodal Multi-task LearningMd Kowsher, Haris Mansoor, Nusrat Prottasha, Ozlem Garibay 等ICML 2026
- Predicting Through Generation: Why Generation Is Better for PredictionMd. Kowsher, Nusrat Jahan Prottasha, Prakash Bhat, Chun-Nam Yu 等ACL 2025
- FlowNIB: An Information Bottleneck Analysis of Bidirectional vs. Unidirectional Language ModelsMd Kowsher, Nusrat Jahan Prottasha, Shiyun Xu, Shetu Mohanto 等ICLR 2026
- Localized Low-Rank Adaptation within Clustered Parameter SubspacesJiahao Xiong, Yihe Liu, Xianming Hu, Hongbo Zhao 等ACL 2026
它引用的顶会 Paper23
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- QLoRA: Efficient Finetuning of Quantized LLMsTim Dettmers, Artidoro Pagnoni, Ari Holtzman, Luke ZettlemoyerNeurIPS 2023 · 被引用 5,863 次
- Deberta: decoding-Enhanced Bert with Disentangled AttentionPengcheng He, Xiaodong Liu, Jianfeng Gao, Weizhu ChenICLR 2021 · 被引用 3,729 次
- PIQA: Reasoning about Physical Commonsense in Natural LanguageYonatan Bisk, Rowan Zellers, Ronan Le Bras, Jianfeng Gao 等AAAI 2020 · 被引用 2,916 次
- Few-Shot Parameter-Efficient Fine-Tuning is Better and Cheaper than In-Context LearningHaokun Liu, Derek Tam, Mohammed Muqeeth, Jay Mohta 等NeurIPS 2022 · 被引用 1,483 次
相关 Paper
- LoRA Training in the NTK Regime has No Spurious Local MinimaUijeong Jang, Jason D. Lee, Ernest K. RyuICML 2024 · 被引用 41 次
- Linearization Explains Fine-Tuning in Large Language ModelsZahra Rahimi Afzal, Tara Esmaeilbeig, Mojtaba Soltanalian, Mesrob I. OhannessianNeurIPS 2025 · 被引用 5 次
- RoSA: Accurate Parameter-Efficient Fine-Tuning via Robust AdaptationMahdi Nikdan, Soroush Tabesh, Elvir Crncevic, Dan AlistarhICML 2024 · 被引用 53 次
- RoseLoRA: Row and Column-wise Sparse Low-rank Adaptation of Pre-trained Language Model for Knowledge Editing and Fine-tuningHaoyu Wang, Tianci Liu, Ruirui Li, Monica Xiao Cheng 等EMNLP 2024 · 被引用 6 次
- Robust Federated Finetuning of LLMs via Alternating Optimization of LoRAShuangyi Chen, Yuanxin Guo, Yue Ju, Hardik Dalal 等NeurIPS 2025 · 被引用 26 次
