MTA: A Merge-then-Adapt Framework for Personalized Large Language Models
Xiaopeng Li, Yuanjin Zheng, Wanyu Wang, Wenlin Zhang, Pengyue Jia, Yingyi Zhang, Haiying He, Mengyang Ma, Yiqi Wang, Maolin Wang, Xuetao Wei, Xiangyu Zhao
摘要
Personalized Large Language Models (PLLMs) aim to align model outputs with individual user preferences, a crucial capability for usercentric applications. However, the prevalent approach of fine-tuning a separate module for each user faces two major limitations: (1) storage costs scale linearly with the number of users, rendering the method unscalable; and (2) fine-tuning a static model from scratch often yields suboptimal performance for users with sparse data. To address these challenges, we propose MTA, a Merge-then-Adapt framework for PLLMs. MTA comprises three key stages. First, we construct a shared Meta-LoRA Bank by selecting anchor users and pre-training meta-personalization traits within meta-LoRA modules. Second, to ensure scalability and enable dynamic personalization combination beyond static models, we introduce an Adaptive LoRA Fusion stage. This stage retrieves and dynamically merges the most relevant anchor meta-LoRAs to synthesize a user-specific adapter on the fly, thereby removing the need to maintain a dedicated, persistently stored per-user adapter for each user and enabling flexible personalization. Third, we propose a LoRA Stacking for Few-Shot Personalization stage, which optionally applies an additional ultra-low-rank, lightweight residual LoRA module on top of the merged LoRA. This stacked module captures user-specific residual signals under few-shot settings. Extensive experiments on the LaMP benchmark demonstrate that our approach outperforms existing SOTA methods across multiple tasks. Our code is available at https://github. com/Applied-Machine-Learning-Lab/ ACL2026_MTA .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper11
- DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding SharingPengcheng He, Jianfeng Gao, Weizhu ChenICLR 2023 · 被引用 394 次
- One Chatbot Per Person: Creating Personalized Chatbots based on Implicit User ProfilesZhengyi Ma, Zhicheng Dou, Yutao Zhu, Hanxun Zhong 等SIGIR 2021 · 被引用 67 次
- LLM4Rerank: LLM-based Auto-Reranking Framework for RecommendationsJingtong Gao, Bo Chen, Xiangyu Zhao, Weiwen Liu 等WWW 2025 · 被引用 50 次
- PROPER: A Progressive Learning Framework for Personalized Large Language Models with Group-Level AdaptationLinhai Zhang, Jialong Wu, Deyu Zhou, Yulan HeACL 2025 · 被引用 14 次
- Harnessing Large Language Models for Knowledge Graph Question Answering via Adaptive Multi-Aspect Retrieval-AugmentationDerong Xu, Xinhang Li, Ziheng Zhang, Zhenxi Lin 等AAAI 2025 · 被引用 14 次
相关 Paper
- One Adapts to Any: Meta Reward Modeling for Personalized LLM AlignmentHongru Cai, Yongqi Li, Tiezheng Yu, Fengbin Zhu 等SIGIR 2026
- Personalized LoRA for Human-Centered Text UnderstandingYou Zhang, Jin Wang, Liang-Chih Yu, Dan Xu 等AAAI 2024 · 被引用 22 次
- PRISP: Privacy-Safe Few-Shot Personalization via Lightweight AdaptationJunho Park, Dohoon Kim, Taesup MoonACL 2026 · 被引用 1 次
- Unraveling LoRA Interference: Orthogonal Subspaces for Robust Model MergingHaobo Zhang, Jiayu ZhouACL 2025
- Heterogeneous Customizable Personalized Federated Fine-Tuning Approach for Large Language Modelsxin tong, Baojiang cuiICML 2026 · 被引用 199 次
