DoFIT: Domain-aware Federated Instruction Tuning with Alleviated Catastrophic Forgetting
Binqian Xu, Xiangbo Shu, Haiyang Mei, Zechen Bai, Basura Fernando, Mike Zheng Shou, Jinhui Tang
Abstract
Federated Instruction Tuning (FIT) advances collaborative training on decentralized data, crucially enhancing model’s capability and safeguarding data privacy. However, existing FIT methods are dedicated to handling data heterogeneity across different clients (i.e., client-aware data heterogeneity), while ignoring the variation between data from different domains (i.e., domain-aware data heterogeneity). When scarce data needs supplementation from related fields, these methods lack the ability to handle domain heterogeneity in cross-domain training. This leads to domain-information catastrophic forgetting in collaborative training and therefore makes model perform sub-optimally on the individual domain. To address this issue, we introduce DoFIT , a new Do main-aware FIT framework that alleviates catastrophic forgetting through two new designs. First, to reduce interference information from the other domain, DoFIT finely aggregates overlapping weights across domains on the inter-domain server side. Second, to retain more domain information, DoFIT initializes intra-domain weights by incorporating inter-domain information into a less-conflicted parameter space. Experimental results on diverse datasets consistently demonstrate that DoFIT excels in cross-domain collaborative training and exhibits significant advantages over conventional FIT methods in alleviating catastrophic forgetting.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on7
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- Few-Shot Parameter-Efficient Fine-Tuning is Better and Cheaper than In-Context LearningHaokun Liu, Derek Tam, Mohammed Muqeeth, Jay Mohta et al.NeurIPS 2022 · 1,483 citations
- Dual-Personalizing Adapter for Federated Foundation ModelsYiyuan Yang, Guodong Long, Tao Shen, Jing Jiang et al.NeurIPS 2024 · 84 citations
- Principled Federated Domain Adaptation: Gradient Projection and Auto-WeightingEnyi Jiang, Yibo Jacky Zhang, Sanmi KoyejoICLR 2024 · 10 citations
Related papers
- Learn from Others and Be Yourself in Heterogeneous Federated LearningWenke Huang, Mang Ye, Bo DuCVPR 2022 · 254 citations
- Splitting with Importance-aware Updating for Heterogeneous Federated Learning with Large Language ModelsYangxu Liao, Wenke Huang, Guancheng Wan, Jian Liang et al.ICML 2025
- Decoupling Shared and Personalized Knowledge: A Dual-Branch Federated Learning Framework for Multi-Domain with Non-IID DataYiran Pang, Zhen Ni, Xiangnan ZhongAAAI 2026
- Towards Efficient Replay in Federated Incremental LearningYichen Li, Qunwei Li, Haozhao Wang, Ruixuan Li et al.CVPR 2024
- Towards Fast and Stable Federated Learning: Confronting Heterogeneity via Knowledge AnchorJinqian Chen, Jihua Zhu, Qinghai ZhengACM MM 2023 · 5 citations
