DoFIT: Domain-aware Federated Instruction Tuning with Alleviated Catastrophic Forgetting
Binqian Xu, Xiangbo Shu, Haiyang Mei, Zechen Bai, Basura Fernando, Mike Zheng Shou, Jinhui Tang
摘要
Federated Instruction Tuning (FIT) advances collaborative training on decentralized data, crucially enhancing model’s capability and safeguarding data privacy. However, existing FIT methods are dedicated to handling data heterogeneity across different clients (i.e., client-aware data heterogeneity), while ignoring the variation between data from different domains (i.e., domain-aware data heterogeneity). When scarce data needs supplementation from related fields, these methods lack the ability to handle domain heterogeneity in cross-domain training. This leads to domain-information catastrophic forgetting in collaborative training and therefore makes model perform sub-optimally on the individual domain. To address this issue, we introduce DoFIT , a new Do main-aware FIT framework that alleviates catastrophic forgetting through two new designs. First, to reduce interference information from the other domain, DoFIT finely aggregates overlapping weights across domains on the inter-domain server side. Second, to retain more domain information, DoFIT initializes intra-domain weights by incorporating inter-domain information into a less-conflicted parameter space. Experimental results on diverse datasets consistently demonstrate that DoFIT excels in cross-domain collaborative training and exhibits significant advantages over conventional FIT methods in alleviating catastrophic forgetting.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper7
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida 等NeurIPS 2022 · 被引用 24,707 次
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- Few-Shot Parameter-Efficient Fine-Tuning is Better and Cheaper than In-Context LearningHaokun Liu, Derek Tam, Mohammed Muqeeth, Jay Mohta 等NeurIPS 2022 · 被引用 1,483 次
- Dual-Personalizing Adapter for Federated Foundation ModelsYiyuan Yang, Guodong Long, Tao Shen, Jing Jiang 等NeurIPS 2024 · 被引用 84 次
- Principled Federated Domain Adaptation: Gradient Projection and Auto-WeightingEnyi Jiang, Yibo Jacky Zhang, Sanmi KoyejoICLR 2024 · 被引用 10 次
相关 Paper
- Learn from Others and Be Yourself in Heterogeneous Federated LearningWenke Huang, Mang Ye, Bo DuCVPR 2022 · 被引用 254 次
- Splitting with Importance-aware Updating for Heterogeneous Federated Learning with Large Language ModelsYangxu Liao, Wenke Huang, Guancheng Wan, Jian Liang 等ICML 2025
- Decoupling Shared and Personalized Knowledge: A Dual-Branch Federated Learning Framework for Multi-Domain with Non-IID DataYiran Pang, Zhen Ni, Xiangnan ZhongAAAI 2026
- Towards Efficient Replay in Federated Incremental LearningYichen Li, Qunwei Li, Haozhao Wang, Ruixuan Li 等CVPR 2024
- Towards Fast and Stable Federated Learning: Confronting Heterogeneity via Knowledge AnchorJinqian Chen, Jihua Zhu, Qinghai ZhengACM MM 2023 · 被引用 5 次
