Federated Learning from Vision-Language Foundation Models: Theoretical Analysis and Method
Bikang Pan, Wei Huang, Ye Shi
摘要
Integrating pretrained vision-language foundation models like CLIP into federated learning has attracted significant attention for enhancing generalization across diverse tasks. Typically, federated learning of vision-language models employs prompt learning to reduce communication and computational costs, i.e., promptbased federated learning. However, there is limited theoretical analysis to understand the performance of prompt-based federated learning. In this work, we construct a theoretical analysis framework for prompt-based federated learning via feature learning theory. Specifically, we monitor the evolution of signal learning and noise memorization in prompt-based federated learning, demonstrating that performance can be assessed by the ratio of task-relevant to task-irrelevant coefficients. Furthermore, we draw an analogy between income and risk in portfolio optimization and the task-relevant and task-irrelevant terms in feature learning. Leveraging inspiration from portfolio optimization that combining two independent assets will maintain the income while reducing the risk, we introduce two prompts: global prompt and local prompt to construct a prompt portfolio to balance the generalization and personalization. Consequently, we showed the performance advantage of the prompt portfolio and derived the optimal mixing coefficient. These theoretical claims have been further supported by empirical experiments. Our code is available at: https://github.com/PanBikang/PromptFolio.git .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper9
- Latte: Collaborative Test-Time Adaptation of Vision-Language Models in Federated LearningWenxuan Bao, Ruxi Deng, Ruizhong Qiu, Tianxin Wei 等ICCV 2025 · 被引用 13 次
- Global Prompt Refinement with Non-Interfering Attention Masking for One-Shot Federated LearningZhuang Qi, Pan Yu, Lei Meng, Sijin Zhou 等NeurIPS 2025 · 被引用 4 次
- Federated Vision-Language-Recommendation with Personalized FusionZhiwei Li, Guodong Long, Jing Jiang, Chengqi Zhang 等AAAI 2026 · 被引用 3 次
- pFedMMA: Personalized Federated Fine-Tuning with Multi-Modal Adapter for Vision-Language ModelsSajjad Ghiasvand, Mahnoosh Alizadeh, Ramtin PedarsaniICLR 2026 · 被引用 3 次
- C2Prompt: Class-aware Client Knowledge Interaction for Federated Continual LearningKunlun Xu, Yibo Feng, Jiangmeng Li, Yongsheng Qi 等NeurIPS 2025 · 被引用 2 次
它引用的顶会 Paper25
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Moment Matching for Multi-Source Domain AdaptationXingchao Peng, Qinxun Bai, Xide Xia, Zijun Huang 等ICCV 2019 · 被引用 2,239 次
- Personalized Cross-Silo Federated Learning on Non-IID DataYutao Huang, Lingyang Chu, Zirui Zhou, Lanjun Wang 等AAAI 2021 · 被引用 816 次
- What Makes Multi-Modal Learning Better than Single (Provably)Yu Huang, Chenzhuang Du, Zihui Xue, Xuanyao Chen 等NeurIPS 2021 · 被引用 404 次
- Toward Understanding the Feature Learning Process of Self-supervised Contrastive LearningZixin Wen, Yuanzhi LiICML 2021 · 被引用 162 次
相关 Paper
- Harmonizing Generalization and Personalization in Federated Prompt LearningTianyu Cui, Hongxia Li, Jingya Wang, Ye ShiICML 2024 · 被引用 31 次
- FedPHA: Federated Prompt Learning for Heterogeneous Client AdaptationChengying Fang, Wenke Huang, Guancheng Wan, Yihao Yang 等ICML 2025
- Mixture of Experts Made Personalized: Federated Prompt Learning for Vision-Language ModelsJun Luo, Chen Chen, Shandong WuICLR 2025
- FedDEAP: Adaptive Dual-Prompt Tuning for Multi-Domain Federated LearningYubin Zheng, Pak-Hei Yeung, Jing Xia, Tianjie Ju 等ACM MM 2025
- pFedPrompt: Learning Personalized Prompt for Vision-Language Models in Federated LearningTao Guo, Song Guo, Junxiao WangWWW 2023 · 被引用 101 次
