FedSum: Data-Efficient Federated Learning Under Data Scarcity Scenario for Text Summarization
Zhiyong Ma, Zhengping Li, Yuanjie Shi, Jian Chen
摘要
Text summarization task extracts salient information from a large amount of text for productivity enhancement. However, most existing methods heavily rely on training models from ample and centrally stored data which is infeasible to collect in practice, due to privacy concerns and data scarcity nature under several settings (e.g., edge computing or cold starting). The main challenge lies in constructing the privacy-preserving and well-behaved summarization model under the data scarcity scenario, where the data scarcity nature will lead to the knowledge shortage of the model while magnifying the impact of data bias, causing performance degeneration. To tackle this challenge, previous studies attempt to complement samples or improve the efficiency of data. The former is usually associated with high computing costs or has a large dependence on empirical settings, while the latter might not effective due to the lack of consideration of data bias. In this work, we propose FedSum which extends the standard FL framework from depth and breadth to further extract prime and diversified knowledge from limited resources for text summarization. For depth extension, we introduce a Data Partition method to cooperatively recognize the prime samples that are more significant and unbiased, and the Data skip mechanism is introduced to help the model further focus on those prime samples during the local training process. For breadth extension, FedSum extends the source of knowledge and develops the summarization model by extracting knowledge from the data samples, hidden spaces, and globally received parameters. Extensive experiments on four benchmark datasets verify the promising improvement of FedSum compared to baselines, and show its generalizability, scalability, and robustness.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper23
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi 等ICML 2020 · 被引用 3,875 次
- Federated Learning on Non-IID Data Silos: An Experimental StudyQinbin Li, Yiqun Diao, Quan Chen, Bingsheng HeICDE 2022 · 被引用 1,110 次
- Exploiting Shared Representations for Personalized Federated LearningLiam Collins, Hamed Hassani, Aryan Mokhtari, Sanjay ShakkottaiICML 2021 · 被引用 1,081 次
- FedProto: Federated Prototype Learning across Heterogeneous ClientsYue Tan, Guodong Long, Lu Liu, Tianyi Zhou 等AAAI 2022 · 被引用 851 次
- Prototypical Contrastive Learning of Unsupervised RepresentationsJunnan Li, Pan Zhou, Caiming Xiong, Steven C. H. HoiICLR 2021 · 被引用 484 次
相关 Paper
- FedED: Federated Learning via Ensemble Distillation for Medical Relation ExtractionDianbo Sui, Yubo Chen, Jun Zhao, Yantao Jia 等EMNLP 2020 · 被引用 126 次
- Flick: Empowering Federated Learning with Commonsense KnowledgeRan Zhu, Mingkun Yang, Shiqiang Wang, Jie Yang 等NeurIPS 2025
- Few-shot Query-Focused Summarization with Prefix-MergingRuifeng Yuan, Zili Wang, Ziqiang Cao, Wenjie LiEMNLP 2022 · 被引用 6 次
- FLea: Addressing Data Scarcity and Label Skew in Federated Learning via Privacy-preserving Feature AugmentationTong Xia, Abhirup Ghosh, Xinchi Qiu, Cecilia MascoloKDD 2024 · 被引用 4 次
- DapperFL: Domain Adaptive Federated Learning with Model Fusion Pruning for Edge DevicesYongzhe Jia, Xuyun Zhang, Hongsheng Hu, Kim-Kwang Raymond Choo 等NeurIPS 2024 · 被引用 14 次
