Data Valuation and Detections in Federated Learning
Wenqian Li, Shuran Fu, Fengrui Zhang, Yan Pang
摘要
Federated Learning (FL) enables collaborative model training while preserving the privacy of raw data. A challenge in this framework is the fair and efficient valuation of data, which is crucial for incentivizing clients to contribute high-quality data in the FL task. In scenarios involving numerous data clients within FL, it is often the case that only a subset of clients and datasets are pertinent to a specific learning task, while others might have either a negative or negligible impact on the model training process. This paper introduces a novel privacy-preserving method for evaluating client contributions and selecting relevant datasets without a pre-specified training algorithm in an FL task. Our proposed approach, FedBary, utilizes Wasserstein distance within the federated context, offering a new solution for data valuation in the FL framework. This method ensures transparent data valuation and efficient computation of the Wasserstein barycenter and reduces the dependence on validation datasets. Through extensive empirical experiments and theoretical analyses, we demonstrate the advantages of this data valuation method as a promising avenue for FL research. Codes are available at https://github.com/muz1lee/MOTdata .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- OpenFGL: A Comprehensive Benchmark for Federated Graph LearningXunkai Li, Yinlin Zhu, Boyang Pang, Guochen Yan 等VLDB 2025 · 被引用 12 次
- From Points to Coalitions: Hierarchical Contrastive Shapley Values for Prioritizing Data SamplesCanran Xiao, Jiabao Dou, Zhiming Lin, Zong Ke 等AAAI 2026 · 被引用 7 次
- Fed-ADE: Adaptive Learning Rate for Federated Post-adaptation under Distribution ShiftHeewon Park, Mugon Joe, Miru Kim, Kyungjin Im 等CVPR 2026 · 被引用 2 次
- When Sample Selection Bias Precipitates Model CollapseXinbao Qiao, Xianglong Du, Wei Liu, Jingqi Zhang 等ICML 2026
- FedRAC: Rolling Submodel Allocation for Collaborative Fairness in Federated LearningZihui Wang, Yuhang Fu, Mengmeng Du, Zhimin Yuan 等CVPR 2026
它引用的顶会 Paper13
- Estimating Training Data Influence by Tracing Gradient DescentGarima Pruthi, Frederick Liu, Satyen Kale, Mukund SundararajanNeurIPS 2020 · 被引用 784 次
- What Neural Networks Memorize and Why: Discovering the Long Tail via Influence EstimationVitaly Feldman, Chiyuan ZhangNeurIPS 2020 · 被引用 674 次
- Geometric Dataset Distances via Optimal TransportDavid Alvarez-Melis, Nicolò FusiNeurIPS 2020 · 被引用 267 次
- Data Valuation using Reinforcement LearningJinsung Yoon, Sercan Ömer Arik, Tomas PfisterICML 2020 · 被引用 236 次
- Robust Federated Learning: The Case of Affine Distribution ShiftsAmirhossein Reisizadeh, Farzan Farnia, Ramtin Pedarsani, Ali JadbabaieNeurIPS 2020 · 被引用 196 次
相关 Paper
- Fortifying Federated Learning Towards Trustworthiness via Auditable Data Valuation and Verifiable Client ContributionK. Naveen Kumar, Ranjeet Ranjan Jha, C. Krishna Mohan, Ravindra Babu TallamrajuCVPR 2025
- Is Your Data Relevant?: Dynamic Selection of Relevant Data for Federated LearningLokesh Nagalapatti, Ruhi Sharma Mittal, Ramasuri NarayanamAAAI 2022 · 被引用 32 次
- Contributions Estimation in Federated Learning: A Comprehensive Experimental EvaluationYiwei Chen, Kaiyu Li, Guoliang Li, Yong WangVLDB 2024 · 被引用 17 次
- PriCAF: Privacy-Preserving Contribution Assessment in Federated Learning Before Model TrainingYixin Xu, Hao Wu, Jingzhou Zhu, Fengyuan Xu 等ACM MM 2025
- CoAst: Validation-Free Contribution Assessment for Federated Learning based on Cross-Round ValuationHao Wu, Likun Zhang, Shucheng Li, Fengyuan Xu 等ACM MM 2024 · 被引用 2 次
