Improving Fairness for Data Valuation in Horizontal Federated Learning
Zhenan Fan, Huang Fang, Zirui Zhou, Jian Pei, Michael P. Friedlander, Changxin Liu, Yong Zhang
摘要
Federated learning is an emerging decentralized machine learning scheme that allows multiple data owners to work collaboratively while ensuring data privacy. The success of federated learning depends largely on the participation of data owners. To sustain and encourage data owners' participation, it is crucial to fairly evaluate the quality of the data provided by the data owners as well as their contribution to the final model and reward them correspondingly. Federated Shapley value, recently proposed by Wang et al. [Federated Learning, 2020], is a measure for data value under the framework of federated learning that satisfies many desired properties for data valuation. However, there are still factors of potential unfairness in the design of federated Shapley value because two data owners with the same local data may not receive the same evaluation. We propose a new measure called completed federated Shapley value to improve the fairness of federated Shapley value. The design depends on completing a matrix consisting of all the possible contributions by different subsets of the data owners. It is shown under mild conditions that this matrix is approximately low-rank by leveraging concepts and tools from optimization. Both theoretical analysis and empirical evaluation verify that the proposed measure does improve fairness in many circumstances.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Fair and Efficient Contribution Valuation for Vertical Federated LearningZhenan Fan, Huang Fang, Xinglu Wang, Zirui Zhou 等ICLR 2024 · 被引用 33 次
- PeFAD: A Parameter-Efficient Federated Framework for Time Series Anomaly DetectionRonghui Xu, Hao Miao, Senzhang Wang, Philip S. Yu 等KDD 2024 · 被引用 32 次
- Contributions Estimation in Federated Learning: A Comprehensive Experimental EvaluationYiwei Chen, Kaiyu Li, Guoliang Li, Yong WangVLDB 2024 · 被引用 17 次
- Secure and Verifiable Data Collaboration with Low-Cost Zero-Knowledge ProofsYizheng Zhu, Yuncheng Wu, Zhaojing Luo, Beng Chin Ooi 等VLDB 2024 · 被引用 14 次
- Fairness in model-sharing gamesKate Donahue, Jon M. KleinbergWWW 2023 · 被引用 12 次
它引用的顶会 Paper4
- On the Convergence of FedAvg on Non-IID DataXiang Li, Kaixuan Huang, Wenhao Yang, Shusen Wang 等ICLR 2020 · 被引用 2,930 次
- Fair Resource Allocation in Federated LearningTian Li, Maziar Sanjabi, Ahmad Beirami, Virginia SmithICLR 2020 · 被引用 971 次
- Data Valuation using Reinforcement LearningJinsung Yoon, Sercan Ömer Arik, Tomas PfisterICML 2020 · 被引用 236 次
- A Distributional Framework For Data ValuationAmirata Ghorbani, Michael P. Kim, James ZouICML 2020 · 被引用 152 次
相关 Paper
- FairFed: Improving Fairness and Efficiency of Contribution Evaluation in Federated Learning via Cooperative Shapley ValueYiqi Liu, Shan Chang, Ye Liu, Bo Li 等INFOCOM 2024 · 被引用 20 次
- Efficient Participant Contribution Evaluation for Horizontal and Vertical Federated LearningJunhao Wang, Lan Zhang, Anran Li, Xuanke You 等ICDE 2022 · 被引用 40 次
- Efficient Data Valuation Approximation in Federated Learning: A Sampling-Based ApproachShuyue Wei, Yongxin Tong, Zimu Zhou, Tianran He 等ICDE 2025
- Ripple Shapley: Data Influence Attribution in One Federated Training RunDewen Zeng, Wenlong Tian, Haozhao Wang, Jianfeng Lu 等AAAI 2026 · 被引用 1 次
- Collaborative Causal Inference with Fair IncentivesRui Qiao, Xinyi Xu, Bryan Kian Hsiang LowICML 2023 · 被引用 8 次
