Incentivizing Truthfulness and Collaborative Fairness in Bayesian Learning
Rachael Hwee Ling Sim, Jue Fan, Xiao Tian, Xinyi Xu, Patrick Jaillet, Bryan Kian Hsiang Low
摘要
Collaborative machine learning involves training high-quality models using datasets from a number of sources. To incentivize sources to share data, existing data valuation methods fairly reward each source based on its data submitted as is. However, as these methods do not verify nor incentivize data truthfulness, the sources can manipulate their data (e.g., by submitting duplicated or noisy data) to artificially increase their valuations and rewards or prevent others from benefiting. This paper presents the first mechanism that provably ensures ( F ) collaborative fairness and incentivizes ( T ) truthfulness at equilibrium for Bayesian models. Our mechanism combines semivalues (e.g., Shapley value), which ensure fairness, and a truthful data valuation function (DVF) based on a validation set that is unknown to the sources. As semivalues are influenced by others' data, we introduce an additional condition to prove that a source can maximize its expected data values in coalitions and semivalues by submitting a dataset that captures its true knowledge. Additionally, we discuss the implications and suitable relaxations of ( F ) and ( T ) when the mediator has a limited budget for rewards or lacks a validation set. Our theoretical findings are validated on synthetic and real-world datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper12
- Hidden Trigger Backdoor AttacksAniruddha Saha, Akshayvarun Subramanya, Hamed PirsiavashAAAI 2020 · 被引用 743 次
- Collaborative Machine Learning with Incentive-Aware Model RewardsRachael Hwee Ling Sim, Yehong Zhang, Mun Choon Chan, Bryan Kian Hsiang LowICML 2020 · 被引用 158 次
- Gradient Driven Rewards to Guarantee Fairness in Collaborative Machine LearningXinyi Xu, Lingjuan Lyu, Xingjun Ma, Chenglin Miao 等NeurIPS 2021 · 被引用 133 次
- Validation Free and Replication Robust Volume-based Data ValuationXinyi Xu, Zhaoxuan Wu, Chuan Sheng Foo, Bryan Kian Hsiang LowNeurIPS 2021 · 被引用 89 次
- Approximating the Shapley Value without Marginal ContributionsPatrick Kolpaczki, Viktor Bengs, Maximilian Muschalik, Eyke HüllermeierAAAI 2024 · 被引用 43 次
相关 Paper
- Incentives in Private Collaborative Machine LearningRachael Hwee Ling Sim, Yehong Zhang, Nghia Hoang, Xinyi Xu 等NeurIPS 2023 · 被引用 12 次
- Collaborative Causal Inference with Fair IncentivesRui Qiao, Xinyi Xu, Bryan Kian Hsiang LowICML 2023 · 被引用 8 次
- Truthful Data Acquisition via Peer PredictionYiling Chen, Yiheng Shen, Shuran ZhengNeurIPS 2020 · 被引用 35 次
- A Cramér-von Mises Approach to Incentivizing Truthful Data SharingAlex Clinton, Thomas Zeng, Yiding Chen, Xiaojin Zhu 等NeurIPS 2025 · 被引用 2 次
- FedServing: A Federated Prediction Serving Framework Based on Incentive MechanismJiasi Weng, Jian Weng, Hongwei Huang, Chengjun Cai 等INFOCOM 2021 · 被引用 37 次
