Towards Fair Truth Discovery from Biased Crowdsourced Answers
Yanying Li, Haipei Sun, Wendy Hui Wang
摘要
Crowdsourcing systems have gained considerable interest and adoption in recent years. One important research problem for crowdsourcing systems is truth discovery, which aims to aggregate noisy answers contributed by the workers to obtain the correct answer (truth) of each task. However, since the collected answers are highly prone to the workers' biases, aggregating these biased answers without proper treatment will unavoidably lead to discriminatory truth discovery results for particular race, gender and political groups. To address this challenge, in this paper, first, we define a new fairness notion named θ-disparity for truth discovery. Intuitively, θ-disparity bounds the difference in the probabilities that the truth of both protected and unprotected groups being predicted to be positive. Second, we design three fairness enhancing methods, namely Pre-TD, FairTD, and Post-TD, for truth discovery. Pre-TD is a pre-processing method that removes the bias in workers' answers before truth discovery. FairTD is an in-processing method that incorporates fairness into the truth discovery process. And Post-TD is a post-processing method that applies additional treatment on the discovered truth to make it satisfy θ-disparity. We perform an extensive set of experiments on both synthetic and real-world crowdsourcing datasets. Our results demonstrate that among the three fairness enhancing methods, FairTD produces the best accuracy with θ-disparity. In some settings, the accuracy of FairTD is even better than truth discovery without fairness, as it removes some low-quality answers as side effects.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper6
- Crowdsourcing Subjective Annotations Using Pairwise Comparisons Reduces Bias and Error Compared to the Majority-vote MethodHasti Narimanzadeh, Arash Badie Modiri, Iuliia G. Smirnova, Ted Hsuan Yun ChenCSCW 2023 · 被引用 20 次
- Chameleon: Foundation Models for Fairness-aware Multi-modal Data Augmentation to Enhance Coverage of MinoritiesMahdi Erfanian, H. V. Jagadish, Abolfazl AsudehVLDB 2024 · 被引用 10 次
- Product Question Answering in E-Commerce: A SurveyYang Deng, Wenxuan Zhang, Qian Yu, Wai LamACL 2023 · 被引用 9 次
- Noisy Interactive Graph SearchQianhao Cong, Jing Tang, Kai Han, Yuming Huang 等KDD 2022 · 被引用 4 次
- Optimal Fair Aggregation of Crowdsourced Noisy Labels using Demographic Parity ConstraintsGabriel Singer, Samuel Gruffaz, Olivier VO VAN, Nicolas Vayatis 等ICML 2026
相关 Paper
- RCTD: Reputation-Constrained Truth Discovery in Sybil Attack Crowdsourcing EnvironmentXing Jin, Zhihai Gong, Jiuchuan Jiang, Chao Wang 等KDD 2024 · 被引用 2 次
- Towards Personalized Privacy-Preserving Incentive for Truth Discovery in Crowdsourced Binary-Choice Question AnsweringPeng Sun, Zhibo Wang, Yunhe Feng, Liantao Wu 等INFOCOM 2020 · 被引用 39 次
- Frustratingly Easy Truth DiscoveryReshef Meir, Ofra Amir, Omer Ben-Porat, Tsviel Ben Shabat 等AAAI 2023 · 被引用 2 次
- Can The Crowd Identify Misinformation Objectively?: The Effects of Judgment Scale and Assessor's BackgroundKevin Roitero, Michael Soprano, Shaoyang Fan, Damiano Spina 等SIGIR 2020 · 被引用 2 次
- Fair Sequential Selection Using Supervised Learning ModelsMohammad Mahdi Khalili, Xueru Zhang, Mahed AbroshanNeurIPS 2021 · 被引用 25 次
