Quality Control in Crowdsourcing based on Fine-Grained Behavioral Features
Weiping Pei, Zhiju Yang, Monchu Chen, Chuan Yue
摘要
Crowdsourcing is popular for large-scale data collection and labeling, but a major challenge is on detecting low-quality submissions. Recent studies have demonstrated that behavioral features of workers are highly correlated with data quality and can be useful in quality control. However, these studies primarily leveraged coarsely extracted behavioral features, and did not further explore quality control at the fine-grained level, i.e., the annotation unit level. In this paper, we investigate the feasibility and benefits of using fine-grained behavioral features, which are the behavioral features finely extracted from a worker's individual interactions with each single unit in a subtask, for quality control in crowdsourcing. We design and implement a framework named Fine-grained Behavior-based Quality Control (FBQC) that specifically extracts fine-grained behavioral features to provide three quality control mechanisms: (1) quality prediction for objective tasks, (2) suspicious behavior detection for subjective tasks, and (3) unsupervised worker categorization. Using the FBQC framework, we conduct two real-world crowdsourcing experiments and demonstrate that using fine-grained behavioral features is feasible and beneficial in all three quality control mechanisms. Our work provides clues and implications for helping job requesters or crowdsourcing platforms to further achieve better quality control.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper2
- The State of Pilot Study Reporting in Crowdsourcing: A Reflection on Best Practices and GuidelinesJonas Oppenlaender, Tahir Abbas, Ujwal GadirajuCSCW 2024 · 被引用 10 次
- Verifying or Clarifying? User Preferences for Mobile Crowdsourcing in Response to Seemingly Inconsistent Sensor DataYou-Hsuan Chiang, Je-Wei Hsu, Hsin-Lun Chiu, Chung-En Liu 等CSCW 2025
相关 Paper
- Improving Data Quality via Pre-Task Participant Screening in Crowdsourced GUI ExperimentsTakaya Miyama, Satoshi Nakamura, Shota YamanakaCHI 2026 · 被引用 2 次
- CrowdMOT: Crowdsourcing Strategies for Tracking Multiple Objects in VideosSamreen Anjum, Chi Lin, Danna GurariCSCW 2020 · 被引用 5 次
- Detecting and Preventing Confused Labels in Crowdsourced DataEvgeny Krivosheev, Siarhei Bykau, Fabio Casati, Sunil PrabhakarVLDB 2020 · 被引用 12 次
- Label Aggregation for Composite Crowd Tasks by Worker Ability Constraint SatisfactionJiyi LiAAAI 2025 · 被引用 1 次
- A Probabilistic Graphical Model for Analyzing the Subjective Visual Quality Assessment Data from CrowdsourcingJing Li, Suiyi Ling, Junle Wang, Patrick Le CalletACM MM 2020 · 被引用 23 次
