LabelAId: Just-in-time AI Interventions for Improving Human Labeling Quality and Domain Knowledge in Crowdsourcing Systems
Chu Li, Zhihan Zhang, Michael Saugstad, Esteban Safranchik, Chaitanyashareef Kulkarni, Xiaoyu Huang, Shwetak N. Patel, Vikram Iyer, Tim Althoff, Jon E. Froehlich
摘要
Crowdsourcing platforms have transformed distributed problem-solving, yet quality control remains a persistent challenge. Traditional quality control measures, such as prescreening workers and refining instructions, often focus solely on optimizing economic output. This paper explores just-in-time AI interventions to enhance both labeling quality and domain-specific knowledge among crowdworkers. We introduce LabelAId, an advanced inference model combining Programmatic Weak Supervision (PWS) with FT-Transformers to infer label correctness based on user behavior and domain knowledge. Our technical evaluation shows that our LabelAId pipeline consistently outperforms state-of-the-art ML baselines, improving mistake inference accuracy by 36.7% with 50 downstream samples. We then implemented LabelAId into Project Sidewalk, an open-source crowdsourcing platform for urban accessibility. A between-subjects study with 34 participants demonstrates that LabelAId significantly enhances label precision without compromising efficiency while also increasing labeler confidence. We discuss LabelAId’s success factors, limitations, and its generalizability to other crowdsourced science domains.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Living Sustainability: In-Context Interactive Environmental Impact CommunicationZhihan Zhang, Puvarin Thavikulwat, Alexander Le Metzger, Yuxuan Mei 等UbiComp 2025 · 被引用 5 次
- CoKnowledge: Supporting Assimilation of Time-synced Collective Knowledge in Online Science VideosYuanhao Zhang, Yumeng Wang, Xiyuan Wang, Changyang He 等CHI 2025 · 被引用 4 次
它引用的顶会 Paper11
- MLP-Mixer: An all-MLP Architecture for VisionIlya O. Tolstikhin, Neil Houlsby, Alexander Kolesnikov, Lucas Beyer 等NeurIPS 2021 · 被引用 3,862 次
- Revisiting Deep Learning Models for Tabular DataYury Gorishniy, Ivan Rubachev, Valentin Khrulkov, Artem BabenkoNeurIPS 2021 · 被引用 1,847 次
- To Trust or to Think: Cognitive Forcing Functions Can Reduce Overreliance on AI in AI-assisted Decision-makingZana Buçinca, Maja Barbara Malaya, Krzysztof Z. GajosCSCW 2021 · 被引用 962 次
- Does the Whole Exceed its Parts? The Effect of AI Explanations on Complementary Team PerformanceGagan Bansal, Tongshuang Wu, Joyce Zhou, Raymond Fok 等CHI 2021 · 被引用 713 次
- On the Use of Multi-sensory Cues in Symmetric and Asymmetric Shared Collaborative Virtual SpacesSungchul Jung, Nawam Karki, Max W. J. Slutter, Robert W. LindemanCSCW 2021 · 被引用 33 次
相关 Paper
- "I never realized sidewalks were a big deal": A Case Study of a Community-Driven Sidewalk Accessibility Assessment using Project SidewalkChu Li, Katrina Oi Yau Ma, Michael Saugstad, Kie Fujii 等CHI 2024 · 被引用 12 次
- Refining Labeling Functions with Limited Labeled DataChenjie Li, Amir Gilad, Boris Glavic, Zhengjie Miao 等KDD 2025 · 被引用 1 次
- Learning Hyper Label Model for Programmatic Weak SupervisionRenzhi Wu, Shen-En Chen, Jieyu Zhang, Xu ChuICLR 2023 · 被引用 2 次
- Weak Supervision Performance Evaluation via Partial IdentificationFelipe Maia Polo, Subha Maity, Mikhail Yurochkin, Moulinath Banerjee 等NeurIPS 2024 · 被引用 6 次
- Active Label Correction for Semantic Segmentation with Foundation ModelsHoyoung Kim, Sehyun Hwang, Suha Kwak, Jungseul OkICML 2024 · 被引用 5 次
