The Influences of Task Design on Crowdsourced Judgement: A Case Study of Recidivism Risk Evaluation
Xiaoni Duan, Chien-Ju Ho, Ming Yin
摘要
Crowdsourcing is widely used to solicit judgement from people in diverse applications ranging from evaluating information quality to rating gig worker performance. To encourage the crowd to put in genuine effort in the judgement tasks, various ways to structure and organize these tasks have been explored, though the understandings of how these task design choices influence the crowd's judgement are still largely lacking. In this paper, using recidivism risk evaluation as an example, we conduct a randomized experiment to examine the effects of two common designs of crowdsourcing judgement tasks-encouraging the crowd to deliberate and providing feedback to the crowd-on the quality, strictness, and fairness of the crowd's recidivism risk judgements. Our results show that different designs of the judgement tasks significantly affect the strictness of the crowd's judgements. Moreover, task designs also have the potential to significantly influence how fairly the crowd judges defendants from different racial groups, on those cases where the crowd exhibits substantial in-group bias. Finally, we find that the impacts of task designs on the judgement also vary with the crowd workers' own characteristics, such as their cognitive reflection levels. Together, these results highlight the importance of obtaining a nuanced understanding on the relationship between task designs and properties of the crowdsourced judgements. CCS CONCEPTS • Human-centered computing → Empirical studies in HCI; Empirical studies in collaborative and social computing.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Are Two Heads Better Than One in AI-Assisted Decision Making? Comparing the Behavior and Performance of Groups and Individuals in Human-AI Collaborative Recidivism Risk AssessmentChun-Wei Chiang, Zhuoran Lu, Zhuoyan Li, Ming YinCHI 2023 · 被引用 56 次
- Encoding Human Behavior in Information Design through Deep LearningGuanghui Yu, Wei Tang, Saumik Narayanan, Chien-Ju HoNeurIPS 2023 · 被引用 8 次
它引用的顶会 Paper5
- "Everyone wants to do the model work, not the data work": Data Cascades in High-Stakes AINithya Sambasivan, Shivani Kapania, Hannah Highfill, Diana Akrong 等CHI 2021 · 被引用 725 次
- Algorithmic Risk Assessments Can Alter Human Decision-Making Processes in High-Stakes Government ContextsBen Green, Yiling ChenCSCW 2021 · 被引用 63 次
- Can Online Juries Make Consistent, Repeatable Decisions?Xinlan Emily Hu, Mark E. Whiting, Michael S. BernsteinCHI 2021 · 被引用 60 次
- Reputation Agent: Prompting Fair Reviews in Gig MarketsCarlos Toxtli, Angela Richmond-Fuller, Saiph SavageWWW 2020 · 被引用 53 次
- Do I Look Like a Criminal? Examining how Race Presentation Impacts Human Judgement of RecidivismKeri Mallari, Kori Inkpen, Paul Johns, Sarah Tan 等CHI 2020 · 被引用 23 次
相关 Paper
- The Impact of Algorithmic Risk Assessments on Human Predictions and its Analysis via Crowdsourcing StudiesRiccardo Fogliato, Alexandra Chouldechova, Zachary C. LiptonCSCW 2021 · 被引用 25 次
- Hardhats and Bungaloos: Comparing Crowdsourced Design Feedback with Peer Design Feedback in the ClassroomJonas Oppenlaender, Elina Kuosmanen, Andrés Lucero, Simo HosioCHI 2021 · 被引用 12 次
- What Ingredients Make for an Effective Crowdsourcing Protocol for Difficult NLU Data Collection Tasks?Nikita Nangia, Saku Sugawara, Harsh Trivedi, Alex Warstadt 等ACL 2021
- Aligning Crowdworker Perspectives and Feedback Outcomes in Crowd-Feedback System DesignSaskia Haug, Ivo Benke, Alexander MaedcheCSCW 2023 · 被引用 9 次
- Estimating Conversational Styles in Conversational Microtask CrowdsourcingSihang Qiu, Ujwal Gadiraju, Alessandro BozzonCSCW 2020 · 被引用 19 次
