Constructing a Classification Scheme - and its Consequences: A Field Study of Learning to Label Data for Computer Vision in a Hospital Intensive Care Unit
Melissa A. Valentine, Roger E. Bohn, Amanda L. Pratt, Prachee Jain, Sara J. Singer, Michael S. Bernstein
摘要
Research on data annotation for artificial intelligence (AI) has demonstrated that biases, power, and culture impact the ways that annotators apply labels to data and subsequently affect downstream AI systems. However, annotators can only apply labels that are available to them in the annotation classification scheme. Drawing on a 3-year ethnographic study of an R&D collaboration between medical and AI researchers, we argue that the construction of the classification schema itself -- decisions about what kinds of data can and cannot be collected, what activities can and cannot be detected in the data, what the possible annotation classes ought to be, and the rules by which an item ought to be classified into each class -- dramatically shape the annotation process, and through it, the AI. We draw on Bowker and Star's [9] classification theory to detail how the creation of a training data codebook for a computer vision algorithm in hospital intensive care units (ICUs) evolved from its original, clinically-driven goal of classifying complex clinical activities into a narrower goal of identifying physical objects and simpler activities in the ICU. This work reinforces how trade-offs and decisions made long before annotators begin labeling data are highly consequential to the resulting AI system.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- Between Subjectivity and Imposition: Power Dynamics in Data Annotation for Computer VisionMilagros Miceli, Martin Schuessler, Tianling YangCSCW 2020 · 被引用 148 次
- PaTAT: Human-AI Collaborative Qualitative Coding with Explainable Interactive Rule SynthesisSimret Araya Gebreegziabher, Zheng Zhang, Xiaohang Tang, Yihao Meng 等CHI 2023 · 被引用 70 次
- Whose AI Dream? In search of the aspiration in data annotationDing Wang, Shantanu Prabhat, Nithya SambasivanCHI 2022 · 被引用 66 次
- Tracing and Visualizing Human-ML/AI Collaborative Processes through Artifacts of Data WorkJen Rogers, Anamaria CrisanCHI 2023 · 被引用 10 次
- Products of Positionality: How Tech Workers Shape Identity Concepts in Computer VisionMorgan Klaus Scheuerman, Jed R. BrubakerCHI 2024 · 被引用 20 次
