Modeling and Aggregation of Complex Annotations via Annotation Distances
Alexander Braylan, Matthew Lease
摘要
Modeling annotators and their labels is valuable for ensuring collected data quality. Though many models have been proposed for binary or categorical labels, prior methods do not generalize to complex annotations (e.g., open-ended text, multivariate, or structured responses) without devising new models for each specific task. To obviate the need for task-specific modeling, we propose to model distances between labels, rather than the labels themselves. Our models are largely agnostic to the distance function; we leave it to the requesters to specify an appropriate distance function for their given annotation task. We propose three models of annotation quality, including a Bayesian hierarchical extension of multidimensional scaling which can be trained in an unsupervised or semi-supervised manner. Results show the generality and effectiveness of our models across diverse complex annotation tasks: sequence labeling, translation, syntactic parsing, and ranking.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Measuring Annotator Agreement Generally across Complex Structured, Multi-object, and Free-text Annotation TasksAlexander Braylan, Omar Alonso, Matthew LeaseWWW 2022 · 被引用 34 次
- Reliance and Automation for Human-AI Collaborative Data Labeling Conflict ResolutionMichelle Brachman, Zahra Ashktorab, Michael Desmond, Evelyn Duesterwald 等CSCW 2022 · 被引用 16 次
- Hate Personified: Investigating the role of LLMs in content moderationSarah Masud, Sahajpreet Singh, Viktor Hangya, Alexander Fraser 等EMNLP 2024 · 被引用 6 次
- HybridEval: A Human-AI Collaborative Approach for Evaluating Design Ideas at ScaleSepideh Mesbah, Ines Arous, Jie Yang, Alessandro BozzonWWW 2023 · 被引用 5 次
- Frustratingly Easy Truth DiscoveryReshef Meir, Ofra Amir, Omer Ben-Porat, Tsviel Ben Shabat 等AAAI 2023 · 被引用 2 次
相关 Paper
- Aggregating Complex Annotations via Merging and MatchingAlexander Braylan, Matthew LeaseKDD 2021 · 被引用 8 次
- Noise Correction on Subjective DatasetsUthman Jinadu, Yi DingACL 2024 · 被引用 2 次
- Architectural Sweet Spots for Modeling Human Label Variation by the Example of Argument Quality: It's Best to Relate Perspectives!Philipp Heinisch, Matthias Orlikowski, Julia Romberg, Philipp CimianoEMNLP 2023 · 被引用 1 次
- Predicting User Preferences of Dimensionality Reduction Embedding QualityCristina Morariu, Adrien Bibal, René Cutura, Benoît Frénay 等IEEE VIS 2022 · 被引用 13 次
- CalCo: A Hierarchical Bayesian Framework for Scalable Human-LLM Hybrid LabelingViet-An Nguyen, Xu Chen, Udi WeinsbergKDD 2026
