Modeling and Aggregation of Complex Annotations via Annotation Distances
Alexander Braylan, Matthew Lease
Abstract
Modeling annotators and their labels is valuable for ensuring collected data quality. Though many models have been proposed for binary or categorical labels, prior methods do not generalize to complex annotations (e.g., open-ended text, multivariate, or structured responses) without devising new models for each specific task. To obviate the need for task-specific modeling, we propose to model distances between labels, rather than the labels themselves. Our models are largely agnostic to the distance function; we leave it to the requesters to specify an appropriate distance function for their given annotation task. We propose three models of annotation quality, including a Bayesian hierarchical extension of multidimensional scaling which can be trained in an unsupervised or semi-supervised manner. Results show the generality and effectiveness of our models across diverse complex annotation tasks: sequence labeling, translation, syntactic parsing, and ranking.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 27f74473-fac2-48f0-854d-e805321fa191Cited by top-tier papers6
- Measuring Annotator Agreement Generally across Complex Structured, Multi-object, and Free-text Annotation TasksAlexander Braylan, Omar Alonso, Matthew LeaseWWW 2022 · 34 citations
- Reliance and Automation for Human-AI Collaborative Data Labeling Conflict ResolutionMichelle Brachman, Zahra Ashktorab, Michael Desmond, Evelyn Duesterwald et al.CSCW 2022 · 16 citations
- Hate Personified: Investigating the role of LLMs in content moderationSarah Masud, Sahajpreet Singh, Viktor Hangya, Alexander Fraser et al.EMNLP 2024 · 6 citations
- HybridEval: A Human-AI Collaborative Approach for Evaluating Design Ideas at ScaleSepideh Mesbah, Ines Arous, Jie Yang, Alessandro BozzonWWW 2023 · 5 citations
- Frustratingly Easy Truth DiscoveryReshef Meir, Ofra Amir, Omer Ben-Porat, Tsviel Ben Shabat et al.AAAI 2023 · 2 citations
Related papers
- Aggregating Complex Annotations via Merging and MatchingAlexander Braylan, Matthew LeaseKDD 2021 · 8 citations
- Noise Correction on Subjective DatasetsUthman Jinadu, Yi DingACL 2024 · 2 citations
- Architectural Sweet Spots for Modeling Human Label Variation by the Example of Argument Quality: It's Best to Relate Perspectives!Philipp Heinisch, Matthias Orlikowski, Julia Romberg, Philipp CimianoEMNLP 2023 · 1 citation
- Predicting User Preferences of Dimensionality Reduction Embedding QualityCristina Morariu, Adrien Bibal, René Cutura, Benoît Frénay et al.IEEE VIS 2022 · 13 citations
- CalCo: A Hierarchical Bayesian Framework for Scalable Human-LLM Hybrid LabelingViet-An Nguyen, Xu Chen, Udi WeinsbergKDD 2026
