Human-LLM Hybrid Text Answer Aggregation for Crowd Annotations
Jiyi Li
摘要
The quality is a crucial issue for crowd annotations. Answer aggregation is an important type of solution. The aggregated answers estimated from multiple crowd answers to the same instance are the eventually collected annotations, rather than the individual crowd answers themselves. Recently, the capability of Large Language Models (LLMs) on data annotation tasks has attracted interest from researchers. Most of the existing studies mainly focus on the average performance of individual crowd workers; several recent works studied the scenarios of aggregation on categorical labels and LLMs used as label creators. However, the scenario of aggregation on text answers and the role of LLMs as aggregators are not yet well-studied. In this paper, we investigate the capability of LLMs as aggregators in the scenario of close-ended crowd text answer aggregation. We propose a human-LLM hybrid text answer aggregation method with a Creator-Aggregator Multi-Stage (CAMS) crowdsourcing framework. We make the experiments based on public crowdsourcing datasets. The results show the effectiveness of our approach based on the collaboration of crowd workers and LLMs.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Prompt Candidates, then Distill: A Teacher-Student Framework for LLM-driven Data AnnotationMingxuan Xia, Haobo Wang, Yixuan Li, Zewei Yu 等ACL 2025 · 被引用 4 次
- Label Aggregation for Composite Crowd Tasks by Worker Ability Constraint SatisfactionJiyi LiAAAI 2025 · 被引用 1 次
它引用的顶会 Paper7
- AI Chains: Transparent and Controllable Human-AI Interaction by Chaining Large Language Model PromptsTongshuang Wu, Michael Terry, Carrie Jun CaiCHI 2022 · 被引用 465 次
- If in a Crowdsourced Data Annotation Pipeline, a GPT-4Zeyu He, Chieh-Yang Huang, Chien-Kuang Cornelia Ding, Shaurya Rohatgi 等CHI 2024 · 被引用 31 次
- ChatGPT to Replace Crowdsourcing of Paraphrases for Intent Classification: Higher Diversity and Comparable Model RobustnessJán Cegin, Jakub Simko, Peter BrusilovskyEMNLP 2023 · 被引用 25 次
- Rank Aggregation via Heterogeneous Thurstone Preference ModelsTao Jin, Pan Xu, Quanquan Gu, Farzad FarnoudAAAI 2020 · 被引用 19 次
- Context-based Collective Preference Aggregation for Prioritizing Crowd Opinions in Social Decision-makingJiyi LiWWW 2022 · 被引用 3 次
相关 Paper
- Human-LLM Collaborative Annotation Through Effective Verification of LLM LabelsXinru Wang, Hannah Kim, Sajjadur Rahman, Kushan Mitra 等CHI 2024 · 被引用 127 次
- Next Generation Active Learning: Mixture of LLMs in the LoopYuanyuan Qi, Xiaohao Yang, Jueqing Lu, Guoxiang Guo 等AAAI 2026
- SkillAggregation: Reference-free LLM-Dependent AggregationGuangzhi Sun, Anmol Kagrecha, Potsawee Manakul, Philip C. Woodland 等ACL 2025 · 被引用 5 次
- Evaluating LLM-contaminated Crowdsourcing Data Without Ground TruthYichi Zhang, Jinlong Pang, Zhaowei Zhu, Yang LiuNeurIPS 2025 · 被引用 3 次
- Wisdom of the Crowd, Without the Crowd: A Socratic LLM for Asynchronous Deliberation on Perspectivist DataMalik Khadar, Daniel Runningen, Julia Tang, Stevie Chancellor 等CSCW 2025 · 被引用 2 次
