CrowdGP: a Gaussian Process Model for Inferring Relevance from Crowd Annotations
Dan Li, Zhaochun Ren, Evangelos Kanoulas
摘要
Test collection has been a crucial factor for developing information retrieval systems. Constructing a test collection requires annotators to assess the relevance of massive query-document pairs. Relevance annotations acquired through crowdsourcing platforms alleviate the enormous cost of this process but they are often noisy. Existing models to denoise crowd annotations mostly assume that annotations are generated independently, based on which a probabilistic graphical model is designed to model the annotation generation process. However, tasks are often correlated with each other in reality. It is an understudied problem whether and how task correlation helps in denoising crowd annotations. In this paper, we relax the independence assumption to model task correlation in terms of relevance. We propose a new crowd annotation generation model named CrowdGP, where true relevance labels, annotator competence, annotator’s bias towards relevancy, task difficulty, and task’s bias towards relevancy are modelled through a Gaussian process and multiple Gaussian variables respectively. The CrowdGP model shows better performance in terms of interring true relevance labels compared with state-of-the-art baselines on two crowdsourcing datasets on relevance. The experiments also demonstrate its effectiveness in terms of selecting new tasks for future crowd annotation, which is a new functionality of CrowdGP. Ablation studies indicate that the effectiveness is attributed to the modelling of task correlation based on the auxiliary information of tasks and the prior relevance information of documents to queries.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Extending Label Aggregation Models with a Gaussian Process to Denoise Crowdsourcing LabelsDan Li, Maarten de RijkeSIGIR 2023 · 被引用 21 次
- Hate Personified: Investigating the role of LLMs in content moderationSarah Masud, Sahajpreet Singh, Viktor Hangya, Alexander Fraser 等EMNLP 2024 · 被引用 6 次
相关 Paper
- A Probabilistic Graphical Model for Analyzing the Subjective Visual Quality Assessment Data from CrowdsourcingJing Li, Suiyi Ling, Junle Wang, Patrick Le CalletACM MM 2020 · 被引用 23 次
- Toward Annotator Group Bias in CrowdsourcingHaochen Liu, Joseph Thekinen, Sinem Mollaoglu, Da Tang 等ACL 2022 · 被引用 20 次
- Semi-Supervised Multi-Label Learning from Crowds via Deep Sequential Generative ModelWanli Shi, Victor S. Sheng, Xiang Li, Bin GuKDD 2020 · 被引用 12 次
- Cross-Domain-Aware Worker Selection with Training for Crowdsourced AnnotationYushi Sun, Jiachuan Wang, Peng Cheng, Libin Zheng 等ICDE 2024 · 被引用 3 次
- Hierarchical Crowdsourcing for Data Labeling with Heterogeneous CrowdHaodi Zhang, Wenxi Huang, Zhenhan Su, Junyang Chen 等ICDE 2023 · 被引用 4 次
