CrowdGP: a Gaussian Process Model for Inferring Relevance from Crowd Annotations
Dan Li, Zhaochun Ren, Evangelos Kanoulas
Abstract
Test collection has been a crucial factor for developing information retrieval systems. Constructing a test collection requires annotators to assess the relevance of massive query-document pairs. Relevance annotations acquired through crowdsourcing platforms alleviate the enormous cost of this process but they are often noisy. Existing models to denoise crowd annotations mostly assume that annotations are generated independently, based on which a probabilistic graphical model is designed to model the annotation generation process. However, tasks are often correlated with each other in reality. It is an understudied problem whether and how task correlation helps in denoising crowd annotations. In this paper, we relax the independence assumption to model task correlation in terms of relevance. We propose a new crowd annotation generation model named CrowdGP, where true relevance labels, annotator competence, annotator’s bias towards relevancy, task difficulty, and task’s bias towards relevancy are modelled through a Gaussian process and multiple Gaussian variables respectively. The CrowdGP model shows better performance in terms of interring true relevance labels compared with state-of-the-art baselines on two crowdsourcing datasets on relevance. The experiments also demonstrate its effectiveness in terms of selecting new tasks for future crowd annotation, which is a new functionality of CrowdGP. Ablation studies indicate that the effectiveness is attributed to the modelling of task correlation based on the auxiliary information of tasks and the prior relevance information of documents to queries.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Extending Label Aggregation Models with a Gaussian Process to Denoise Crowdsourcing LabelsDan Li, Maarten de RijkeSIGIR 2023 · 21 citations
- Hate Personified: Investigating the role of LLMs in content moderationSarah Masud, Sahajpreet Singh, Viktor Hangya, Alexander Fraser et al.EMNLP 2024 · 6 citations
Related papers
- A Probabilistic Graphical Model for Analyzing the Subjective Visual Quality Assessment Data from CrowdsourcingJing Li, Suiyi Ling, Junle Wang, Patrick Le CalletACM MM 2020 · 23 citations
- Toward Annotator Group Bias in CrowdsourcingHaochen Liu, Joseph Thekinen, Sinem Mollaoglu, Da Tang et al.ACL 2022 · 20 citations
- Semi-Supervised Multi-Label Learning from Crowds via Deep Sequential Generative ModelWanli Shi, Victor S. Sheng, Xiang Li, Bin GuKDD 2020 · 12 citations
- Cross-Domain-Aware Worker Selection with Training for Crowdsourced AnnotationYushi Sun, Jiachuan Wang, Peng Cheng, Libin Zheng et al.ICDE 2024 · 3 citations
- Hierarchical Crowdsourcing for Data Labeling with Heterogeneous CrowdHaodi Zhang, Wenxi Huang, Zhenhan Su, Junyang Chen et al.ICDE 2023 · 4 citations
