Let the Prototype Guide You: Robust Aggregation of Sparse Multi-Class Annotations via Annotator Prototype Learning
Ju Chen, Jun Feng, Shenyu Zhang
摘要
Truth inference is a critical technique for aggregating noisy and biased multi-class classification annotations. State-of-the-art approaches model each annotator using an individual confusion matrix. While well-grounded, they suffer from two fundamental bottlenecks: 1) confusion matrices are underfit when annotators label only a small subset of tasks or when classes are imbalanced, and 2) a single confusion matrix per annotator is inadequate for capturing complex annotator behaviors, leading to class-level collapse when tasks are extremely difficult. Simultaneously addressing these challenges is non-trivial, as it demands both robustness to data sparsity and sufficient expressiveness for complex annotator patterns. In this paper, we propose CPBCC (Class-specific Prototype-driven Bayesian Classifier Combination), which creatively models annotators through a dual-pathway architecture: (i) learning class-specific prototype annotation patterns across all annotators, and (ii) learning annotator-specific weights over prototypes. This framework addresses the bottlenecks and achieves a robust yet rich annotator characterization. Experiments across 10 real-world datasets spanning five domains demonstrate that CPBCC yields a 26% accuracy improvement in the best case, and boosts average accuracy from 68.73% to 74.11%. Our source code is available at https://github.com/JuJuCHEN-HHU/CPBCC_PTBCC.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper11
- Learning from Crowds by Modeling Common ConfusionsZhendong Chu, Jing Ma, Hongning WangAAAI 2021 · 被引用 60 次
- Open Knowledge Enrichment for Long-tail EntitiesErmei Cao, Difeng Wang, Jiacheng Huang, Wei HuWWW 2020 · 被引用 51 次
- Adversarial Learning from CrowdsPengpeng Chen, Hailong Sun, Yongqiang Yang, Zhijun ChenAAAI 2022 · 被引用 16 次
- Crowdsourcing via Annotator Co-occurrence Imputation and Provable Symmetric Nonnegative Matrix FactorizationShahana Ibrahim, Xiao FuICML 2021 · 被引用 12 次
- DHG-Bench: A Comprehensive Benchmark for Deep Hypergraph LearningFan Li, Xiaoyang Wang, Wenjie Zhang, Ying Zhang 等ICLR 2026 · 被引用 9 次
相关 Paper
- Coupled Confusion Correction: Learning from Crowds with Sparse AnnotationsHansong Zhang, Shikun Li, Dan Zeng, Chenggang Yan 等AAAI 2024 · 被引用 23 次
- Coupled-View Deep Classifier Learning from Multiple Noisy AnnotatorsShikun Li, Shiming Ge, Yingying Hua, Chunhui Zhang 等AAAI 2020 · 被引用 30 次
- Aggregating Complex Annotations via Merging and MatchingAlexander Braylan, Matthew LeaseKDD 2021 · 被引用 8 次
- QuMAB: Query-based Multi-annotator Behavior Pattern LearningLiyun Zhang, Zheng Lian, Hong Liu, Takanori Takebe 等AAAI 2026 · 被引用 3 次
- Noisy Label Learning with Instance-Dependent Outliers: Identifiability via Crowd WisdomTri Nguyen, Shahana Ibrahim, Xiao FuNeurIPS 2024 · 被引用 14 次
