KaLOS finds Consensus: A Meta-Algorithm for Evaluating Inter-Annotator Agreement in Complex Vision Tasks
David Tschirschwitz, Volker Rodehorst
2026年份
1被引次数
摘要
Progress in object detection benchmarks is stagnating. It is limited not by architectures but by the inability to distinguish model improvements from label noise. To restore trust in benchmarking the field requires rigorous quantification of annotation consistency to ensure the reliability of evaluation data. However, standard statistical metrics fail to handle the instance correspondence problem inherent to vision tasks. Furthermore, validating new agreement metrics remains circular because no objective ground truth for agreement exists. This forces reliance on unverifiable heuristics.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper7
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- DETRs with Collaborative Hybrid Assignments TrainingZhuofan Zong, Guanglu Song, Yu LiuICCV 2023 · 被引用 594 次
- Scaling Open-Vocabulary Object DetectionMatthias Minderer, Alexey A. Gritsenko, Neil HoulsbyNeurIPS 2023 · 被引用 482 次
- Towards Robust Adaptive Object Detection under Noisy AnnotationsXinyu Liu, Wuyang Li, Qiushi Yang, Baopu Li 等CVPR 2022 · 被引用 40 次
- Measuring Annotator Agreement Generally across Complex Structured, Multi-object, and Free-text Annotation TasksAlexander Braylan, Omar Alonso, Matthew LeaseWWW 2022 · 被引用 34 次
相关 Paper
- A Theory of Dynamic BenchmarksAli Shirali, Rediet Abebe, Moritz HardtICLR 2023 · 被引用 1 次
- Discrepancy Ratio: Evaluating Model Performance When Even Experts Disagree on the TruthIgor Lovchinsky, Alon Daks, Israel Malkin, Pouya Samangouei 等ICLR 2020 · 被引用 11 次
- Pixel-level Quality Assessment for Oriented Object DetectionYunhui Zhu, Buliao HuangAAAI 2026
- Formally Exploring Visual Anomaly Detection Evaluation MetricsNasar Iqbal, Dennis Wagner, Philipp Liznerski, Nabeel Hussain Syed 等ICML 2026
- Are We Overconfident in Models and Results for Semi-Supervised 3D Medical Image Segmentation?Jun Li, ZIWEI QINICML 2026
