Estimating Agreement by Chance for Sequence Annotation
Diya Li, Carolyn P. Rosé, Ao Yuan, Chunxiao Zhou
摘要
In the field of natural language processing, correction of performance assessment for chance agreement plays a crucial role in evaluating the reliability of annotations. However, there is a notable dearth of research focusing on chance correction for assessing the reliability of sequence annotation tasks, despite their widespread prevalence in the field. To address this gap, this paper introduces a novel model for generating random annotations, which serves as the foundation for estimating chance agreement in sequence annotation tasks. Utilizing the proposed randomization model and a related comparison approach, we successfully derive the analytical form of the distribution, enabling the computation of the probable location of each annotated text segment and subsequent chance agreement estimation. Through a combination simulation and corpus-based evaluation, we successfully assess its applicability and validate its accuracy and efficacy.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
相关 Paper
- Evaluating Sequence Labeling on the basis of Information TheoryEnrique Amigó, Elena Álvarez Mellado, Julio Gonzalo, Jorge Carrillo-de-AlbornozACL 2025
- Overcoming Copyright Barriers in Corpus Distribution Through Non-Reversible HashingArthur Amalvy, Vincent Labatut, Xavier Bost, Hen-Hsen HuangACL 2026
- Near-Negative Distinction: Giving a Second Life to Human Evaluation DatasetsPhilippe Laban, Chien-Sheng Wu, Wenhao Liu, Caiming XiongEMNLP 2022 · 被引用 4 次
- We Need to Talk About Reproducibility in NLP Model ComparisonYan Xue, Xuefei Cao, Xingli Yang, Yu Wang 等EMNLP 2023 · 被引用 2 次
- Evaluating Evaluation Metrics: A Framework for Analyzing NLG Evaluation Metrics using Measurement TheoryZiang Xiao, Susu Zhang, Vivian Lai, Q. Vera LiaoEMNLP 2023 · 被引用 6 次
