Efficient Pairwise Annotation of Argument Quality
Lukas Gienapp, Benno Stein, Matthias Hagen, Martin Potthast
Abstract
We present an efficient annotation framework for argument quality, a feature difficult to be measured reliably as per previous work. A stochastic transitivity model is combined with an effective sampling strategy to infer highquality labels with low effort from crowdsourced pairwise judgments. The model's capabilities are showcased by compiling Webis-ArgQuality-20, an argument quality corpus that comprises scores for rhetorical, logical, dialectical, and overall quality inferred from a total of 41,859 pairwise judgments among 1,271 arguments. With up to 93% cost savings, our approach significantly outperforms existing annotation procedures. Furthermore, novel insight into argument quality is provided through statistical analysis, and a new aggregation method to infer overall quality from individual quality dimensions is proposed.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 991dbf21-e5e3-4d36-8a9e-7aa580b27580Cited by top-tier papers6
- The Viability of Crowdsourcing for RAG EvaluationLukas Gienapp, Tim Hagen, Maik Fröbe, Matthias Hagen et al.SIGIR 2025 · 7 citations
- Data-Driven Insight Synthesis for Multi-Dimensional DataJunjie Xing, Xinyu Wang, H. V. JagadishVLDB 2024 · 5 citations
- Let's discuss! Quality Dimensions and Annotated Datasets for Computational Argument Quality AssessmentRositsa V. Ivanova, Thomas Huber, Christina NiklausEMNLP 2024 · 2 citations
- A Multi-persona Framework for Argument Quality AssessmentBojun Jin, Jianzhu Bao, Yufang Hou, Yang Sun et al.ACL 2025
- LLM-based Rewriting of Inappropriate Argumentation using Reinforcement Learning from Machine FeedbackTimon Ziegenbein, Gabriella Skitalinskaya, Alireza Bayat Makou, Henning WachsmuthACL 2024
Related papers
- A Large-Scale Dataset for Argument Quality Ranking: Construction and AnalysisShai Gretz, Roni Friedman, Edo Cohen-Karlik, Assaf Toledo et al.AAAI 2020 · 148 citations
- A Probabilistic Graphical Model for Analyzing the Subjective Visual Quality Assessment Data from CrowdsourcingJing Li, Suiyi Ling, Junle Wang, Patrick Le CalletACM MM 2020 · 23 citations
- Fine-Grained Argument Unit Recognition and ClassificationDietrich Trautmann, Johannes Daxenberger, Christian Stab, Hinrich Schütze et al.AAAI 2020 · 70 citations
- Modeling and Aggregation of Complex Annotations via Annotation DistancesAlexander Braylan, Matthew LeaseWWW 2020 · 15 citations
- Human Rationales as Attribution Priors for Explainable Stance DetectionSahil Jayaram, Emily AllawayEMNLP 2021 · 17 citations
