Fine-Grained Argument Unit Recognition and Classification
Dietrich Trautmann, Johannes Daxenberger, Christian Stab, Hinrich Schütze, Iryna Gurevych
Abstract
Prior work has commonly defined argument retrieval from heterogeneous document collections as a sentence-level classification task. Consequently, argument retrieval suffers both from low recall and from sentence segmentation errors making it difficult for humans and machines to consume the arguments. In this work, we argue that the task should be performed on a more fine-grained level of sequence labeling. For this, we define the task as Argument Unit Recognition and Classification (AURC). We present a dataset of arguments from heterogeneous sources annotated as spans of tokens within a sentence, as well as with a corresponding stance. We show that and how such difficult argument annotations can be effectively collected through crowdsourcing with high inter-annotator agreement. The new benchmark, AURC-8, contains up to 15% more arguments per topic as compared to annotations on the sentence level. We identify a number of methods targeted at AURC sequence labeling, achieving close to human performance on known domains. Further analysis also reveals that, contrary to previous approaches, our methods are more robust against sentence segmentation errors. We publicly release our code and the AURC-8 dataset.1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b10d0910-0f81-4cdd-abd3-501682f88835Cited by top-tier papers6
- Argument Mining Driven Analysis of Peer-ReviewsMichael Fromm, Evgeniy Faerman, Max Berrendorf, Siddharth Bhargava et al.AAAI 2021 · 36 citations
- Empowering the Fact-checkers! Automatic Identification of Claim Spans on TwitterMegha Sundriyal, Atharva Kulkarni, Vaibhav Pulastya, Md. Shad Akhtar et al.EMNLP 2022 · 12 citations
- HARGAN: Heterogeneous Argument Attention Network for Persuasiveness PredictionKuo Yu Huang, Hen-Hsen Huang, Hsin-Hsi ChenAAAI 2021 · 10 citations
- Superlim: A Swedish Language Understanding Evaluation BenchmarkAleksandrs Berdicevskis, Gerlof Bouma, Robin Kurtz, Felix Morger et al.EMNLP 2023 · 2 citations
- DREAM: Deployment of Recombination and Ensembles in Argument MiningFlorian Ruosch, Cristina Sarasua, Abraham BernsteinEMNLP 2023
Related papers
- Diversity Over Size: On the Effect of Sample and Topic Sizes for Topic-Dependent Argument Mining DatasetsBenjamin Schiller, Johannes Daxenberger, Andreas Waldis, Iryna GurevychEMNLP 2024 · 3 citations
- Corpus Wide Argument Mining - A Working SolutionLiat Ein-Dor, Eyal Shnarch, Lena Dankin, Alon Halfon et al.AAAI 2020 · 70 citations
- Few-Shot Document-Level Event Argument ExtractionXianjun Yang, Yujie Lu, Linda R. PetzoldACL 2023 · 5 citations
- A Large-Scale Dataset for Argument Quality Ranking: Construction and AnalysisShai Gretz, Roni Friedman, Edo Cohen-Karlik, Assaf Toledo et al.AAAI 2020 · 148 citations
- Limited Generalizability in Argument Mining: State-Of-The-Art Models Learn Datasets, Not ArgumentsMarc Feger, Katarina Boland, Stefan DietzeACL 2025
