Corpus Wide Argument Mining - A Working Solution
Liat Ein-Dor, Eyal Shnarch, Lena Dankin, Alon Halfon, Benjamin Sznajder, Ariel Gera, Carlos Alzate, Martin Gleize, Leshem Choshen, Yufang Hou, Yonatan Bilu, Ranit Aharonov, Noam Slonim
Abstract
One of the main tasks in argument mining is the retrieval of argumentative content pertaining to a given topic. Most previous work addressed this task by retrieving a relatively small number of relevant documents as the initial source for such content. This line of research yielded moderate success, which is of limited use in a real-world system. Furthermore, for such a system to yield a comprehensive set of relevant arguments, over a wide range of topics, it requires leveraging a large and diverse corpus in an appropriate manner. Here we present a first end-to-end high-precision, corpus-wide argument mining system. This is made possible by combining sentence-level queries over an appropriate indexing of a very large corpus of newspaper articles, with an iterative annotation scheme. This scheme addresses the inherent label bias in the data and pinpoints the regions of the sample space whose manual labeling is required to obtain high-precision among top-ranked candidates.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 94497b05-df34-47ed-a3dd-af41497d22e0Cited by top-tier papers8
- Cluster & Tune: Boost Cold Start Performance in Text ClassificationEyal Shnarch, Ariel Gera, Alon Halfon, Lena Dankin et al.ACL 2022 · 27 citations
- Diversity Over Size: On the Effect of Sample and Topic Sizes for Topic-Dependent Argument Mining DatasetsBenjamin Schiller, Johannes Daxenberger, Andreas Waldis, Iryna GurevychEMNLP 2024 · 3 citations
- Label-Efficient Model Selection for Text GenerationShir Ashury-Tahan, Ariel Gera, Benjamin Sznajder, Leshem Choshen et al.ACL 2024 · 1 citation
- Debatable Intelligence: Benchmarking LLM Judges via Debate Speech EvaluationNoy Sternlicht, Ariel Gera, Roy Bar-Haim, Tom Hope et al.EMNLP 2025 · 1 citation
- The Moral Debater: A Study on the Computational Generation of Morally Framed ArgumentsMilad Alshomary, Roxanne El Baff, Timon Gurcke, Henning WachsmuthACL 2022
Related papers
- Fine-Grained Argument Unit Recognition and ClassificationDietrich Trautmann, Johannes Daxenberger, Christian Stab, Hinrich Schütze et al.AAAI 2020 · 70 citations
- ArgAnalysis35K : A large-scale dataset for Argument Quality AnalysisOmkar Joshi, Priya Pitre, Yashodhara HaribhaktaACL 2023 · 4 citations
- A Generative Model for End-to-End Argument Mining with Reconstructed Positional Encoding and Constrained Pointer MechanismJianzhu Bao, Yuhang He, Yang Sun, Bin Liang et al.EMNLP 2022 · 15 citations
- A Large-Scale Dataset for Argument Quality Ranking: Construction and AnalysisShai Gretz, Roni Friedman, Edo Cohen-Karlik, Assaf Toledo et al.AAAI 2020 · 148 citations
- A Neural Transition-based Model for Argumentation MiningJianzhu Bao, Chuang Fan, Jipeng Wu, Yixue Dang et al.ACL 2021
