Finding Authentic Counterhate Arguments: A Case Study with Public Figures
Abdullah Albanyan, Ahmed Hassan, Eduardo Blanco
Abstract
We explore authentic counterhate arguments for online hateful content toward individuals. Previous efforts are limited to counterhate to fight against hateful content toward groups. Thus, we present a corpus of 54,816 hateful tweet-paragraph pairs, where the paragraphs are candidate counterhate arguments. The counterhate arguments are retrieved from 2,500 online articles from multiple sources. We propose a methodology that assures the authenticity of the counter argument and its specificity to the individual of interest. We show that finding arguments in online articles is an efficient alternative to counterhate generation approaches that may hallucinate unsupported arguments. We also present linguistic insights on the language used in counterhate arguments. Experimental results show promising results. It is more challenging, however, to identify counterhate arguments for hateful content toward individuals not included in the training set. Authentic counterhate argument from Dailypost.ng: www.dailypost.ng/2012/05/11/ [...] -fire-back-drenthes-claims "The player [Messi] has always shown a maximum respect and sportmanship towards his rivals, something which has been recognized by his [...]" Authentic counterhate argument from Quora.com: www.quora.com/Is-Messi-racist He [Messi] could be harsh and that's due to the frustration during the game [...], it's all love from Messi.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 64ef3850-42f9-40fe-964a-4c547ab6f92fCited by top-tier papers1
Ask how each one uses itBuilds on1
Related papers
- A Fine-Grained Taxonomy of Replies to Hate SpeechXinchen Yu, Ashley Zhao, Eduardo Blanco, Lingzi HongEMNLP 2023 · 1 citation
- Pinpointing Fine-Grained Relationships between Hateful Tweets and RepliesAbdullah Albanyan, Eduardo BlancoAAAI 2022 · 9 citations
- Is Safer Better? The Impact of Guardrails on the Argumentative Strength of LLMs in Hate Speech CounteringHelena Bonaldi, Greta Damo, Nicolás Benjamín Ocampo, Elena Cabrio et al.EMNLP 2024 · 2 citations
- Counterspeakers' Perspectives: Unveiling Barriers and AI Needs in the Fight against Online HateJimin Mun, Cathy Buerger, Jenny T. Liang, Joshua Garland et al.CHI 2024 · 12 citations
- PersonaHate: A Scalable Persona-Based Data Synthesis Pipeline for Hate Speech AnalysisXinyu Zhang, Ziqing Yang, Michael Backes, Yang ZhangCCS 2026
