Finding Authentic Counterhate Arguments: A Case Study with Public Figures
Abdullah Albanyan, Ahmed Hassan, Eduardo Blanco
摘要
We explore authentic counterhate arguments for online hateful content toward individuals. Previous efforts are limited to counterhate to fight against hateful content toward groups. Thus, we present a corpus of 54,816 hateful tweet-paragraph pairs, where the paragraphs are candidate counterhate arguments. The counterhate arguments are retrieved from 2,500 online articles from multiple sources. We propose a methodology that assures the authenticity of the counter argument and its specificity to the individual of interest. We show that finding arguments in online articles is an efficient alternative to counterhate generation approaches that may hallucinate unsupported arguments. We also present linguistic insights on the language used in counterhate arguments. Experimental results show promising results. It is more challenging, however, to identify counterhate arguments for hateful content toward individuals not included in the training set. Authentic counterhate argument from Dailypost.ng: www.dailypost.ng/2012/05/11/ [...] -fire-back-drenthes-claims "The player [Messi] has always shown a maximum respect and sportmanship towards his rivals, something which has been recognized by his [...]" Authentic counterhate argument from Quora.com: www.quora.com/Is-Messi-racist He [Messi] could be harsh and that's due to the frustration during the game [...], it's all love from Messi.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper1
相关 Paper
- A Fine-Grained Taxonomy of Replies to Hate SpeechXinchen Yu, Ashley Zhao, Eduardo Blanco, Lingzi HongEMNLP 2023 · 被引用 1 次
- Pinpointing Fine-Grained Relationships between Hateful Tweets and RepliesAbdullah Albanyan, Eduardo BlancoAAAI 2022 · 被引用 9 次
- Is Safer Better? The Impact of Guardrails on the Argumentative Strength of LLMs in Hate Speech CounteringHelena Bonaldi, Greta Damo, Nicolás Benjamín Ocampo, Elena Cabrio 等EMNLP 2024 · 被引用 2 次
- Counterspeakers' Perspectives: Unveiling Barriers and AI Needs in the Fight against Online HateJimin Mun, Cathy Buerger, Jenny T. Liang, Joshua Garland 等CHI 2024 · 被引用 12 次
- PersonaHate: A Scalable Persona-Based Data Synthesis Pipeline for Hate Speech AnalysisXinyu Zhang, Ziqing Yang, Michael Backes, Yang ZhangCCS 2026
