Hateful Word in Context Classification
Sanne Hoeken, Sina Zarrieß, Özge Alaçam
摘要
Hate speech detection is a prevalent research field, yet it remains underexplored at the level of word meaning. This is significant, as terms used to convey hate often involve non-standard or novel usages which might be overlooked by commonly leveraged LMs trained on general language use. In this paper, we introduce the Hateful Word in Context Classification (HateWiC) task and present a dataset of 4000 WiC-instances, each labeled by three annotators. Our analyses and computational exploration focus on the interplay between the subjective nature (context-dependent connotations) and the descriptive nature (as described in dictionary definitions) of hateful word senses. HateWiC annotations confirm that hatefulness of a word in context does not always derive from the sense definition alone. We explore the prediction of both majority and individual annotator labels, and we experiment with modeling context- and sense-based inputs. Our findings indicate that including definitions proves effective overall, yet not in cases where hateful connotations vary. Conversely, including annotator demographics becomes more important for mitigating performance drop in subjective hate prediction.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Eyes Don't Lie: Subjective Hate Annotation and Detection with GazeÖzge Alaçam, Sanne Hoeken, Sina ZarrießEMNLP 2024 · 被引用 1 次
- Disentangling Subjectivity and Uncertainty for Hate Speech Annotation and Modeling using GazeÖzge Alaçam, Sanne Hoeken, Andreas Säuberli, Hannes Gröner 等EMNLP 2025
它引用的顶会 Paper8
- Interpreting Pretrained Contextualized Representations via Reductions to Static EmbeddingsRishi Bommasani, Kelly Davis, Claire CardieACL 2020 · 被引用 137 次
- Analysing Lexical Semantic Change with Contextualised Word RepresentationsMario Giulianelli, Marco Del Tredici, Raquel FernándezACL 2020 · 被引用 118 次
- Probing Pretrained Language Models for Lexical SemanticsIvan Vulic, Edoardo Maria Ponti, Robert Litschko, Goran Glavas 等EMNLP 2020 · 被引用 26 次
- Moving Down the Long Tail of Word Sense Disambiguation with Gloss Informed Bi-encodersTerra Blevins, Luke ZettlemoyerACL 2020 · 被引用 19 次
- From Dogwhistles to Bullhorns: Unveiling Coded Rhetoric with Language ModelsJulia Mendelsohn, Ronan Le Bras, Yejin Choi, Maarten SapACL 2023 · 被引用 14 次
相关 Paper
- XL-WiC: A Multilingual Benchmark for Evaluating Semantic ContextualizationAlessandro Raganato, Tommaso Pasini, José Camacho-Collados, Mohammad Taher PilehvarEMNLP 2020 · 被引用 2 次
- AUTALIC: A Dataset for Anti-AUTistic Ableist Language In ContextNaba Rizvi, Harper Strickland, Daniel Gitelman, Alexis Morales Flores 等ACL 2025 · 被引用 5 次
- Hate Speech Detection Based on Sentiment Knowledge SharingXianbing Zhou, Yang Yong, Xiaochao Fan, Ge Ren 等ACL 2021
- When the Majority is Wrong: Modeling Annotator Disagreement for Subjective TasksEve Fleisig, Rediet Abebe, Dan KleinEMNLP 2023 · 被引用 11 次
- Comparative Evaluation of Label-Agnostic Selection Bias in Multilingual Hate Speech DatasetsNedjma Ousidhoum, Yangqiu Song, Dit-Yan YeungEMNLP 2020 · 被引用 22 次
