BanHADEX: Towards Explainable HAte Speech Detection in Bangla Using Human Annotated EXplanation
Faisal Hossain Raquib, Akm Moshiur Rahman Mazumder, Md Fahim, Md. Tahmid Hasan Fuad, Md Farhan Ishmam, Faria Sultana, M. Ashraful Amin, Amin Ahsan Ali, Akmmahbubur Rahman
Abstract
Online safety in low-resource languages hinges not only on accurate hate speech detection but also on transparent, culturally grounded explanations. Yet prior work in Bangla largely focuses on hate classification while overlooking interpretability. We address this gap by introducing BANHADEX, the first hate explainability dataset in Bangla with human-annotated labels. BANHADEX contains 19,203 YouTube comments spanning April 2024-June 2025, annotated for binary hate classification with seven fine-grained hate categories, seven target groups, and concise explanations for each sample. Our data pipeline relies on a two-stage annotation protocol that uses majority voting for robust labeling. Our rich suite of experiments on open and closed-source LLMs reveals that explanation-guided LoRA substantially outperforms both classification and explanation quality across prompting and finetuning strategies. BANHADEX establishes the groundwork for faithful interpretability and safer moderation in linguistically rich yet underresourced languages. The code and dataset are publicly available at: https://github.com/ MOSHIIUR/BANHADEX .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fc61d9f4-451f-4b2b-8721-68c03a83989dBuilds on4
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- Efficient Memory Management for Large Language Model Serving with PagedAttentionWoosuk Kwon, Zhuohan Li, Siyuan Zhuang, Ying Sheng et al.SOSP 2023 · 1,016 citations
- Improving Automatic VQA Evaluation Using Large Language ModelsOscar Mañas, Benno Krojer, Aishwarya AgrawalAAAI 2024 · 59 citations
- BANMIME : Misogyny Detection with Metaphor Explanation on Bangla MemesMd Ayon Mia, Akm Moshiur Rahman Mazumder, Khadiza Sultana Sayma, Md Fahim et al.EMNLP 2025
Related papers
- LLM-Based Multi-Task Bangla Hate Speech Detection: Type, Severity, and TargetMd. Arid Hasan, Firoj Alam, Md Fahad Hossain, Usman Naseem et al.ACL 2026
- Deciphering Hate: Identifying Hateful Memes and Their TargetsEftekhar Hossain, Omar Sharif, Mohammed Moshiul Hoque, Sarah Masud PreumACL 2024 · 9 citations
- Spanning the Spectrum of Hatred Detection: A Persian Multi-Label Hate Speech Dataset with Annotator RationalesZahra Delbari, Nafise Sadat Moosavi, Mohammad Taher PilehvarAAAI 2024 · 11 citations
- SMARTER: A Data-efficient Framework to Improve Toxicity Detection with Explanation via Self-augmenting Large Language ModelsHuy Nghiem, Advik Sachdeva, Hal Daumé IIIACL 2026 · 1 citation
- MemeIntel: Explainable Detection of Propagandistic and Hateful MemesMohamed Bayan Kmainasi, Abul Hasnat, Md. Arid Hasan, Ali Ezzat Shahroor et al.EMNLP 2025 · 1 citation
