Unsupervised Discovery of Implicit Gender Bias
Anjalie Field, Yulia Tsvetkov
Abstract
Despite their prevalence in society, social biases are difficult to identify, primarily because human judgements in this domain can be unreliable. We take an unsupervised approach to identifying gender bias against women at a comment level and present a model that can surface text likely to contain bias. Our main challenge is forcing the model to focus on signs of implicit bias, rather than other artifacts in the data. Thus, our methodology involves reducing the influence of confounds through propensity matching and adversarial learning. Our analysis shows how biased comments directed towards female politicians contain mixed criticisms, while comments directed towards other female public figures focus on appearance and sexualization. Ultimately, our work offers a way to capture subtle biases in various domains without relying on subjective human judgements. 1 I love tennis! Tennis is great! Do I look ok? Bro <title>, golf is better UR hot! Me too <3 UR hot! Canada's got no game
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b36c803c-6bf2-45de-8907-54a14253f6b3Cited by top-tier papers10
- Controlled Text Generation as Continuous Optimization with Multiple ConstraintsSachin Kumar, Eric Malmi, Aliaksei Severyn, Yulia TsvetkovNeurIPS 2021 · 91 citations
- Uncovering Latent Biases in Text: Method and Application to Peer ReviewEmaad A. Manzoor, Nihar B. ShahAAAI 2021 · 42 citations
- Improving Generalizability in Implicitly Abusive Language Detection with Concept Activation VectorsIsar Nejadgholi, Kathleen C. Fraser, Svetlana KiritchenkoACL 2022 · 27 citations
- Gradient-based Constrained Sampling from Language ModelsSachin Kumar, Biswajit Paria, Yulia TsvetkovEMNLP 2022 · 22 citations
- Transcending the "Male Code": Implicit Masculine Biases in NLP ContextsKatie Seaborn, Shruti Chandra, Thibault FabreCHI 2023 · 16 citations
Builds on2
- Social Bias Frames: Reasoning about Social and Power Implications of LanguageMaarten Sap, Saadia Gabriel, Lianhui Qin, Dan Jurafsky et al.ACL 2020 · 16 citations
- Text and Causal Inference: A Review of Using Text to Remove Confounding from Causal EstimatesKatherine A. Keith, David D. Jensen, Brendan O'ConnorACL 2020 · 16 citations
Related papers
- "Fifty Shades of Bias": Normative Ratings of Gender Bias in GPT Generated English TextRishav Hada, Agrima Seth, Harshita Diddee, Kalika BaliEMNLP 2023 · 10 citations
- Intrinsic Bias Metrics Do Not Correlate with Application BiasSeraphina Goldfarb-Tarrant, Rebecca Marchant, Ricardo Muñoz Sánchez, Mugdha Pandya et al.ACL 2021
- Multi-Dimensional Gender Bias ClassificationEmily Dinan, Angela Fan, Ledell Wu, Jason Weston et al.EMNLP 2020 · 7 citations
- Discovering and Mitigating Visual Biases Through Keyword ExplanationYounghyun Kim, Sangwoo Mo, Minkyu Kim, Kyungmin Lee et al.CVPR 2024
- Classifier-to-Bias: Toward Unsupervised Automatic Bias Detection for Visual ClassifiersQuentin Guimard, Moreno D'Incà, Massimiliano Mancini, Elisa RicciCVPR 2025
