Assessing the Reliability of Word Embedding Gender Bias Measures
Yupei Du, Qixiang Fang, Dong Nguyen
Abstract
Various measures have been proposed to quantify human-like social biases in word embeddings. However, bias scores based on these measures can suffer from measurement error. One indication of measurement quality is reliability, concerning the extent to which a measure produces consistent results. In this paper, we assess three types of reliability of word embedding gender bias measures, namely testretest reliability, inter-rater consistency and internal consistency. Specifically, we investigate the consistency of bias scores across different choices of random seeds, scoring rules and words. Furthermore, we analyse the effects of various factors on these measures' reliability scores. Our findings inform better design of word embedding gender bias measures. Moreover, we urge researchers to be more critical about the application of such measures. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ccbeba3d-4aac-4440-9813-8c33912fa635Cited by top-tier papers1
Ask how each one uses itBuilds on3
- Exploring the Linear Subspace Hypothesis in Gender Bias MitigationFrancisco Vargas, Ryan CotterellEMNLP 2020 · 2 citations
- Bad Seeds: Evaluating Lexical Methods for Bias MeasurementMaria Antoniak, David MimnoACL 2021
- Intrinsic Bias Metrics Do Not Correlate with Application BiasSeraphina Goldfarb-Tarrant, Rebecca Marchant, Ricardo Muñoz Sánchez, Mugdha Pandya et al.ACL 2021
Related papers
- A Causal Inference Method for Reducing Gender Bias in Word Embedding RelationsZekun Yang, Juan FengAAAI 2020 · 40 citations
- When do Word Embeddings Accurately Reflect Surveys on our Beliefs About People?Kenneth Joseph, Jonathan H. MorganACL 2020 · 5 citations
- Understanding Gender Bias in Knowledge Base EmbeddingsYupei Du, Qi Zheng, Yuanbin Wu, Man Lan et al.ACL 2022
- A General Framework for Implicit and Explicit Debiasing of Distributional Word Vector SpacesAnne Lauscher, Goran Glavas, Simone Paolo Ponzetto, Ivan VulicAAAI 2020 · 68 citations
- ValNorm Quantifies Semantics to Reveal Consistent Valence Biases Across Languages and Over CenturiesAutumn Toney, Aylin CaliskanEMNLP 2021
