Low Frequency Names Exhibit Bias and Overfitting in Contextualizing Language Models
Robert Wolfe, Aylin Caliskan
摘要
We use a dataset of U.S. first names with labels based on predominant gender and racial group to examine the effect of training corpus frequency on tokenization, contextualization, similarity to initial representation, and bias in BERT, GPT-2, T5, and XLNet. We show that predominantly female and non-white names are less frequent in the training corpora of these four language models. We find that infrequent names are more self-similar across contexts, with Spearman's ρ between frequency and self-similarity as low as -.763. Infrequent names are also less similar to initial representation, with Spearman's ρ between frequency and linear centered kernel alignment (CKA) similarity to initial representation as high as .702. Moreover, we find Spearman's ρ between racial bias and name frequency in BERT of .492, indicating that lower-frequency minority group names are more associated with unpleasantness. Representations of infrequent names undergo more processing, but are more self-similar, indicating that models rely on less context-informed representations of uncommon and minority names which are overfit to a lower number of observed contexts.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- See Widely, Think Wisely: Toward Designing a Generative Multi-agent System to Burst Filter BubblesYu Zhang, Jingwei Sun, Li Feng, Cen Yao 等CHI 2024 · 被引用 38 次
- Do All Languages Cost the Same? Tokenization in the Era of Commercial Language ModelsOrevaoghene Ahia, Sachin Kumar, Hila Gonen, Jungo Kasai 等EMNLP 2023 · 被引用 24 次
- VAST: The Valence-Assessing Semantics Test for Contextualizing Language ModelsRobert Wolfe, Aylin CaliskanAAAI 2022 · 被引用 17 次
- Contrastive Visual Semantic Pretraining Magnifies the Semantics of Natural Language RepresentationsRobert Wolfe, Aylin CaliskanACL 2022 · 被引用 16 次
- Mitigating Frequency Bias and Anisotropy in Language Model Pre-Training with Syntactic SmoothingRichard Diehl Martinez, Zébulon Goriely, Andrew Caines, Paula Buttery 等EMNLP 2024 · 被引用 1 次
它引用的顶会 Paper4
- Interpreting Pretrained Contextualized Representations via Reductions to Static EmbeddingsRishi Bommasani, Kelly Davis, Claire CardieACL 2020 · 被引用 137 次
- Double-Hard Debias: Tailoring Word Embeddings for Gender Bias MitigationTianlu Wang, Xi Victoria Lin, Nazneen Fatema Rajani, Bryan McCann 等ACL 2020 · 被引用 42 次
- BPE-Dropout: Simple and Effective Subword RegularizationIvan Provilkov, Dmitrii Emelianenko, Elena VoitaACL 2020 · 被引用 17 次
- ValNorm Quantifies Semantics to Reveal Consistent Valence Biases Across Languages and Over CenturiesAutumn Toney, Aylin CaliskanEMNLP 2021
相关 Paper
- A Rose by Any Other Name would not Smell as Sweet: Social Bias in Names MistranslationSandra Sandoval, Jieyu Zhao, Marine Carpuat, Hal Daumé IIIEMNLP 2023 · 被引用 2 次
- On the Mutual Influence of Gender and Occupation in LLM RepresentationsHaozhe An, Connor Baumler, Abhilasha Sancheti, Rachel RudingerACL 2025
- Frequency Effects on Syntactic Rule Learning in TransformersJason Wei, Dan Garrette, Tal Linzen, Ellie PavlickEMNLP 2021 · 被引用 40 次
- "You Gotta be a Doctor, Lin" : An Investigation of Name-Based Bias of Large Language Models in Employment RecommendationsHuy Nghiem, John Prindle, Jieyu Zhao, Hal Daumé IIIEMNLP 2024 · 被引用 4 次
- Data Caricatures: On the Representation of African American Language in Pretraining CorporaNicholas Deas, Blake Vente, Amith Ananthram, Jessica Grieser 等ACL 2025
