Discovering Differences in the Representation of People using Contextualized Semantic Axes
Li Lucy, Divya Tadimeti, David Bamman
Abstract
A common paradigm for identifying semantic differences across social and temporal contexts is the use of static word embeddings and their distances. In particular, past work has compared embeddings against "semantic axes" that represent two opposing concepts. We extend this paradigm to BERT embeddings, and construct contextualized axes that mitigate the pitfall where antonyms have neighboring representations. We validate and demonstrate these axes on two people-centric datasets: occupations from Wikipedia, and multi-platform discussions in extremist, men's communities over fourteen years. In both studies, contextualized semantic axes can characterize differences among instances of the same word type. In the latter study, we show that references to women and the contexts around them have become more detestable over time.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4ee4b76c-4a04-420f-af7d-650ebab1747bCited by top-tier papers4
- CoMPosT: Characterizing and Evaluating Caricature in LLM SimulationsMyra Cheng, Tiziano Piccardi, Diyi YangEMNLP 2023 · 33 citations
- WikiBio: a Semantic Resource for the Intersectional Analysis of Biographical EventsMarco Antonio Stranisci, Rossana Damiano, Enrico Mensa, Viviana Patti et al.ACL 2023 · 2 citations
- Adaptive Axes: A Pipeline for In-domain Social Stereotype AnalysisQingcheng Zeng, Mingyu Jin, Rob VoigtEMNLP 2024
- Who Holds the Pen? Caricature and Perspective in LLM Retellings of HistoryLubna Zahan Lamia, Mabsur Fatin Bin Hossain, Md. Mosaddek KhanEMNLP 2025
Builds on12
- Interpreting Pretrained Contextualized Representations via Reductions to Static EmbeddingsRishi Bommasani, Kelly Davis, Claire CardieACL 2020 · 137 citations
- Analysing Lexical Semantic Change with Contextualised Word RepresentationsMario Giulianelli, Marco Del Tredici, Raquel FernándezACL 2020 · 118 citations
- Do Platform Migrations Compromise Content Moderation? Evidence from r/The_Donald and r/IncelsManoel Horta Ribeiro, Shagun Jhaver, Savvas Zannettou, Jeremy Blackburn et al.CSCW 2021 · 102 citations
- Don't Stop Pretraining: Adapt Language Models to Domains and TasksSuchin Gururangan, Ana Marasovic, Swabha Swayamdipta, Kyle Lo et al.ACL 2020 · 93 citations
- "Go eat a bat, Chang!": On the Emergence of Sinophobic Behavior on Web Communities in the Face of COVID-19Fatemeh Tahmasbi, Leonard Schild, Chen Ling, Jeremy Blackburn et al.WWW 2021 · 92 citations
Related papers
- Measure and Evaluation of Semantic Divergence across Two LanguagesSyrielle Montariol, Alexandre AllauzenACL 2021
- Inducing lexicons of in-group language with socio-temporal contextChristine de KockACL 2025
- Towards Debiasing Sentence RepresentationsPaul Pu Liang, Irene Mengze Li, Emily Zheng, Yao Chong Lim et al.ACL 2020 · 149 citations
- Dynamic Contextualized Word EmbeddingsValentin Hofmann, Janet B. Pierrehumbert, Hinrich SchützeACL 2021
- A Multidimensional Framework for Evaluating Lexical Semantic Change with Social Science ApplicationsNaomi Baes, Nick Haslam, Ekaterina VylomovaACL 2024 · 5 citations
