Discovering Differences in the Representation of People using Contextualized Semantic Axes
Li Lucy, Divya Tadimeti, David Bamman
摘要
A common paradigm for identifying semantic differences across social and temporal contexts is the use of static word embeddings and their distances. In particular, past work has compared embeddings against "semantic axes" that represent two opposing concepts. We extend this paradigm to BERT embeddings, and construct contextualized axes that mitigate the pitfall where antonyms have neighboring representations. We validate and demonstrate these axes on two people-centric datasets: occupations from Wikipedia, and multi-platform discussions in extremist, men's communities over fourteen years. In both studies, contextualized semantic axes can characterize differences among instances of the same word type. In the latter study, we show that references to women and the contexts around them have become more detestable over time.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- CoMPosT: Characterizing and Evaluating Caricature in LLM SimulationsMyra Cheng, Tiziano Piccardi, Diyi YangEMNLP 2023 · 被引用 33 次
- WikiBio: a Semantic Resource for the Intersectional Analysis of Biographical EventsMarco Antonio Stranisci, Rossana Damiano, Enrico Mensa, Viviana Patti 等ACL 2023 · 被引用 2 次
- Adaptive Axes: A Pipeline for In-domain Social Stereotype AnalysisQingcheng Zeng, Mingyu Jin, Rob VoigtEMNLP 2024
- Who Holds the Pen? Caricature and Perspective in LLM Retellings of HistoryLubna Zahan Lamia, Mabsur Fatin Bin Hossain, Md. Mosaddek KhanEMNLP 2025
它引用的顶会 Paper12
- Interpreting Pretrained Contextualized Representations via Reductions to Static EmbeddingsRishi Bommasani, Kelly Davis, Claire CardieACL 2020 · 被引用 137 次
- Analysing Lexical Semantic Change with Contextualised Word RepresentationsMario Giulianelli, Marco Del Tredici, Raquel FernándezACL 2020 · 被引用 118 次
- Do Platform Migrations Compromise Content Moderation? Evidence from r/The_Donald and r/IncelsManoel Horta Ribeiro, Shagun Jhaver, Savvas Zannettou, Jeremy Blackburn 等CSCW 2021 · 被引用 102 次
- Don't Stop Pretraining: Adapt Language Models to Domains and TasksSuchin Gururangan, Ana Marasovic, Swabha Swayamdipta, Kyle Lo 等ACL 2020 · 被引用 93 次
- "Go eat a bat, Chang!": On the Emergence of Sinophobic Behavior on Web Communities in the Face of COVID-19Fatemeh Tahmasbi, Leonard Schild, Chen Ling, Jeremy Blackburn 等WWW 2021 · 被引用 92 次
相关 Paper
- Measure and Evaluation of Semantic Divergence across Two LanguagesSyrielle Montariol, Alexandre AllauzenACL 2021
- Inducing lexicons of in-group language with socio-temporal contextChristine de KockACL 2025
- Towards Debiasing Sentence RepresentationsPaul Pu Liang, Irene Mengze Li, Emily Zheng, Yao Chong Lim 等ACL 2020 · 被引用 149 次
- Dynamic Contextualized Word EmbeddingsValentin Hofmann, Janet B. Pierrehumbert, Hinrich SchützeACL 2021
- A Multidimensional Framework for Evaluating Lexical Semantic Change with Social Science ApplicationsNaomi Baes, Nick Haslam, Ekaterina VylomovaACL 2024 · 被引用 5 次
