Harms of Gender Exclusivity and Challenges in Non-Binary Representation in Language Technologies
Sunipa Dev, Masoud Monajatipoor, Anaelia Ovalle, Arjun Subramonian, Jeff M. Phillips, Kai-Wei Chang
摘要
Content Warning: This paper contains examples of stereotypes and associations, misgendering, erasure, and other harms that could be offensive and triggering to trans and nonbinary individuals. Gender is widely discussed in the context of language tasks and when examining the stereotypes propagated by language models. However, current discussions primarily treat gender as binary, which can perpetuate harms such as the cyclical erasure of non-binary gender identities. These harms are driven by model and dataset biases, which are consequences of the non-recognition and lack of understanding of non-binary genders in society. In this paper, we explain the complexity of gender and language around it, and survey non-binary persons to understand harms associated with the treatment of gender as binary in English language technologies. We also detail how current language representations (e.g., GloVe, BERT) capture and perpetuate these harms and related challenges that need to be acknowledged and addressed for representations to equitably encode gender information.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper36
- GLM-130B: An Open Bilingual Pre-trained ModelAohan Zeng, Xiao Liu, Zhengxiao Du, Zihan Wang 等ICLR 2023 · 被引用 295 次
- QuRating: Selecting High-Quality Data for Training Language ModelsAlexander Wettig, Aatmik Gupta, Saumya Malik, Danqi ChenICML 2024 · 被引用 138 次
- HRS-Bench: Holistic, Reliable and Scalable Benchmark for Text-to-Image ModelsEslam Mohamed Bakr, Pengzhan Sun, Xiaoqian Shen, Faizan Farooq Khan 等ICCV 2023 · 被引用 115 次
- Designing Responsible AI: Adaptations of UX Practice to Meet Responsible AI ChallengesQiaosi Wang, Michael Madaio, Shaun K. Kane, Shivani Kapania 等CHI 2023 · 被引用 90 次
- Controllable Text Generation with Neurally-Decomposed OracleTao Meng, Sidi Lu, Nanyun Peng, Kai-Wei ChangNeurIPS 2022 · 被引用 45 次
它引用的顶会 Paper8
- Balanced Datasets Are Not Enough: Estimating and Mitigating Gender Bias in Deep Image RepresentationsTianlu Wang, Jieyu Zhao, Mark Yatskar, Kai-Wei Chang 等ICCV 2019 · 被引用 469 次
- On Measuring and Mitigating Biased Inferences of Word EmbeddingsSunipa Dev, Tao Li, Jeff M. Phillips, Vivek SrikumarAAAI 2020 · 被引用 195 次
- Language (Technology) is Power: A Critical Survey of "Bias" in NLPSu Lin Blodgett, Solon Barocas, Hal Daumé III, Hanna M. WallachACL 2020 · 被引用 68 次
- Revisiting Gendered Web Forms: An Evaluation of Gender Inputs with (Non-)Binary PeopleMorgan Klaus Scheuerman, Jialun Aaron Jiang, Katta Spiel, Jed R. BrubakerCHI 2021 · 被引用 42 次
- OSCaR: Orthogonal Subspace Correction and Rectification of Biases in Word EmbeddingsSunipa Dev, Tao Li, Jeff M. Phillips, Vivek SrikumarEMNLP 2021 · 被引用 31 次
相关 Paper
- MISGENDERED: Limits of Large Language Models in Understanding PronounsTamanna Hossain, Sunipa Dev, Sameer SinghACL 2023 · 被引用 8 次
- "Fifty Shades of Bias": Normative Ratings of Gender Bias in GPT Generated English TextRishav Hada, Agrima Seth, Harshita Diddee, Kalika BaliEMNLP 2023 · 被引用 10 次
- Toward Gender-Inclusive Coreference ResolutionYang Trista Cao, Hal Daumé IIIACL 2020 · 被引用 20 次
- Detecting Gender Stereotypes: Lexicon vs. Supervised Learning MethodsJenna Cryan, Shiliang Tang, Xinyi Zhang, Miriam J. Metzger 等CHI 2020 · 被引用 43 次
- Gendered Mental Health Stigma in Masked Language ModelsInna W. Lin, Lucille Njoo, Anjalie Field, Ashish Sharma 等EMNLP 2022 · 被引用 15 次
