Type B Reflexivization as an Unambiguous Testbed for Multilingual Multi-Task Gender Bias
Ana Valeria González-Garduño, Maria Barrett, Rasmus Hvingelby, Kellie Webster, Anders Søgaard
摘要
The one-sided focus on English in previous studies of gender bias in NLP misses out on opportunities in other languages: English challenge datasets such as GAP and Wino-Gender highlight model preferences that are "hallucinatory", e.g., disambiguating genderambiguous occurrences of 'doctor' as male doctors. We show that for languages with type B reflexivization, e.g., Swedish and Russian, we can construct multi-task challenge datasets for detecting gender bias that lead to unambiguously wrong model predictions: In these languages, the direct translation of 'the doctor removed his mask' is not ambiguous between a coreferential reading and a disjoint reading. Instead, the coreferential reading requires a non-gendered pronoun, and the gendered, possessive pronouns are anti-reflexive. We present a multilingual, multi-task challenge dataset, which spans four languages and four NLP tasks and focuses only on this phenomenon. We find evidence for gender bias across all task-language combinations and correlate model bias with national labor market statistics.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Under the Morphosyntactic Lens: A Multifaceted Evaluation of Gender Bias in Speech TranslationBeatrice Savoldi, Marco Gaido, Luisa Bentivogli, Matteo Negri 等ACL 2022 · 被引用 30 次
- Investigating Failures of Automatic Translationin the Case of Unambiguous GenderAdi Renduchintala, Adina WilliamsACL 2022 · 被引用 28 次
- What the Harm? Quantifying the Tangible Impact of Gender Bias in Machine Translation with a Human-centered StudyBeatrice Savoldi, Sara Papi, Matteo Negri, Ana Guerberof Arenas 等EMNLP 2024 · 被引用 1 次
相关 Paper
- Reducing Gender Bias in Neural Machine Translation as a Domain Adaptation ProblemDanielle Saunders, Bill ByrneACL 2020 · 被引用 7 次
- A Survey on Zero Pronoun TranslationLongyue Wang, Siyou Liu, Mingzhou Xu, Linfeng Song 等ACL 2023 · 被引用 5 次
- Superlim: A Swedish Language Understanding Evaluation BenchmarkAleksandrs Berdicevskis, Gerlof Bouma, Robin Kurtz, Felix Morger 等EMNLP 2023 · 被引用 2 次
- Gender Bias in Multilingual Embeddings and Cross-Lingual TransferJieyu Zhao, Subhabrata Mukherjee, Saghar Hosseini, Kai-Wei Chang 等ACL 2020 · 被引用 59 次
- MORPHOGEN: A Multilingual Benchmark for Evaluating Gender-Aware Morphological GenerationMehul Agarwal, Aditya Aggarwal, Arnav Goel, Medha Hira 等ACL 2026
