Exploiting Biased Models to De-bias Text: A Gender-Fair Rewriting Model
Chantal Amrhein, Florian Schottmann, Rico Sennrich, Samuel Läubli
Abstract
Natural language generation models reproduce and often amplify the biases present in their training data. Previous research explored using sequence-to-sequence rewriting models to transform biased model outputs (or original texts) into more gender-fair language by creating pseudo training data through linguistic rules. However, this approach is not practical for languages with more complex morphology than English. We hypothesise that creating training data in the reverse direction, i.e. starting from gender-fair text, is easier for morphologically complex languages and show that it matches the performance of state-of-the-art rewriting models for English. To eliminate the rule-based nature of data creation, we instead propose using machine translation models to create gender-biased text from real gender-fair text via round-trip translation. Our approach allows us to train a rewriting model for German without the need for elaborate handcrafted rules. The outputs of this model increased genderfairness as shown in a human evaluation study. 1 * Work done during an internship at Textshuttle. 1 We publicly release our data and code here: https:// github.com/textshuttle/exploiting-bias-to-debias Issue 1: Rewritings that affect other dependencies, e.g. when rewriting third-person subject pronouns in English, verbs in present tense need to be pluralised (e.g. "she knows" rewritten as "they know").
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 928fa110-234e-4ad3-b374-f0b558691c2cCited by top-tier papers4
- Hi Guys or Hi Folks? Benchmarking Gender-Neutral Machine Translation with the GeNTE CorpusAndrea Piergentili, Beatrice Savoldi, Dennis Fucci, Matteo Negri et al.EMNLP 2023 · 3 citations
- What the Harm? Quantifying the Tangible Impact of Gender Bias in Machine Translation with a Human-centered StudyBeatrice Savoldi, Sara Papi, Matteo Negri, Ana Guerberof Arenas et al.EMNLP 2024 · 1 citation
- The Lou Dataset - Exploring the Impact of Gender-Fair Language in German Text ClassificationAndreas Waldis, Joel Birrer, Anne Lauscher, Iryna GurevychEMNLP 2024 · 1 citation
- Mind the Inclusivity Gap: Multilingual Gender-Neutral Translation Evaluation with mGeNTEBeatrice Savoldi, Giuseppe Attanasio, Eleonora Cupin, Eleni Gkovedarou et al.EMNLP 2025
Builds on6
- Perturbation Augmentation for Fairer NLPRebecca Qian, Candace Ross, Jude Fernandes, Eric Michael Smith et al.EMNLP 2022 · 54 citations
- Investigating Failures of Automatic Translationin the Case of Unambiguous GenderAdi Renduchintala, Adina WilliamsACL 2022 · 28 citations
- Dialect-robust Evaluation of Generated TextJiao Sun, Thibault Sellam, Elizabeth Clark, Tu Vu et al.ACL 2023 · 11 citations
- Reducing Gender Bias in Neural Machine Translation as a Domain Adaptation ProblemDanielle Saunders, Bill ByrneACL 2020 · 7 citations
- StereoSet: Measuring stereotypical bias in pretrained language modelsMoin Nadeem, Anna Bethke, Siva ReddyACL 2021
Related papers
- GFST: Gender-Filtered Self-Training for More Accurate Gender in TranslationPrafulla Kumar Choubey, Anna Currey, Prashant Mathur, Georgiana DinuEMNLP 2021 · 7 citations
- A Tale of Pronouns: Interpretability Informs Gender Bias Mitigation for Fairer Instruction-Tuned Machine TranslationGiuseppe Attanasio, Flor Miriam Plaza del Arco, Debora Nozza, Anne LauscherEMNLP 2023 · 3 citations
- MT-GenEval: A Counterfactual and Contextual Dataset for Evaluating Gender Accuracy in Machine TranslationAnna Currey, Maria Nadejde, Raghavendra Reddy Pappagari, Mia Mayer et al.EMNLP 2022 · 22 citations
- Measuring and Mitigating Name Biases in Neural Machine TranslationJun Wang, Benjamin I. P. Rubinstein, Trevor CohnACL 2022 · 31 citations
- Gender in Danger? Evaluating Speech Translation Technology on the MuST-SHE CorpusLuisa Bentivogli, Beatrice Savoldi, Matteo Negri, Mattia Antonino Di Gangi et al.ACL 2020 · 40 citations
