Lune

ACL2023顶会

Exploiting Biased Models to De-bias Text: A Gender-Fair Rewriting Model

Chantal Amrhein, Florian Schottmann, Rico Sennrich, Samuel Läubli

2023年份
7被引次数
4顶会引用

摘要

Natural language generation models reproduce and often amplify the biases present in their training data. Previous research explored using sequence-to-sequence rewriting models to transform biased model outputs (or original texts) into more gender-fair language by creating pseudo training data through linguistic rules. However, this approach is not practical for languages with more complex morphology than English. We hypothesise that creating training data in the reverse direction, i.e. starting from gender-fair text, is easier for morphologically complex languages and show that it matches the performance of state-of-the-art rewriting models for English. To eliminate the rule-based nature of data creation, we instead propose using machine translation models to create gender-biased text from real gender-fair text via round-trip translation. Our approach allows us to train a rewriting model for German without the need for elaborate handcrafted rules. The outputs of this model increased genderfairness as shown in a human evaluation study. 1 * Work done during an internship at Textshuttle. 1 We publicly release our data and code here: https:// github.com/textshuttle/exploiting-bias-to-debias Issue 1: Rewritings that affect other dependencies, e.g. when rewriting third-person subject pronouns in English, verbs in present tense need to be pluralised (e.g. "she knows" rewritten as "they know").

问问这篇 Paper

智能体会读完全文。

Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。

可以从这些问题问起

智能体调用

Luneget_paper_fulltext

在 Lune 里问

免费开始,无需绑卡

引用它的顶会 Paper4

问问它们各自怎么用它

它引用的顶会 Paper6

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖