It's Morphin' Time! Combating Linguistic Discrimination with Inflectional Perturbations
Samson Tan, Shafiq R. Joty, Min-Yen Kan, Richard Socher
2020年份
88被引次数
21顶会引用
摘要
Training on only perfect Standard English corpora predisposes pre-trained neural networks to discriminate against minorities from nonstandard linguistic backgrounds (e.g., African American Vernacular English, Colloquial Singapore English, etc.). We perturb the inflectional morphology of words to craft plausible and semantically similar adversarial examples that expose these biases in popular NLP models, e.g., BERT and Transformer, and show that adversarially fine-tuning them for a single epoch significantly improves robustness without sacrificing performance on clean data. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper21
- Tree of Attacks: Jailbreaking Black-Box LLMs AutomaticallyAnay Mehrotra, Manolis Zampetakis, Paul Kassianik, Blaine Nelson 等NeurIPS 2024 · 被引用 835 次
- Incorporating Hierarchy into Text Encoder: a Contrastive Learning Approach for Hierarchical Text ClassificationZihan Wang, Peiyi Wang, Lianzhe Huang, Xin Sun 等ACL 2022 · 被引用 157 次
- Spinning Language Models: Risks of Propaganda-As-A-Service and CountermeasuresEugene Bagdasaryan, Vitaly ShmatikovS&P 2022 · 被引用 94 次
- Prompting GPT-3 To Be ReliableChenglei Si, Zhe Gan, Zhengyuan Yang, Shuohang Wang 等ICLR 2023 · 被引用 68 次
- VALUE: Understanding Dialect Disparity in NLUCaleb Ziems, Jiaao Chen, Camille Harris, Jessica Anderson 等ACL 2022 · 被引用 57 次
相关 Paper
- Mind Your Inflections! Improving NLP for Non-Standard Englishes with Base-Inflection EncodingSamson Tan, Shafiq R. Joty, Lav R. Varshney, Min-Yen KanEMNLP 2020 · 被引用 26 次
- Evaluation of African American Language Bias in Natural Language GenerationNicholas Deas, Jessica Grieser, Shana Kleiner, Desmond Patton 等EMNLP 2023 · 被引用 15 次
- The King Is Naked: On the Notion of Robustness for Natural Language ProcessingEmanuele La Malfa, Marta KwiatkowskaAAAI 2022 · 被引用 31 次
- Token-Aware Virtual Adversarial Training in Natural Language UnderstandingLinyang Li, Xipeng QiuAAAI 2021 · 被引用 54 次
- DADA: Dialect Adaptation via Dynamic Aggregation of Linguistic RulesYanchen Liu, William Barr Held, Diyi YangEMNLP 2023 · 被引用 6 次
