Feature-Level Debiased Natural Language Understanding
Yougang Lyu, Piji Li, Yechang Yang, Maarten de Rijke, Pengjie Ren, Yukun Zhao, Dawei Yin, Zhaochun Ren
摘要
Natural language understanding (NLU) models often rely on dataset biases rather than intended task-relevant features to achieve high performance on specific datasets. As a result, these models perform poorly on datasets outside the training distribution. Some recent studies address this issue by reducing the weights of biased samples during the training process. However, these methods still encode biased latent features in representations and neglect the dynamic nature of bias, which hinders model prediction. We propose an NLU debiasing method, named debiasing contrastive learning (DCT), to simultaneously alleviate the above problems based on contrastive learning. We devise a debiasing, positive sampling strategy to mitigate biased latent features by selecting the least similar biased positive samples. We also propose a dynamic negative sampling strategy to capture the dynamic influence of biases by employing a bias-only model to dynamically select the most similar biased negative samples. We conduct experiments on three NLU benchmark datasets. Experimental results show that DCT outperforms state-of-the-art baselines on out-of-distribution datasets while maintaining in-distribution performance. We also verify that DCT can reduce biased latent features from the model's representations.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Navigate Beyond Shortcuts: Debiased Learning through the Lens of Neural CollapseYining Wang, Junjie Sun, Chenyue Wang, Mi Zhang 等CVPR 2024 · 被引用 6 次
- Improving Bias Mitigation through Bias Experts in Natural Language UnderstandingEojin Jeon, Mingyu Lee, Juhyeong Park, Yeachan Kim 等EMNLP 2023 · 被引用 3 次
- KnowTuning: Knowledge-aware Fine-tuning for Large Language ModelsYougang Lyu, Lingyong Yan, Shuaiqiang Wang, Haibo Shi 等EMNLP 2024 · 被引用 3 次
- IBADR: an Iterative Bias-Aware Dataset Refinement Framework for Debiasing NLU modelsXiaoyue Wang, Xin Liu, Lijie Wang, Yaoxiang Wang 等EMNLP 2023 · 被引用 2 次
- FairFlow: Mitigating Dataset Biases through Undecided Learning for Natural Language UnderstandingJiali Cheng, Hadi AmiriEMNLP 2024 · 被引用 2 次
它引用的顶会 Paper18
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Supervised Contrastive LearningPrannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna 等NeurIPS 2020 · 被引用 7,049 次
- WinoGrande: An Adversarial Winograd Schema Challenge at ScaleKeisuke Sakaguchi, Ronan Le Bras, Chandra Bhagavatula, Yejin ChoiAAAI 2020 · 被引用 3,037 次
- SimCSE: Simple Contrastive Learning of Sentence EmbeddingsTianyu Gao, Xingcheng Yao, Danqi ChenEMNLP 2021 · 被引用 2,496 次
- Climbing towards NLU: On Meaning, Form, and Understanding in the Age of DataEmily M. Bender, Alexander KollerACL 2020 · 被引用 914 次
相关 Paper
- Towards Stable Natural Language Understanding via Information Entropy Guided DebiasingLi Du, Xiao Ding, Zhouhao Sun, Ting Liu 等ACL 2023 · 被引用 1 次
- End-to-End Bias Mitigation by Modelling Biases in CorporaRabeeh Karimi Mahabadi, Yonatan Belinkov, James HendersonACL 2020 · 被引用 136 次
- Unbiased Classification through Bias-Contrastive and Bias-Balanced LearningYoungkyu Hong, Eunho YangNeurIPS 2021 · 被引用 94 次
- Counterexample Contrastive Learning for Spurious Correlation EliminationJinqiang Wang, Rui Hu, Chaoquan Jiang, Rui Hu 等ACM MM 2022 · 被引用 3 次
- Mind the Trade-off: Debiasing NLU Models without Degrading the In-distribution PerformancePrasetya Ajie Utama, Nafise Sadat Moosavi, Iryna GurevychACL 2020 · 被引用 11 次
