Invariant Language Modeling
Maxime Peyrard, Sarvjeet Singh Ghotra, Martin Josifoski, Vidhan Agarwal, Barun Patra, Dean Carignan, Emre Kiciman, Saurabh Tiwary, Robert West
Abstract
Large pretrained language models are critical components of modern NLP pipelines. Yet, they suffer from spurious correlations, poor out-of-domain generalization, and biases. Inspired by recent progress in causal machine learning, in particular the invariant risk minimization (IRM) paradigm, we propose invariant language modeling, a framework for learning invariant representations that generalize better across multiple environments. In particular, we adapt a game-theoretic formulation of IRM (IRM-games) to language models, where the invariance emerges from a specific training schedule in which all the environments compete to optimize their own environmentspecific loss by updating subsets of the model in a round-robin fashion. We focus on controlled experiments to precisely demonstrate the ability of our method to (i) remove structured noise, (ii) ignore specific spurious correlations without affecting global performance, and (iii) achieve better out-of-domain generalization. These benefits come with a negligible computational overhead compared to standard training, do not require changing the local loss, and can be applied to any language model. We believe this framework is promising to help mitigate spurious correlations and biases in language models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d4920307-182c-4707-9ec2-29caf67e32eaCited by top-tier papers9
- Causal-Debias: Unifying Debiasing in Pretrained Language Models and Fine-tuning via Causal Invariant LearningFan Zhou, Yuzhou Mao, Liu Yu, Yi Yang et al.ACL 2023 · 21 citations
- PIDformer: Transformer Meets Control TheoryTam Minh Nguyen, César A. Uribe, Tan Minh Nguyen, Richard G. BaraniukICML 2024 · 13 citations
- Out-of-Distribution Generalization in Natural Language Processing: Past, Present, and FutureLinyi Yang, Yaoxian Song, Xuan Ren, Chenyang Lyu et al.EMNLP 2023 · 12 citations
- Fuse to Forget: Bias Reduction and Selective Memorization through Model FusionKerem Zaman, Leshem Choshen, Shashank SrivastavaEMNLP 2024 · 4 citations
- Causally Motivated Personalized Federated Invariant Learning with Shortcut-Averse Information-Theoretic RegularizationXueyang Tang, Song Guo, Jingcai Guo, Jie Zhang et al.ICML 2024 · 4 citations
Builds on6
- Domain Generalization with MixStyleKaiyang Zhou, Yongxin Yang, Yu Qiao, Tao XiangICLR 2021 · 986 citations
- Domain Generalization via Entropy RegularizationShanshan Zhao, Mingming Gong, Tongliang Liu, Huan Fu et al.NeurIPS 2020 · 327 citations
- Invariant Risk Minimization GamesKartik Ahuja, Karthikeyan Shanmugam, Kush R. Varshney, Amit DhurandharICML 2020 · 289 citations
- Evaluating the Robustness of Neural Language Models to Input PerturbationsMilad Moradi, Matthias SamwaldEMNLP 2021 · 64 citations
- Better than Average: Paired Evaluation of NLP systemsMaxime Peyrard, Wei Zhao, Steffen Eger, Robert WestACL 2021
Related papers
- Learning Optimal Features via Partial InvarianceMoulik Choraria, Ibtihal Ferwana, Ankur Mani, Lav R. VarshneyAAAI 2023 · 3 citations
- A Causal Marriage between VLM and IRM from Understanding to ReasoningZiliang Chen, Tianang Xiao, jusheng zhang, Yongsen Zheng et al.CVPR 2026
- Context is EnvironmentSharut Gupta, Stefanie Jegelka, David Lopez-Paz, Kartik AhujaICLR 2024
- IRM - when it works and when it doesn't: A test case of natural language inferenceYana Dranker, He He, Yonatan BelinkovNeurIPS 2021 · 22 citations
- What Is Missing in IRM Training and Evaluation? Challenges and SolutionsYihua Zhang, Pranay Sharma, Parikshit Ram, Mingyi Hong et al.ICLR 2023
