Learning Graph Invariance by Harnessing Spuriosity
Tianjun Yao, Yongqiang Chen, Kai Hu, Tongliang Liu, Kun Zhang, Zhiqiang Shen
Abstract
Recently, graph invariant learning has become the de facto approach to tackle the Out-of-Distribution (OOD) generalization failure in graph representation learning. They generally follow the framework of invariant risk minimization to capture the invariance of graph data from different environments. Despite some success, it remains unclear to what extent existing approaches have captured invariant features for OOD generalization on graphs. In this work, we find that representative OOD methods such as IRM and VRex, and their variants on graph invariant learning may have captured a limited set of invariant features. To tackle this challenge, we propose LIRS, a novel learning framework designed to Learn graph Invariance by Removing Spurious features. Different from most existing approaches that directly learn the invariant features, LIRS takes an indirect approach by first learning the spurious features and then removing them from the ERM-learned features. We demonstrate that learning the invariant graph features in an indirect way enables the model to capture a more comprehensive set of invariant features, leading to better OOD generalization performance in novel environments. Notably, LIRS surpasses the second-best method by as much as 25.50% across all competitive baselines, underscoring its efficacy in OOD generalization. 1 Assumption 1 posits that one or more substructure patterns in G c are not only stably associated with the target label Y across different environments but also possess sufficient predictive power to accurately determine Y . This assumption is well-aligned with real-world scenarios. For instance, in the GOODHIV dataset (Gui et al., 2022; Hu et al., 2020; Wu et al., 2018) , a molecule's ability to inhibit HIV may depend on the presence of several functional groups interacting with various parts of the virus. Moreover, recent study also provides empirical evidence for graph applications, suggesting that multiple substructures remain stable and predictive of the targets (see Appendix E.2 in Bui et al. ( 2024 )), thereby supporting the validity of Assumption 1. Therefore, when the OOD algorithms are able to learn a broader set of invariant substructures (features), they will generalize more effectively across different environments.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 98a0b7b6-2adf-4b0a-9c11-a97d9ebc0322Cited by top-tier papers5
- Quantifying Distributional Invariance in Causal Subgraph for IRM-Free Graph GeneralizationYang Qiu, Yixiong Zou, Jun Wang, Wei Liu et al.NeurIPS 2025 · 3 citations
- Hierarchical Graph Tokenization for Molecule-Language AlignmentYongqiang Chen, Quanming Yao, Juzheng Zhang, James Cheng et al.ICML 2025 · 2 citations
- Rethinking Graph Generalization through the Lens of Sharpness-Aware MinimizationYang Qiu, Yixiong Zou, Jun WangWWW 2026
- Pruning Spurious Subgraphs for Graph Out-of-Distribution GeneralizationTianjun Yao, Haoxuan Li, Yongqiang Chen, Tongliang Liu et al.NeurIPS 2025
- Diverse and Sparse Mixture-of-Experts for Causal Subgraph-Based Out-of-Distribution Graph LearningJerry Sun, Mohamed Abubakr Hassan, Yaoyu Zhang, Wanying Zhang et al.ICLR 2026
Builds on45
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Supervised Contrastive LearningPrannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna et al.NeurIPS 2020 · 7,049 citations
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong et al.NeurIPS 2020 · 3,935 citations
- Graph Contrastive Learning with AugmentationsYuning You, Tianlong Chen, Yongduo Sui, Ting Chen et al.NeurIPS 2020 · 3,042 citations
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie et al.ICML 2021 · 1,773 citations
Related papers
- Does Invariant Graph Learning via Environment Augmentation Learn Invariance?Yongqiang Chen, Yatao Bian, Kaiwen Zhou, Binghui Xie et al.NeurIPS 2023 · 71 citations
- A Unified Invariant Learning Framework for Graph ClassificationYongduo Sui, Jie Sun, Shuyao Wang, Zemin Liu et al.KDD 2025 · 1 citation
- A Structure-aware Invariant Learning Framework for Node-level Graph OOD GeneralizationRuiwen Yuan, Yongqiang Tang, Wensheng ZhangKDD 2025 · 4 citations
- Dissecting the Failure of Invariant Learning on GraphsQixun Wang, Yifei Wang, Yisen Wang, Xianghua YingNeurIPS 2024 · 10 citations
- Empowering Graph Invariance Learning with Deep Spurious InfomaxTianjun Yao, Yongqiang Chen, Zhenhao Chen, Kai Hu et al.ICML 2024 · 20 citations
