Graph Mixture of Experts: Learning on Large-Scale Graphs with Explicit Diversity Modeling
Haotao Wang, Ziyu Jiang, Yuning You, Yan Han, Gaowen Liu, Jayanth Srinivasa, Ramana Kompella, Zhangyang Wang
Abstract
Graph neural networks (GNNs) have found extensive applications in learning from graph data. However, real-world graphs often possess diverse structures and comprise nodes and edges of varying types. To bolster the generalization capacity of GNNs, it has become customary to augment training graph structures through techniques like graph augmentations and large-scale pre-training on a wider array of graphs. Balancing this diversity while avoiding increased computational costs and the notorious trainability issues of GNNs is crucial. This study introduces the concept of Mixture-of-Experts (MoE) to GNNs, with the aim of augmenting their capacity to adapt to a diverse range of training graph structures, without incurring explosive computational overhead. The proposed Graph Mixture of Experts (GMoE) model empowers individual nodes in the graph to dynamically and adaptively select more general information aggregation experts. These experts are trained to capture distinct subgroups of graph structures and to incorporate information with varying hop sizes, where those with larger hop sizes specialize in gathering information over longer distances. The effectiveness of GMoE is validated through a series of experiments on a diverse set of tasks, including graph, node, and link prediction, using the OGB benchmark. Notably, it enhances ROC-AUC by in ogbg-molhiv and by in ogbg-molbbbp, when compared to the non-MoE baselines. Our code is publicly available at https://github.com/VITA-Group/Graph-Mixture-of-Experts.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e87c11fb-8350-4137-bc0b-e4c74970e618Cited by top-tier papers31
- UniGraph2: Learning a Unified Embedding Space to Bind Multimodal GraphsYufei He, Yuan Sui, Xiaoxin He, Yue Liu et al.WWW 2025 · 37 citations
- GraphMETRO: Mitigating Complex Graph Distribution Shifts via Mixture of Aligned ExpertsShirley Wu, Kaidi Cao, Bruno Ribeiro, James Y. Zou et al.NeurIPS 2024 · 27 citations
- Mixture of In-Context Experts Enhance LLMs' Long Context AwarenessHongzhan Lin, Ang Lv, Yuhan Chen, Chen Zhu et al.NeurIPS 2024 · 25 citations
- Mixture of Link Predictors on GraphsLi Ma, Haoyu Han, Juanhui Li, Harry Shomer et al.NeurIPS 2024 · 23 citations
- Graph Mixture of Experts and Memory-augmented Routers for Multivariate Time Series Anomaly DetectionXiaoyu Huang, Weidong Chen, Bo Hu, Zhendong MaoAAAI 2025 · 22 citations
Builds on23
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong et al.NeurIPS 2020 · 3,935 citations
- Graph Contrastive Learning with AugmentationsYuning You, Tianlong Chen, Yongduo Sui, Ting Chen et al.NeurIPS 2020 · 3,042 citations
- GShard: Scaling Giant Models with Conditional Computation and Automatic ShardingDmitry Lepikhin, HyoukJoong Lee, Yuanzhong Xu, Dehao Chen et al.ICLR 2021 · 1,954 citations
- Strategies for Pre-training Graph Neural NetworksWeihua Hu, Bowen Liu, Joseph Gomes, Marinka Zitnik et al.ICLR 2020 · 1,744 citations
- Do Transformers Really Perform Badly for Graph Representation?Chengxuan Ying, Tianle Cai, Shengjie Luo, Shuxin Zheng et al.NeurIPS 2021 · 1,632 citations
Related papers
- One For All: Achieving Adaptive Graph Neural Networks via Mixture of Message PassingZhaojun Luo, Jintang Li, Yuchang Zhu, Yun Fu et al.KDD 2026
- Graph Sparsification via Mixture of GraphsGuibin Zhang, Xiangguo Sun, Yanwei Yue, Chonghe Jiang et al.ICLR 2025
- Self-Adaptive Graph Mixture of ModelsMohit Meena, Yash Punjabi, Abhishek A, Vishal Sharma et al.AAAI 2026
- Discriminative Mixture-of-Experts on Graphs with Reliable Expert FusionHaoyue Deng, Menghui Wang, Yunlong Zhou, Jingyi Liu et al.ICML 2026
- Mixture of Scope Experts at Test: Generalizing Deeper Graph Neural Networks with Shallow VariantsGangda Deng, Hongkuan Zhou, Rajgopal Kannan, Viktor PrasannaNeurIPS 2025 · 1 citation
