Mixture of Weak and Strong Experts on Graphs
Hanqing Zeng, Hanjia Lyu, Diyi Hu, Yinglong Xia, Jiebo Luo
摘要
Realistic graphs contain both (1) rich self-features of nodes and (2) informative structures of neighborhoods, jointly handled by a Graph Neural Network (GNN) in the typical setup. We propose to decouple the two modalities by Mixture of weak and strong experts (Mowst), where the weak expert is a light-weight Multilayer Perceptron (MLP), and the strong expert is an off-the-shelf GNN. To adapt the experts' collaboration to different target nodes, we propose a "confidence" mechanism based on the dispersion of the weak expert's prediction logits. The strong expert is conditionally activated in the low-confidence region when either the node's classification relies on neighborhood information, or the weak expert has low model quality. We reveal interesting training dynamics by analyzing the influence of the confidence function on loss: our training algorithm encourages the specialization of each expert by effectively generating soft splitting of the graph. In addition, our "confidence" design imposes a desirable bias toward the strong expert to benefit from GNN's better generalization capability. Mowst is easy to optimize and achieves strong expressive power, with a computation cost comparable to a single GNN. Empirically, Mowst on 4 backbone GNN architectures show significant accuracy improvement on 6 standard node classification benchmarks, including both homophilous and heterophilous graphs (https://github.com/facebookresearch/mowst-gnn).
Strong Expert (GNN)
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- From Coarse to Fine: Enable Comprehensive Graph Self-supervised Learning with Multi-granular Semantic EnsembleQianlong Wen, Mingxuan Ju, Zhongyu Ouyang, Chuxu Zhang 等ICML 2024 · 被引用 9 次
- Enhanced Expert Merging for Mixture-of-Experts in Graph Foundation ModelsLei Liu, Xingyu Xia, Qianqian Xie, Ben Liu 等NeurIPS 2025 · 被引用 4 次
- S'MoRE: Structural Mixture of Residual Experts for Parameter-Efficient LLM Fine-tuningHanqing Zeng, Yinglong Xia, Zhuokai Zhao, Chuan Jiang 等NeurIPS 2025 · 被引用 3 次
- Unifying and Enhancing Graph Transformers via a Hierarchical Mask FrameworkYujie Xing, Xiao Wang, Bin Wu, Hai Huang 等NeurIPS 2025 · 被引用 2 次
- Mixture of Scope Experts at Test: Generalizing Deeper Graph Neural Networks with Shallow VariantsGangda Deng, Hongkuan Zhou, Rajgopal Kannan, Viktor PrasannaNeurIPS 2025 · 被引用 1 次
它引用的顶会 Paper30
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong 等NeurIPS 2020 · 被引用 3,935 次
- GShard: Scaling Giant Models with Conditional Computation and Automatic ShardingDmitry Lepikhin, HyoukJoong Lee, Yuanzhong Xu, Dehao Chen 等ICLR 2021 · 被引用 1,954 次
- Simple and Deep Graph Convolutional NetworksMing Chen, Zhewei Wei, Zengfeng Huang, Bolin Ding 等ICML 2020 · 被引用 1,910 次
- How Attentive are Graph Attention Networks?Shaked Brody, Uri Alon, Eran YahavICLR 2022 · 被引用 1,717 次
- Beyond Homophily in Graph Neural Networks: Current Limitations and Effective DesignsJiong Zhu, Yujun Yan, Lingxiao Zhao, Mark Heimann 等NeurIPS 2020 · 被引用 1,490 次
相关 Paper
- One For All: Achieving Adaptive Graph Neural Networks via Mixture of Message PassingZhaojun Luo, Jintang Li, Yuchang Zhu, Yun Fu 等KDD 2026
- Mixture of Link Predictors on GraphsLi Ma, Haoyu Han, Juanhui Li, Harry Shomer 等NeurIPS 2024 · 被引用 23 次
- Graph Mixture of Experts: Learning on Large-Scale Graphs with Explicit Diversity ModelingHaotao Wang, Ziyu Jiang, Yuning You, Yan Han 等NeurIPS 2023 · 被引用 104 次
- Discriminative Mixture-of-Experts on Graphs with Reliable Expert FusionHaoyue Deng, Menghui Wang, Yunlong Zhou, Jingyi Liu 等ICML 2026
- Self-Adaptive Graph Mixture of ModelsMohit Meena, Yash Punjabi, Abhishek A, Vishal Sharma 等AAAI 2026
