Grain: Improving Data Efficiency of Graph Neural Networks via Diversified Influence Maximization
Wentao Zhang, Zhi Yang, Yexin Wang, Yu Shen, Yang Li, Liang Wang, Bin Cui
Abstract
Data selection methods, such as active learning and core-set selection, are useful tools for improving the data efficiency of deep learning models on large-scale datasets. However, recent deep learning models have moved forward from independent and identically distributed data to graph-structured data, such as social networks, e-commerce user-item graphs, and knowledge graphs. This evolution has led to the emergence of Graph Neural Networks (GNNs) that go beyond the models existing data selection methods are designed for. Therefore, we present GRAIN, an efficient framework that opens up a new perspective through connecting data selection in GNNs with social influence maximization. By exploiting the common patterns of GNNs, GRAIN introduces a novel feature propagation concept, a diversified influence maximization objective with novel influence and diversity functions, and a greedy algorithm with an approximation guarantee into a unified framework. Empirical studies on public datasets demonstrate that GRAIN significantly improves both the performance and efficiency of data selection (including active learning and core-set selection) for GNNs. To the best of our knowledge, this is the first attempt to bridge two largely parallel threads of research, data selection, and social influence maximization, in the setting of GNNs, paving new ways for improving data efficiency.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers14
- Zebra: When Temporal Graph Neural Networks Meet Temporal Personalized PageRankYiming Li, Yanyan Shen, Lei Chen, Mingxuan YuanVLDB 2023 · 67 citations
- Algorithm and System Co-design for Efficient Subgraph-based Graph Representation LearningHaoteng Yin, Muhan Zhang, Yanbang Wang, Jianguo Wang et al.VLDB 2022 · 47 citations
- RIM: Reliable Influence-based Active Learning on GraphsWentao Zhang, Yexin Wang, Zhenbang You, Meng Cao et al.NeurIPS 2021 · 43 citations
- Information Gain Propagation: a New Way to Graph Active Learning with Soft LabelsWentao Zhang, Yexin Wang, Zhenbang You, Meng Cao et al.ICLR 2022 · 24 citations
- No Change, No Gain: Empowering Graph Neural Networks with Expected Model Change Maximization for Active LearningZixing Song, Yifei Zhang, Irwin KingNeurIPS 2023 · 21 citations
Builds on10
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li et al.SIGIR 2020 · 4,448 citations
- Open Graph Benchmark: Datasets for Machine Learning on GraphsWeihua Hu, Matthias Fey, Marinka Zitnik, Yuxiao Dong et al.NeurIPS 2020 · 3,935 citations
- Contrastive Multi-View Representation Learning on GraphsKaveh Hassani, Amir Hosein Khas AhmadiICML 2020 · 1,663 citations
- Simple Spectral Graph ConvolutionHao Zhu, Piotr KoniuszICLR 2021 · 352 citations
- Adaptive Graph Encoder for Attributed Graph EmbeddingGanqu Cui, Jie Zhou, Cheng Yang, Zhiyuan LiuKDD 2020 · 224 citations
Related papers
- DeepSN: A Sheaf Neural Framework for Influence MaximizationAsela Hevapathige, Qing Wang, Ahad N. ZehmakanAAAI 2025 · 4 citations
- IMGNN: An Efficient, Effective and Generalizable Algorithm for Influence Maximization in Social NetworksHaotian Zhang, Kai Han, Zhizhuo Yin, Shuang Cui et al.KDD 2026
- NC-ALG: Graph-Based Active Learning Under Noisy CrowdWentao Zhang, Yexin Wang, Zhenbang You, Yang Li et al.ICDE 2024 · 4 citations
- GRAIN: Multi-Granular and Implicit Information Aggregation Graph Neural Network for Heterophilous GraphsSongwei Zhao, Yuan Jiang, Zijing Zhang, Yang Yu et al.AAAI 2025 · 5 citations
- Know Your Neighbors: Subgraph Importance Sampling for Heterophilic Graph Active LearningWenjie Yang, Shengzhong Zhang, Chen Ye, Jiaxing Guo et al.AAAI 2026
