Bonsai: Gradient-free Graph Condensation for Node Classification
Mridul Gupta, Samyak Jain, Vansh Ramani, Hariprasad Kodamana, Sayan Ranu
摘要
Graph condensation has emerged as a promising avenue to enable scalable training of Gnns by compressing the training dataset while preserving essential graph characteristics. Our study uncovers significant shortcomings in current graph condensation techniques. First, the majority of the algorithms paradoxically require training on the full dataset to perform condensation. Second, due to their gradient-emulating approach, these methods require fresh condensation for any change in hyper-parameters or Gnn architecture, limiting their flexibility and reusability. To address these challenges, we present Bonsai, a novel graph condensation method empowered by the observation that computation trees form the fundamental processing units of message-passing Gnns. Bonsai condenses datasets by encoding a careful selection of exemplar trees that maximize the representation of all computation trees in the training set. This unique approach imparts Bonsai as the first linear-time, model-agnostic graph condensation algorithm for node classification that outperforms existing baselines across 7 real-world datasets on accuracy, while being 22 times faster on average. Bonsai is grounded in rigorous mathematical guarantees on the adopted approximation strategies, making it robust to Gnn architectures, datasets, and parameters. *Denotes equal contribution. ¹Some algorithms sparsify the fully-connected graph based on edge weights. But this sparsification process requires training on the fully connected graph itself to identify the pruning threshold. ²Inspired by the art of Bonsai, which transforms large trees into miniature forms while preserving their essence, our graph condensation algorithm gracefully prunes redundant computation trees, creating a condensed graph that is significantly smaller yet maintains comparable performance.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper15
- GraphSAINT: Graph Sampling Based Inductive Learning MethodHanqing Zeng, Hongkuan Zhou, Ajitesh Srivastava, Rajgopal Kannan 等ICLR 2020 · 被引用 1,155 次
- Graph Condensation for Graph Neural NetworksWei Jin, Lingxiao Zhao, Shichang Zhang, Yozen Liu 等ICLR 2022 · 被引用 203 次
- Condensing Graphs via One-Step Gradient MatchingWei Jin, Xianfeng Tang, Haoming Jiang, Zheng Li 等KDD 2022 · 被引用 68 次
- Does Graph Distillation See Like Vision Dataset Counterpart?Beining Yang, Kai Wang, Qingyun Sun, Cheng Ji 等NeurIPS 2023 · 被引用 62 次
- Fast Graph Condensation with Structure-based Neural Tangent KernelLin Wang, Wenqi Fan, Jiatong Li, Yao Ma 等WWW 2024 · 被引用 45 次
相关 Paper
- Mirage: Model-agnostic Graph Distillation for Graph ClassificationMridul Gupta, Sahil Manchanda, Hariprasad Kodamana, Sayan RanuICLR 2024 · 被引用 17 次
- Bi-Directional Multi-Scale Graph Dataset Condensation via Information BottleneckXingcheng Fu, Yisen Gao, Beining Yang, Yuxuan Wu 等AAAI 2025 · 被引用 7 次
- Disentangled Condensation for Large-scale GraphsZhenbang Xiao, Yu Wang, Shunyu Liu, Bingde Hu 等WWW 2025 · 被引用 14 次
- Graph Condensation for Inductive Node Representation LearningXinyi Gao, Tong Chen, Yilong Zang, Wentao Zhang 等ICDE 2024 · 被引用 32 次
- Rethinking and Accelerating Graph Condensation: A Training-Free Approach with Class PartitionXinyi Gao, Guanhua Ye, Tong Chen, Wentao Zhang 等WWW 2025 · 被引用 27 次
