MetaNMP: Leveraging Cartesian-Like Product to Accelerate HGNNs with Near-Memory Processing
Dan Chen, Haiheng He, Hai Jin, Long Zheng, Yu Huang, Xinyang Shen, Xiaofei Liao
摘要
Heterogeneous graph neural networks (HGNNs) based on metapath exhibit powerful capturing of rich structural and semantic information in the heterogeneous graph. HGNNs are highly memory-bound and thus can be accelerated by near-memory processing. However, they also suffer from significant memory footprint (due to storing metapath instances as intermediate data) and severe redundant computation (when vertex features are aggregated among metapath instances). To address these issues, this paper proposes MetaNMP, the first DIMM-based near-memory processing HGNNs accelerator with reduced memory footprint and high performance. Specifically, we first propose a cartesian-like product paradigm to generate all metapath instances on the fly for heterogeneous graphs. In this way, metapath instances no longer need to be stored as intermediate data, avoiding significant memory consumption. We then design a data flow for aggregating vertex features on metapath instances, which aggregates vertex features along the direction of the metapath instances dispersed from the starting vertex to exploit shareable aggregation computations, eliminating most of the redundant computations. Finally, we integrate specialized hardware units in DIMM to accelerate HGNNs with near-memory processing, and introduce a broadcast mechanism for edge data and vertex features to mitigate the inter-DIMM communication. Our evaluation shows that MetaNMP achieves the memory space reduction of 51.9% on average and the performance improvement by 415.18× compared to NVIDIA Tesla V100 GPU.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper6
- Tender: Accelerating Large Language Models via Tensor Decomposition and Runtime RequantizationJungi Lee, Wonbeom Lee, Jaewoong SimISCA 2024 · 被引用 41 次
- ANSMET: Approximate Nearest Neighbor Search with Near-Memory Processing and Hybrid Early TerminationYiwei Li, Yuxin Jin, Boyu Tian, Huanchen Zhang 等ISCA 2025 · 被引用 10 次
- GDR-HGNN: A Heterogeneous Graph Neural Networks Accelerator Frontend with Graph Decoupling and RecouplingRunzhen Xue, Mingyu Yan, Dengke Han, Yihan Teng 等DAC 2024 · 被引用 6 次
- Stream-Based Data Placement for Near-Data Processing with Extended MemoryYiwei Li, Boyu Tian, Yi Ren, Mingyu GaoMICRO 2024 · 被引用 5 次
- Ironman: Accelerating Oblivious Transfer Extension for Privacy-Preserving AI with Near-Memory ProcessingChenqi Lin, Kang Yang, Tianshi Xu, Ling Liang 等MICRO 2025 · 被引用 4 次
相关 Paper
- MetaHG: Enhancing HGNN Systems Leveraging Advanced Metapath Graph AbstractionHaiheng He, Haifeng Liu, Long Zheng, Yu Huang 等EuroSys 2025
- ABC-DIMM: Alleviating the Bottleneck of Communication in DIMM-based Near-Memory Processing with Inter-DIMM BroadcastWeiyi Sun, Zhaoshi Li, Shouyi Yin, Shaojun Wei 等ISCA 2021 · 被引用 42 次
- Simple and Efficient Heterogeneous Graph Neural NetworkXiaocheng Yang, Mingyu Yan, Shirui Pan, Xiaochun Ye 等AAAI 2023 · 被引用 233 次
- MAGNN: Metapath Aggregated Graph Neural Network for Heterogeneous Graph EmbeddingXinyu Fu, Jiani Zhang, Ziqiao Meng, Irwin KingWWW 2020 · 被引用 1,149 次
- Heta: Distributed Training of Heterogeneous Graph Neural NetworksYuchen Zhong, Junwei Su, Chuan Wu, Minjie WangVLDB 2025 · 被引用 2 次
