Centaur: Hybrid Processing in On/Off-chip Memory Architecture for Graph Analytics
Abraham Addisie, Valeria Bertacco
Abstract
The increased use of graph algorithms in diverse fields has highlighted their inefficiencies in current chip-multiprocessor (CMP) architectures, primarily due to their seemingly random-access patterns to off-chip memory. Recently, two families of solutions have been proposed: 1) solutions that offload operations generated by all vertices from the processor cores to off-chip memory; and 2) solutions that offload only operations generated by high-degree vertices to dedicated on-chip memory, while the cores continue to process the work related to the remaining vertices. Neither approach is optimal over the full range of vertex's degrees. Thus, in this work, we propose Centaur, a novel architecture that processes operations on vertex data in on- and off-chip memory. Centaur utilizes a vertex's degree as a proxy to determine whether to process related operations in on- or off-chip memory. Centaur manages to provide up to 4.0× improvement in performance and 3.8× in energy benefits, compared to a baseline CMP, and up to a 2.0× performance boost over state-of-the-art specialized solutions.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Cited by top-tier papers1
Ask how each one uses itRelated papers
- NOVA: A Novel Vertex Management Architecture for Scalable Graph ProcessingMarjan Fariborz, Mahyar Samani, Austin York, S. J. Ben Yoo et al.HPCA 2025 · 2 citations
- CGgraph: An Ultra-fast Graph Processing System on Modern Commodity CPU-GPU Co-processorPengjie Cui, Haotian Liu, Bo Tang, Ye YuanVLDB 2024 · 18 citations
- CoroGraph: Bridging Cache Efficiency and Work Efficiency for Graph Algorithm ExecutionXiangyu Zhi, Xiao Yan, Bo Tang, Ziyao Yin et al.VLDB 2024 · 12 citations
- Hardware-Accelerated Hypergraph Processing with Chain-Driven SchedulingQinggang Wang, Long Zheng, Jingrui Yuan, Yu Huang et al.HPCA 2022 · 9 citations
- LCCG: a locality-centric hardware accelerator for high throughput of concurrent graph processingJin Zhao, Yu Zhang, Xiaofei Liao, Ligang He et al.SC 2021 · 8 citations
