CoroGraph: Bridging Cache Efficiency and Work Efficiency for Graph Algorithm Execution
Xiangyu Zhi, Xiao Yan, Bo Tang, Ziyao Yin, Yanchao Zhu, Minqi Zhou
Abstract
Many systems are designed to run graph algorithms efficiently in memory but they achieve only cache efficiency or work efficiency. We tackle this fundamental trade-off in existing systems by designing CoroGraph, a system that attains both cache efficiency and work efficiency for in-memory graph processing. CoroGraph adopts a novel hybrid execution model , which generates update messages at vertex granularity to prioritize promising vertices for work efficiency, and commits updates at partition granularity to share data access for cache efficiency. To overlap the random memory access of graph algorithms with computation, CoroGraph extensively uses coroutine , i.e., a lightweight function in C++ that can yield and resume with low overhead, to prefetch the required data. A suite of designs are incorporated to reap the full benefits of coroutine, which include prefetch pipeline, cache-friendly graph format, and stop-free synchronization. We compare CoroGraph with five state-of-the-art graph algorithm systems via extensive experiments. The results show that CoroGraph yields shorter algorithm execution time than all baselines in 18 out of 20 cases, and its speedup over the best-performing baseline can be over 2x. Detailed profiling suggests that CoroGraph achieves both cache efficiency and work efficiency with a low memory stall and a small number of processed edges.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers7
- GastCoCo: Graph Storage and Coroutine-Based Prefetch Co-Design for Dynamic Graph ProcessingHongfu Li, Qian Tao, Song Yu, Shufeng Gong et al.VLDB 2024 · 9 citations
- Enabling Window-Based Monotonic Graph Analytics with Reusable Transitional Results for Pattern-Consistent QueriesZheng Chen, Feng Zhang, Yang Chen, Xiaokun Fang et al.VLDB 2024 · 6 citations
- ACGraph: An Efficient Asynchronous Out-of-Core Graph Processing FrameworkDechuang Chen, Sibo Wang, Qintian GuoSIGMOD 2026 · 3 citations
- Efficient Cooperation-Aware Key and Value Management for LLM InferenceQiheng Sun, Hongwei Zhang, Junxu Liu, Haocheng Xia et al.VLDB 2026
- Rule-Based Graph Cleaning with GPUs on a Single MachineWenchao Bai, Wenfei Fan, Shuhao Liu, Kehan Pang et al.SIGMOD 2025
Builds on6
- Classifying Memory Access Patterns for PrefetchingGrant Ayers, Heiner Litz, Christos Kozyrakis, Parthasarathy RanganathanASPLOS 2020 · 83 citations
- ThunderRW: An In-Memory Graph Random Walk EngineShixuan Sun, Yuhang Chen, Shengliang Lu, Bingsheng He et al.VLDB 2021 · 31 citations
- MxTasks: How to Make Efficient Synchronization and Prefetching EasyJan Mühlig, Jens TeubnerSIGMOD 2021 · 11 citations
- Cache-Efficient Fork-Processing Patterns on Large GraphsShengliang Lu, Shixuan Sun, Johns Paul, Yuchen Li et al.SIGMOD 2021 · 10 citations
- Blaze: Fast Graph Processing on Fast SSDsJuno Kim, Steven SwansonSC 2022 · 8 citations
Related papers
- MiniGraph: Querying Big Graphs with a Single MachineXiaoke Zhu, Yang Liu, Shuhao Liu, Wenfei FanVLDB 2023 · 12 citations
- An Efficient Memoization Engine for Concurrent Graph Query ProcessingSen Gao, Shengliang Lu, Shixuan Sun, Yuchen Li et al.ICDE 2025 · 1 citation
- Approaching DRAM performance by using microsecond-latency flash memory for small-sized random read accesses: a new access method and its graph applicationsTomoya Suzuki, Kazuhiro Hiwada, Hirotsugu Kajihara, Shintaro Sano et al.VLDB 2021 · 9 citations
- Centaur: Hybrid Processing in On/Off-chip Memory Architecture for Graph AnalyticsAbraham Addisie, Valeria BertaccoDAC 2020 · 5 citations
- CGgraph: An Ultra-fast Graph Processing System on Modern Commodity CPU-GPU Co-processorPengjie Cui, Haotian Liu, Bo Tang, Ye YuanVLDB 2024 · 18 citations
