Are dynamic memory managers on GPUs slow?: a survey and benchmarks
Martin Winter, Mathias Parger, Daniel Mlakar, Markus Steinberger
摘要
Dynamic memory management on GPUs is generally understood to be a challenging topic. On current GPUs, hundreds of thousands of threads might concurrently allocate new memory or free previously allocated memory. This leads to problems with thread contention, synchronization overhead and fragmentation. Various approaches have been proposed in the last ten years and we set out to evaluate them on a level playing field on modern hardware to answer the question, if dynamic memory managers are as slow as commonly thought of. In this survey paper, we provide a consistent framework to evaluate all publicly available memory managers in a large set of scenarios. We summarize each approach and thoroughly evaluate allocation performance (thread-based as well as warp-based), and look at performance scaling, fragmentation and real-world performance considering a synthetic workload as well as updating dynamic graphs. We discuss the strengths and weaknesses of each approach and provide guidelines for the respective best usage scenario. We provide a unified interface to integrate any of the tested memory managers into an application and switch between them for benchmarking purposes. Given our results, we can dispel some of the dread associated with dynamic memory managers on the GPU.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper8
- Efficient and Scalable Graph Pattern Mining on GPUsXuhao Chen, ArvindOSDI 2022 · 被引用 53 次
- GTS: GPU-based Tree Index for Fast Similarity SearchYifan Zhu, Ruiyao Ma, Baihua Zheng, Xiangyu Ke 等SIGMOD 2024 · 被引用 8 次
- Efficient Maximal Biclique Enumeration on GPUsZhe Pan, Shuibing He, Xu Li, Xuechen Zhang 等SC 2023 · 被引用 8 次
- SIMR: Single Instruction Multiple Request Processing for Energy-Efficient Data Center MicroservicesMahmoud Khairy, Ahmad Alawneh, Aaron Barnes, Timothy G. RogersMICRO 2022 · 被引用 8 次
- Occamy: Memory-efficient GPU Compiler for DNN InferenceJaeho Lee, Shinnung Jeong, Seungbin Song, Kunwoo Kim 等DAC 2023 · 被引用 3 次
相关 Paper
- Gallatin: A General-Purpose GPU Memory ManagerHunter McCoy, Prashant PandeyPPoPP 2024 · 被引用 7 次
- Towards Sufficient GPU-accelerated Dynamic Graph Management: Survey and ExperimentYinnian Lin, Lei Zou, Xunbin SuVLDB 2025 · 被引用 2 次
- MoonBright: A GPU Memory Allocator with Device-Side Page Table Materialization and Deferred TLB CoherenceYangyu Zhang, Lei Chen, Chunwei Xia, Shuaijiang Li 等OSDI 2026
- Memory Harvesting in Multi-GPU Systems with Hierarchical Unified Virtual MemorySangjin Choi, Taeksoo Kim, Jinwoo Jeong, Rachata Ausavarungnirun 等USENIX ATC 2022 · 被引用 28 次
- SuperCollider: Scalable and Effective Data Race Detection for CUDAMark Stephenson, Sana Damani, Mohamed Tarek Ibn Ziad, Anis Ladram 等PLDI 2026
