GPHash: An Efficient Hash Index for GPU with Byte-Granularity Persistent Memory
Menglei Chen, Yu Hua, Zhangyu Chen, Ming Zhang, Gen Dong
摘要
GPU with persistent memory (GPM) enables GPU-powered applications to directly manage the data in persistent memory at the byte granularity. Hash indexes have been widely used to achieve efficient data management. However, conventional hash indexes become inefficient for GPM systems due to warp-agnostic execution manner, high-overhead consistency guarantee, and significant bandwidth gap between PM and GPU. In this paper, we propose GPHash, an efficient hash index for GPM systems with high performance and consistency guarantee. To fully exploit the parallelism of GPU, GPHash executes all index operations in a lock-free and warpcooperative manner. Moreover, by using CAS primitive and slot states, GPHash ensures consistency guarantee with low overhead. To further bridge the bandwidth gap between PM and GPU, GPHash caches hot items in GPU memory while minimizing the overhead for cache management. Extensive evaluations on YCSB and real-world workloads show that GPHash outperforms state-of-the-art CPU-assisted data management approaches and GPM hash indexes by up to 27.62×. Slot 0 Slot 1 Slot 2 Slot 3 (b) Fixed-length large key (> 8 Bytes) 8-byte FP & State Value pointer Key (a) Fixed-length small key (<= 8 Bytes) 8-byte Key & State Value pointer (c) Variable-length keys and values 16-bit FP 48-bit Key-value pair pointer Inter-level shared buckets 63 (a) GPU-conscious and PM-friendly hash table Pointer-based key placement Bucket FPs & States
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper31
- An Empirical Guide to the Behavior and Use of Scalable Persistent MemoryJian Yang, Juno Kim, Morteza Hoseinzadeh, Joseph Izraelevitz 等FAST 2020 · 被引用 470 次
- NeRF in the Dark: High Dynamic Range View Synthesis from Noisy Raw ImagesBen Mildenhall, Peter Hedman, Ricardo Martin-Brualla, Pratul P. Srinivasan 等CVPR 2022 · 被引用 307 次
- Cost-Efficient Large Language Model Serving for Multi-turn Conversations with CachedAttentionBin Gao, Zhuomin He, Puru Sharma, Qingxuan Kang 等USENIX ATC 2024 · 被引用 273 次
- A large scale analysis of hundreds of in-memory cache clusters at TwitterJuncheng Yang, Yao Yue, K. V. RashmiOSDI 2020 · 被引用 245 次
- FlatStore: An Efficient Log-Structured Key-Value Storage Engine for Persistent MemoryYoumin Chen, Youyou Lu, Fan Yang, Qing Wang 等ASPLOS 2020 · 被引用 166 次
相关 Paper
- Exploiting Persistent CPU Cache for Scalable Persistent Hash IndexBowen Zhang, Shengan Zheng, Liangxu Nie, Zhenlin Qi 等ICDE 2024 · 被引用 3 次
- GPH: An Efficient and Effective Perfect Hashing Scheme for GPU ArchitecturesJiaping Cao, Le Xu, Man Lung Yiu, Jianbin Qin 等SIGMOD 2025 · 被引用 4 次
- Lock-free Concurrent Level Hashing for Persistent MemoryZhangyu Chen, Yu Hua, Bo Ding, Pengfei ZuoUSENIX ATC 2020 · 被引用 98 次
- SEPH: Scalable, Efficient, and Predictable Hashing on Persistent MemoryChao Wang, Junliang Hu, Tsun-Yu Yang, Yuhong Liang 等OSDI 2023
- Persistent Memory Hash Indexes: An Experimental EvaluationDaokun Hu, Zhiwen Chen, Jianbing Wu, Jianhua Sun 等VLDB 2021 · 被引用 47 次
