FORGE: Mitigating Synchronization Amplification for Memory-Disaggregated Caching Systems
Zhijun Yang, Yu Hua, Ming Zhang, Menglei Chen, Yixiao Wang
摘要
Disaggregated Memory (DM) architectures offer caching systems the potential for elastic scaling and improved resource utilization by decoupling compute and memory. However, this advantage is undermined by costly cross-node synchronization, which exacerbates the overheads of critical cache operations, including hotness tracking, eviction coordination, and memory defragmentation. To address this challenge, we present FORGE, a caching system tailored for DM that prioritizes synchronization efficiency. FORGE groups cached objects based on similarity and performs group-level synchronizations to amortize overheads. It evicts cold groups via a contention-free and hotness-aware FIFO queue, efficiently sustaining high hit ratios while mitigating memory fragmentation. Leveraging the predictability of FIFO evictions, FORGE adopts a lazy synchronization strategy that updates hotness metrics just-in-time for eviction and offloads this process to on-chip memory in RDMA NICs for acceleration. Extensive evaluations on YCSB and real-world workloads demonstrate that FORGE achieves up to 4.5× higher throughput, 4.0×/7.5× lower P50/P99 latency, and an average of 1.14× higher cache hit ratio compared with state-of-the-art systems.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper36
- Serverless in the Wild: Characterizing and Optimizing the Serverless Workload at a Large Cloud ProviderMohammad Shahrad, Rodrigo Fonseca, Iñigo Goiri, Gohar Irfan Chaudhry 等USENIX ATC 2020 · 被引用 946 次
- Mooncake: Trading More Storage for Less Computation - A KVCache-centric Architecture for Serving LLM ChatbotRuoyu Qin, Zheming Li, Weiran He, Jialei Cui 等FAST 2025 · 被引用 337 次
- Pond: CXL-Based Memory Pooling Systems for Cloud PlatformsHuaicheng Li, Daniel S. Berger, Lisa Hsu, Daniel Ernst 等ASPLOS 2023 · 被引用 328 次
- TPP: Transparent Page Placement for CXL-Enabled Tiered-MemoryHasan Al Maruf, Hao Wang, Abhishek Dhanotia, Johannes Weiner 等ASPLOS 2023 · 被引用 255 次
- A large scale analysis of hundreds of in-memory cache clusters at TwitterJuncheng Yang, Yao Yue, K. V. RashmiOSDI 2020 · 被引用 245 次
相关 Paper
- Shard: A Scalable and Resize-optimized Hash Index on Disaggregated MemoryHantian Zha, Teng Ma, Baotong Lu, Yuansen Wang 等VLDB 2026
- DMTree: Towards Efficient Tree Indexing on Disaggregated Memory via Compute-side Collaborative DesignGuoli Wei, Yongkun Li, Haoze Song, Tao Li 等FAST 2026 · 被引用 1 次
- Fast Distributed Transactions for RDMA-based Disaggregated MemoryHaodi Lu, Haikun Liu, Yujian Zhang, Zhuohui Duan 等USENIX ATC 2025 · 被引用 9 次
- CoRM: Compactable Remote Memory over RDMAKonstantin Taranov, Salvatore Di Girolamo, Torsten HoeflerSIGMOD 2021 · 被引用 18 次
- UniMem: Redesigning Disaggregated Memory within A Unified Local-Remote Memory HierarchyYijie Zhong, Minqiang Zhou, Zhirong Shen, Jiwu ShuUSENIX ATC 2024 · 被引用 7 次
