Mako: a low-pause, high-throughput evacuating collector for memory-disaggregated datacenters
Haoran Ma, Shi Liu, Chenxi Wang, Yifan Qiao, Michael D. Bond, Stephen M. Blackburn, Miryung Kim, Guoqing Harry Xu
Abstract
Resource disaggregation has gained much traction as an emerging datacenter architecture, as it improves resource utilization and simplifies hardware adoption. Under resource disaggregation, different types of resources (memory, CPUs, etc.) are disaggregated into dedicated servers connected by high-speed network fabrics. Memory disaggregation brings efficiency challenges to concurrent garbage collection (GC), which is widely used for latency-sensitive cloud applications, because GC and mutator threads simultaneously run and constantly compete for memory and swap resources.
Mako is a new concurrent and distributed GC designed for memory-disaggregated environments. Key to Mako's success is its ability to offload both tracing and evacuation onto memory servers and run these tasks concurrently when the CPU server executes mutator threads. A major challenge is how to let servers efficiently synchronize as they do not share memory. We tackle this challenge with a set of novel techniques centered around the heap indirection table (HIT), where entries provide one-hop indirection for heap pointers. Our evaluation shows that Mako achieves ∼12ms at the 90th-percentile pause time and outperforms Shenandoah by an average of 3× in throughput.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 11d9cef1-4d75-4474-8741-c5a1eacf2183Cited by top-tier papers13
- Hermit: Low-Latency, High-Throughput, and Transparent Remote Memory via Feedback-Directed AsynchronyYifan Qiao, Chenxi Wang, Zhenyuan Ruan, Adam Belay et al.NSDI 2023 · 50 citations
- MemLiner: Lining up Tracing and Application for a Far-Memory-Friendly RuntimeChenxi Wang, Haoran Ma, Shi Liu, Yifan Qiao et al.OSDI 2022 · 47 citations
- A Tale of Two Paths: Toward a Hybrid Data Plane for Efficient Far-Memory ApplicationsLei Chen, Shi Liu, Chenxi Wang, Haoran Ma et al.OSDI 2024 · 16 citations
- CHIME: A Cache-Efficient and High-Performance Hybrid Index on Disaggregated MemoryXuchuan Luo, Jiacheng Shen, Pengfei Zuo, Xin Wang et al.SOSP 2024 · 10 citations
- DRust: Language-Guided Distributed Shared Memory with Fine Granularity, Full Transparency, and Ultra EfficiencyHaoran Ma, Yifan Qiao, Shi Liu, Shan Yu et al.OSDI 2024 · 8 citations
Builds on4
- AIFM: High-Performance, Application-Integrated Far MemoryZhenyuan Ruan, Malte Schwarzkopf, Marcos K. Aguilera, Adam BelayOSDI 2020 · 224 citations
- Can far memory improve job throughput?Emmanuel Amaro, Christopher Branner-Augmon, Zhihong Luo, Amy Ousterhout et al.EuroSys 2020 · 163 citations
- StRoM: smart remote memoryDavid Sidler, Zeke Wang, Monica Chiosa, Amit Kulkarni et al.EuroSys 2020 · 83 citations
- Semeru: A Memory-Disaggregated Managed RuntimeChenxi Wang, Haoran Ma, Shi Liu, Yuanqi Li et al.OSDI 2020 · 6 citations
Related papers
- Shaving the Peaks: Taming Tail Latency for Managed Workloads via Disaggregated Garbage CollectionHongtao Lyu, Yuhan Li, Mingyu WuOSDI 2026
- Cowbird: Freeing CPUs to Compute by Offloading the Disaggregation of MemoryXinyi Chen, Liangcheng Yu, Vincent Liu, Qizhen ZhangSIGCOMM 2023 · 15 citations
- Deft: A Scalable Tree Index for Disaggregated MemoryJing Wang, Qing Wang, Yuhao Zhang, Jiwu ShuEuroSys 2025 · 7 citations
- Scalable Distributed Inverted List Indexes in Disaggregated MemoryManuel Widmoser, Daniel Kocher, Nikolaus AugstenSIGMOD 2024 · 5 citations
- Rethinking software runtimes for disaggregated memoryIrina Calciu, M. Talha Imran, Ivan Puddu, Sanidhya Kashyap et al.ASPLOS 2021 · 116 citations
