Scalable Far Memory: Balancing Faults and Evictions
Yueyang Pan, Yash Lala, Musa Unal, Yujie Ren, SeungSeob Lee, Abhishek Bhattacharjee, Anurag Khandelwal, Sanidhya Kashyap
Abstract
Page-based far memory systems transparently expand an application's memory capacity beyond a single machine without modifying application code. However, existing systems are tailored to scenarios with low application thread counts, and fail to scale on today's multi-core machines. This makes them unsuitable for data-intensive applications that both rely on far memory support and scale with increasing thread count. Our analysis reveals that this poor scalability stems from inefficient holistic coordination between page fault-in and eviction operations. As thread count increases, current systems encounter scalability bottlenecks in TLB shootdowns, page accounting, and memory allocation.
This paper presents three design principles that address these scalability challenges and enable efficient memory offloading. These principles are always-asynchronous decoupling to handle eviction operations as asynchronously as possible, cross-batch pipelined execution to avoid idle waiting periods, and scalability prioritization to avoid synchronization overheads at high thread counts at the cost of eviction accuracy. We implement these principles in both the Linux kernel and a library OS. Our evaluation shows that this approach increases throughput for batch-processing applications by up to 4.2× and reduces 99 th percentile latency for a latency-critical memcached application by 94.5%.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 33643d75-d936-4234-aed6-b8e68c9ff1d7Cited by top-tier papers2
- Efficient and Scalable Synchronization via Generalized Cache CoherenceYanpeng Yu, Seung-seob Lee, Lin Zhong, Anurag KhandelwalOSDI 2026
- OneSidedMW: Managing Disaggregated Memory Efficiently, Flexibly, and Securely with RNIC OffloadingZixuan Wang, Jinyu Gu, Xingda Wei, Yubin XiaNSDI 2026
Builds on21
- Efficient Memory Management for Large Language Model Serving with PagedAttentionWoosuk Kwon, Zhuohan Li, Siyuan Zhuang, Ying Sheng et al.SOSP 2023 · 1,016 citations
- AIFM: High-Performance, Application-Integrated Far MemoryZhenyuan Ruan, Malte Schwarzkopf, Marcos K. Aguilera, Adam BelayOSDI 2020 · 224 citations
- Effectively Prefetching Remote Memory with LeapHasan Al Maruf, Mosharaf ChowdhuryUSENIX ATC 2020 · 186 citations
- RDMA over Ethernet for Distributed Training at Meta ScaleAdithya Gangidi, Rui Miao, Shengbao Zheng, Sai Jayesh Bondu et al.SIGCOMM 2024 · 171 citations
- Can far memory improve job throughput?Emmanuel Amaro, Christopher Branner-Augmon, Zhihong Luo, Amy Ousterhout et al.EuroSys 2020 · 163 citations
Related papers
- Batch-Aware Unified Memory Management in GPUs for Irregular WorkloadsHyojong Kim, Jaewoong Sim, Prasun Gera, Ramyad Hadidi et al.ASPLOS 2020 · 89 citations
- RaidenSwap: A Multi-Swap Remote System for Multi-core ApplicationsKefan Liu, Ke Liu, Xu Zhang, Hui Yuan et al.EuroSys 2026 · 1 citation
- Meerkat: multicore-scalable replicated transactions following the zero-coordination principleAdriana Szekeres, Michael J. Whittaker, Jialin Li, Naveen Kr. Sharma et al.EuroSys 2020 · 11 citations
- ScaleCache: A Scalable Page Cache for Multiple Solid-State DrivesKiet Tuan Pham, Seokjoo Cho, Sangjin Lee, Lan Anh Nguyen et al.EuroSys 2024 · 10 citations
- PageFlex: Flexible and Efficient User-space Delegation of Linux Paging Policies with eBPFAnil Yelam, Kan Wu, Zhiyuan Guo, Suli Yang et al.USENIX ATC 2025 · 15 citations
