Scalable Far Memory: Balancing Faults and Evictions
Yueyang Pan, Yash Lala, Musa Unal, Yujie Ren, SeungSeob Lee, Abhishek Bhattacharjee, Anurag Khandelwal, Sanidhya Kashyap
摘要
Page-based far memory systems transparently expand an application's memory capacity beyond a single machine without modifying application code. However, existing systems are tailored to scenarios with low application thread counts, and fail to scale on today's multi-core machines. This makes them unsuitable for data-intensive applications that both rely on far memory support and scale with increasing thread count. Our analysis reveals that this poor scalability stems from inefficient holistic coordination between page fault-in and eviction operations. As thread count increases, current systems encounter scalability bottlenecks in TLB shootdowns, page accounting, and memory allocation.
This paper presents three design principles that address these scalability challenges and enable efficient memory offloading. These principles are always-asynchronous decoupling to handle eviction operations as asynchronously as possible, cross-batch pipelined execution to avoid idle waiting periods, and scalability prioritization to avoid synchronization overheads at high thread counts at the cost of eviction accuracy. We implement these principles in both the Linux kernel and a library OS. Our evaluation shows that this approach increases throughput for batch-processing applications by up to 4.2× and reduces 99 th percentile latency for a latency-critical memcached application by 94.5%.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Efficient and Scalable Synchronization via Generalized Cache CoherenceYanpeng Yu, Seung-seob Lee, Lin Zhong, Anurag KhandelwalOSDI 2026
- OneSidedMW: Managing Disaggregated Memory Efficiently, Flexibly, and Securely with RNIC OffloadingZixuan Wang, Jinyu Gu, Xingda Wei, Yubin XiaNSDI 2026
它引用的顶会 Paper21
- Efficient Memory Management for Large Language Model Serving with PagedAttentionWoosuk Kwon, Zhuohan Li, Siyuan Zhuang, Ying Sheng 等SOSP 2023 · 被引用 1,016 次
- AIFM: High-Performance, Application-Integrated Far MemoryZhenyuan Ruan, Malte Schwarzkopf, Marcos K. Aguilera, Adam BelayOSDI 2020 · 被引用 224 次
- Effectively Prefetching Remote Memory with LeapHasan Al Maruf, Mosharaf ChowdhuryUSENIX ATC 2020 · 被引用 186 次
- RDMA over Ethernet for Distributed Training at Meta ScaleAdithya Gangidi, Rui Miao, Shengbao Zheng, Sai Jayesh Bondu 等SIGCOMM 2024 · 被引用 171 次
- Can far memory improve job throughput?Emmanuel Amaro, Christopher Branner-Augmon, Zhihong Luo, Amy Ousterhout 等EuroSys 2020 · 被引用 163 次
相关 Paper
- Batch-Aware Unified Memory Management in GPUs for Irregular WorkloadsHyojong Kim, Jaewoong Sim, Prasun Gera, Ramyad Hadidi 等ASPLOS 2020 · 被引用 89 次
- RaidenSwap: A Multi-Swap Remote System for Multi-core ApplicationsKefan Liu, Ke Liu, Xu Zhang, Hui Yuan 等EuroSys 2026 · 被引用 1 次
- Meerkat: multicore-scalable replicated transactions following the zero-coordination principleAdriana Szekeres, Michael J. Whittaker, Jialin Li, Naveen Kr. Sharma 等EuroSys 2020 · 被引用 11 次
- ScaleCache: A Scalable Page Cache for Multiple Solid-State DrivesKiet Tuan Pham, Seokjoo Cho, Sangjin Lee, Lan Anh Nguyen 等EuroSys 2024 · 被引用 10 次
- PageFlex: Flexible and Efficient User-space Delegation of Linux Paging Policies with eBPFAnil Yelam, Kan Wu, Zhiyuan Guo, Suli Yang 等USENIX ATC 2025 · 被引用 15 次
