PaCaR: Improved Buffered I/O Locality on NUMA Systems with Page Cache Replication
Jérôme Coquisart, Julien Sopena, Redha Gouicem
摘要
Modern systems featuring multiple sockets suffer from the non-uniformity of performance of their memory accesses. Accessing memory of a remote node can lead to doubled latencies and halved bandwidth. This is particularly critical for I/O-intensive applications relying on the Linux page cache, where locality is paramount. Existing solutions migrating threads or memory across nodes fail to address the issue with highly parallel workloads, where a small working set is constantly accessed by multiple threads across nodes.
We present PaCaR, a novel page cache replication mechanism that enhances locality by transparently replicating cached pages across Non-Uniform Memory Access (NUMA) nodes. PaCaR ensures write consistency, adapts to memory pressure, and operates seamlessly without requiring application modifications. Evaluations on a dual socket NUMA server demonstrate up to 1.4x performance improvements in synthetic workloads and up to 25% gains in real-world workloads, showcasing the effectiveness of PaCaR in optimizing NUMA-aware buffered I/O operations.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper7
- Building An Elastic Query Engine on Disaggregated StorageMidhul Vuppalapati, Justin Miron, Rachit Agarwal, Dan Truong 等NSDI 2020 · 被引用 142 次
- TMO: transparent memory offloading in datacentersJohannes Weiner, Niket Agarwal, Dan Schatzberg, Leon Yang 等ASPLOS 2022 · 被引用 103 次
- Mitosis: Transparently Self-Replicating Page-Tables for Large-Memory MachinesReto Achermann, Ashish Panwar, Abhishek Bhattacharjee, Timothy Roscoe 等ASPLOS 2020 · 被引用 62 次
- Efficient Memory Overcommitment for I/O Passthrough Enabled VMs via Fine-grained Page Meta-data ManagementYaohui Wang, Ben Luo, Yibin ShenUSENIX ATC 2023 · 被引用 18 次
- Cloud-scale VM-deflation for Running Interactive Applications On Transient ServersAlexander Fuerst, Ahmed Ali-Eldin, Prashant J. Shenoy, Prateek SharmaHPDC 2020 · 被引用 12 次
相关 Paper
- Scalable and Effective Page-table and TLB management on NUMA SystemsBin Gao, Qingxuan Kang, Hao-Wei Tee, Kyle Timothy Ng Chu 等USENIX ATC 2024 · 被引用 5 次
- StarNUMA: Mitigating NUMA Challenges with Memory PoolingAlbert Cho, Alexandros DaglisMICRO 2024 · 被引用 11 次
- Fast local page-tables for virtualized NUMA servers with vMitosisAshish Panwar, Reto Achermann, Arkaprava Basu, Abhishek Bhattacharjee 等ASPLOS 2021 · 被引用 29 次
- TAPMM: A Traffic-Aware Page Mapping Method for Multi-level NUMA SystemsFengkun Dong, Guoqing Xiao, Haotian Wang, Yikun Hu 等DAC 2024
- NUBA: Non-Uniform Bandwidth GPUsXia Zhao, Magnus Jahre, Yuhua Tang, Guangda Zhang 等ASPLOS 2023 · 被引用 17 次
