Lune

HPCA2026顶会

ARIADNE: Adaptive UVM Management for Efficient GPU Memory Oversubscription

Hyunkyun Shin, Seongtae Bang, Hyungwon Park, Daehoon Kim

2026年份
2被引次数
1顶会引用

摘要

Unified Virtual Memory (UVM) simplifies GPU programming and supports memory oversubscription, but suffers from severe performance degradation under high memory pressure due to page fault overhead and thrashing. Existing approaches such as prefetching, access counter-based migration, and dynamic Zero-copy offer limited benefits and often require hardware or compiler modifications, undermining UVM's portability and ease of deployment. We present ARIADNE, a runtime UVM management framework that preserves UVM's GPU memory abstraction while ensuring high and robust performance under memory oversubscription. ARIADNE is guided by three principles: (1) pipelined fault handling to hide migration latency, (2) Sharing Degree, a runtime metric that captures thread-level access locality without requiring hardware or compiler changes, to inform placement decisions, and (3) dynamic placement of memory regions between GPU memory and Zero-copy based on real-time access patterns. Implemented entirely within NVIDIA's UVM driver, ARIADNE requires no recompilation or hardware modifications and applies transparently to any executable or closed-source GPU UVM applications. Our experimental results show that ARIADNE delivers average speedups of1.9×,5.0×1.9 \times, 5.0 \times, and4.8×4.8 \timesover a state-of-the-art method at1 3 0 %,  1 7 5 %\text{1 3 0 \%, ~} \text{1 7 5 \%}, and 300 % oversubscription, respectively, while effectively preventing thrashing and maintaining near-linear performance scaling.

问问这篇 Paper

问问你的智能体。

Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。

可以从这些问题问起

智能体调用

Lunesearch_papers

在 Lune 里问

免费开始,无需绑卡

lune papers get b555c604-c902-4fa4-9eb4-9555fe76ca79

引用它的顶会 Paper1

问问它们各自怎么用它

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖