Efficient GPU Multitasking with Morphable Kernels
Tingxu Ren, Ruwen Fan, Hao Guo, Minhui Xie, Shiwei Gao, Jiwu Shu, Youyou Lu
摘要
GPU multitasking offers a promising approach to improving hardware utilization by co-locating concurrent workloads on the same device. However, achieving high resource utilization with minimized interference requires fine-grained, adaptive scheduling. Existing scheduling solutions are fundamentally constrained by a rigid assumption: once a kernel is launched, its resource footprint remains fixed throughout its execution. Consequently, they either rely on static resource pre-partitioning, which lacks the flexibility to adapt to rapid workload changes, or adopt kernel-slicing techniques, which achieve fine-grained control at the cost of prohibitive kernel launch overheads.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- Navigator: Dynamic Multi-kernel Scheduling to Improve GPU PerformanceJiho Kim, John Kim, Yongjun ParkDAC 2020 · 被引用 9 次
- BlockMaestro: Enabling Programmer-Transparent Task-based Execution in GPU SystemsAmirAli Abdolrashidi, Hodjat Asghari Esfeden, Ali Jahanshahi, Kaustubh Singh 等ISCA 2021 · 被引用 15 次
- µShare: Non-Intrusive Kernel Co-Locating on NVIDIA GPUsWenhao Huang, Zhaolin Duan, Laiping Zhao, Yuhao Zhang 等HPCA 2026
- Interference-aware Multiplexing for Deep Learning in GPU Clusters: A Middleware ApproachWenyan Chen, Zizhao Mo, Huanle Xu, Kejiang Ye 等SC 2023 · 被引用 21 次
- Tacker: Tensor-CUDA Core Kernel Fusion for Improving the GPU Utilization while Ensuring QoSHan Zhao, Weihao Cui, Quan Chen, Youtao Zhang 等HPCA 2022 · 被引用 42 次
