Symmetry-Preserving Architecture for Multi-NUMA Environments (SPANE): A Deep Reinforcement Learning Approach for Dynamic VM Scheduling
Chan Tin Ping, Yunlong Cheng, Yizhan Zhu, Xiaofeng Gao, Guihai Chen
Abstract
As cloud computing continues to evolve, the adoption of multi-NUMA (Non-Uniform Memory Access) architecture by cloud service providers has introduced new challenges in virtual machine (VM) scheduling. To address these challenges and more accurately reflect the complexities faced by modern cloud environments, we introduce the Dynamic VM Allocation problem in Multi-NUMA PM (DVAMP). We formally define both offline and online versions of DVAMP as mixed-integer lin-ear programming problems, providing a rigorous mathematical foundation for analysis. A tight performance bound for greedy online algorithms is derived, offering insights into the worst-case optimality gap as a function of the number of physical machines and VM lifetime variability. To address the challenges posed by DVAMP, we propose SPANE (Symmetry-Preserving Architecture for Multi-NUMA Environments), a novel deep rein-forcement learning approach that exploits the problem's inherent symmetries. SPANE produces invariant results under arbitrary permutations of physical machine states, enhancing learning effi-ciency and solution quality. Extensive experiments conducted on the Huawei-East-1 dataset demonstrate that SPANE outperforms existing baselines, reducing average VM wait time by 45%. Our work contributes to the field of cloud resource management by providing both theoretical insights and practical solutions for VM scheduling in multi-NUMA environments, addressing a critical gap in the literature and offering improved performance for real-world cloud systems.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5e865f55-4fdd-474e-9761-27f3abf2addfBuilds on2
- MDP Homomorphic Networks: Group Symmetries in Reinforcement LearningElise van der Pol, Daniel E. Worrall, Herke van Hoof, Frans A. Oliehoek et al.NeurIPS 2020 · 203 citations
- An Approximation for Job Scheduling on Cloud with Synchronization and Slowdown ConstraintsDejun Kong, Zhongrui Zhang, Yangguang Shi, Xiaofeng GaoINFOCOM 2023 · 2 citations
Related papers
- A Deep Reinforcement Learning based Online Scheduling Policy for Deep Neural Network Multi-Tenant Multi-Accelerator SystemsFrancesco Giulio Blanco, Enrico Russo, Maurizio Palesi, Davide Patti et al.DAC 2024 · 9 citations
- Towards VM Rescheduling Optimization Through Deep Reinforcement LearningXianzhong Ding, Yunkai Zhang, Binbin Chen, Donghao Ying et al.EuroSys 2025 · 10 citations
- Scheduling of Time-Varying Workloads Using Reinforcement LearningShanka Subhra Mondal, Nikhil Sheoran, Subrata MitraAAAI 2021 · 45 citations
- Learning a Partitioning Advisor for Cloud DatabasesBenjamin Hilprecht, Carsten Binnig, Uwe RöhmSIGMOD 2020 · 64 citations
- Online Placement of Virtual Machines with Prior DataDavid Naori, Danny RazINFOCOM 2020 · 4 citations
