Deft Scheduling of Dynamic Cloud Workflows with Varying Deadlines via Mixture-of-Experts
Ya Shen, Gang Chen, Hui Ma, Mengjie Zhang
Abstract
Workflow scheduling in cloud computing demands the intelligent allocation of dynamically arriving, graph-structured workflows with varying deadlines onto ever-changing virtual machine resources. However, existing deep reinforcement learning (DRL) schedulers remain limited by rigid, single-path inference architectures that struggle to handle diverse scheduling scenarios. We introduce (eadline-prceptive Mixture-o-Expers), an innovative DRL policy architecture that leverages a specialized mixture of experts, each trained to manage different levels of deadline tightness. To our knowledge, DEFT is the first to introduce and validate a Mixture-of-Experts architecture for dynamic cloud workflow scheduling. By adaptively routing decisions through the most appropriate experts, DEFT is capable of meeting a broad spectrum of deadline requirements that no single expert can achieve. Central to DEFT is a gating mechanism that encodes workflow DAGs, task states, and VM conditions, using cross-attention to guide expert activation in a fine-grained, deadline-sensitive manner. Experiments on dynamic cloud workflow benchmarks demonstrate that DEFT significantly reduces execution cost and deadline violations, outperforming multiple state-of-the-art DRL baselines.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on5
- GLaM: Efficient Scaling of Language Models with Mixture-of-ExpertsNan Du, Yanping Huang, Andrew M. Dai, Simon Tong et al.ICML 2022 · 1,173 citations
- MOReL: Model-Based Offline Reinforcement LearningRahul Kidambi, Aravind Rajeswaran, Praneeth Netrapalli, Thorsten JoachimsNeurIPS 2020 · 870 citations
- Multi-Decoder Attention Model with Embedding Glimpse for Solving Vehicle Routing ProblemsLiang Xin, Wen Song, Zhiguang Cao, Jie ZhangAAAI 2021 · 209 citations
- Graph Assisted Offline-Online Deep Reinforcement Learning for Dynamic Workflow SchedulingYifan Yang, Gang Chen, Hui Ma, Cong Zhang et al.ICLR 2025
- SHIELD: Multi-task Multi-distribution Vehicle Routing Solver with Sparsity and HierarchyYong Liang Goh, Zhiguang Cao, Yining Ma, Jianan Zhou et al.ICML 2025
Related papers
- Two Experts Are All You Need for Steering Thinking: Reinforcing Cognitive Effort in MoE Reasoning Models Without Additional TrainingMengru Wang, Xingyu Chen, Yue Wang, Zhiwei He et al.NeurIPS 2025 · 19 citations
- DRAE: Dynamic Retrieval-Augmented Expert Networks for Lifelong Learning and Task Adaptation in RoboticsYayu Long, Kewei Chen, Long Jin, Mingsheng ShangACL 2025 · 6 citations
- Phase-Aware Mixture of Experts for Agentic Reinforcement LearningYang Shengtian, Ziteng Cui, Shuo He, Yewen Li et al.ICML 2026 · 2 citations
- Mirage: Towards Low-interruption Services on Batch GPU Clusters with Reinforcement LearningQiyang Ding, Pengfei Zheng, Shreyas Kudari, Shivaram Venkataraman et al.SC 2023 · 5 citations
- Offline Reinforcement Learning for Mixture-of-Expert Dialogue ManagementDhawal Gupta, Yinlam Chow, Azamat Tulepbergenov, Mohammad Ghavamzadeh et al.NeurIPS 2023 · 7 citations
