Deft Scheduling of Dynamic Cloud Workflows with Varying Deadlines via Mixture-of-Experts
Ya Shen, Gang Chen, Hui Ma, Mengjie Zhang
摘要
Workflow scheduling in cloud computing demands the intelligent allocation of dynamically arriving, graph-structured workflows with varying deadlines onto ever-changing virtual machine resources. However, existing deep reinforcement learning (DRL) schedulers remain limited by rigid, single-path inference architectures that struggle to handle diverse scheduling scenarios. We introduce (eadline-prceptive Mixture-o-Expers), an innovative DRL policy architecture that leverages a specialized mixture of experts, each trained to manage different levels of deadline tightness. To our knowledge, DEFT is the first to introduce and validate a Mixture-of-Experts architecture for dynamic cloud workflow scheduling. By adaptively routing decisions through the most appropriate experts, DEFT is capable of meeting a broad spectrum of deadline requirements that no single expert can achieve. Central to DEFT is a gating mechanism that encodes workflow DAGs, task states, and VM conditions, using cross-attention to guide expert activation in a fine-grained, deadline-sensitive manner. Experiments on dynamic cloud workflow benchmarks demonstrate that DEFT significantly reduces execution cost and deadline violations, outperforming multiple state-of-the-art DRL baselines.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper5
- GLaM: Efficient Scaling of Language Models with Mixture-of-ExpertsNan Du, Yanping Huang, Andrew M. Dai, Simon Tong 等ICML 2022 · 被引用 1,173 次
- MOReL: Model-Based Offline Reinforcement LearningRahul Kidambi, Aravind Rajeswaran, Praneeth Netrapalli, Thorsten JoachimsNeurIPS 2020 · 被引用 870 次
- Multi-Decoder Attention Model with Embedding Glimpse for Solving Vehicle Routing ProblemsLiang Xin, Wen Song, Zhiguang Cao, Jie ZhangAAAI 2021 · 被引用 209 次
- Graph Assisted Offline-Online Deep Reinforcement Learning for Dynamic Workflow SchedulingYifan Yang, Gang Chen, Hui Ma, Cong Zhang 等ICLR 2025
- SHIELD: Multi-task Multi-distribution Vehicle Routing Solver with Sparsity and HierarchyYong Liang Goh, Zhiguang Cao, Yining Ma, Jianan Zhou 等ICML 2025
相关 Paper
- Two Experts Are All You Need for Steering Thinking: Reinforcing Cognitive Effort in MoE Reasoning Models Without Additional TrainingMengru Wang, Xingyu Chen, Yue Wang, Zhiwei He 等NeurIPS 2025 · 被引用 19 次
- DRAE: Dynamic Retrieval-Augmented Expert Networks for Lifelong Learning and Task Adaptation in RoboticsYayu Long, Kewei Chen, Long Jin, Mingsheng ShangACL 2025 · 被引用 6 次
- Phase-Aware Mixture of Experts for Agentic Reinforcement LearningYang Shengtian, Ziteng Cui, Shuo He, Yewen Li 等ICML 2026 · 被引用 2 次
- Mirage: Towards Low-interruption Services on Batch GPU Clusters with Reinforcement LearningQiyang Ding, Pengfei Zheng, Shreyas Kudari, Shivaram Venkataraman 等SC 2023 · 被引用 5 次
- Offline Reinforcement Learning for Mixture-of-Expert Dialogue ManagementDhawal Gupta, Yinlam Chow, Azamat Tulepbergenov, Mohammad Ghavamzadeh 等NeurIPS 2023 · 被引用 7 次
