Learning to Dispatch for Job Shop Scheduling via Deep Reinforcement Learning
Cong Zhang, Wen Song, Zhiguang Cao, Jie Zhang, Puay Siew Tan, Chi Xu
Abstract
Priority dispatching rule (PDR) is widely used for solving real-world Job-shop scheduling problem (JSSP). However, the design of effective PDRs is a tedious task, requiring a myriad of specialized knowledge and often delivering limited performance. In this paper, we propose to automatically learn PDRs via an end-to-end deep reinforcement learning agent. We exploit the disjunctive graph representation of JSSP, and propose a Graph Neural Network based scheme to embed the states encountered during solving. The resulting policy network is size-agnostic, effectively enabling generalization on large-scale instances. Experiments show that the agent can learn high-quality PDRs from scratch with elementary raw features, and demonstrates strong performance against the best existing PDRs. The learned policies also perform well on much larger instances that are unseen in training.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7eae3c61-182b-4002-9572-47f6fde55b1cCited by top-tier papers46
- DIFUSCO: Graph-based Diffusion Solvers for Combinatorial OptimizationZhiqing Sun, Yiming YangNeurIPS 2023 · 356 citations
- Learning to Iteratively Solve Routing Problems with Dual-Aspect Collaborative TransformerYining Ma, Jingwen Li, Zhiguang Cao, Wen Song et al.NeurIPS 2021 · 230 citations
- BQ-NCO: Bisimulation Quotienting for Efficient Neural Combinatorial OptimizationDarko Drakulic, Sofia Michel, Florian Mai, Arnaud Sors et al.NeurIPS 2023 · 124 citations
- Simulation-guided Beam Search for Neural Combinatorial OptimizationJinho Choo, Yeong-Dae Kwon, Jihoon Kim, Jeongwoo Jae et al.NeurIPS 2022 · 123 citations
- Efficient Active Search for Combinatorial Optimization ProblemsAndré Hottung, Yeong-Dae Kwon, Kevin TierneyICLR 2022 · 123 citations
Builds on1
Related papers
- Deep Reinforcement Learning Guided Improvement Heuristic for Job Shop SchedulingCong Zhang, Zhiguang Cao, Wen Song, Yaoxin Wu et al.ICLR 2024 · 28 citations
- RESCHED: Rethinking Flexible Job Shop Scheduling from a Transformer-based Architecture with Simplified StatesXiangjie Xiao, Cong Zhang, Wen Song, Zhiguang CaoICLR 2026 · 2 citations
- Learning Memory-Enhanced Improvement Heuristics for Flexible Job Shop SchedulingJiaqi Wang, Zhiguang Cao, Peng Zhao, Rui Cao et al.NeurIPS 2025
- Towards Generalizable Multi-Policy Optimization with Self-Evolution for Job SchedulingInguk Choi, Woo-Jin Shin, Sang-Hyun Cho, Hyun-Jung KimNeurIPS 2025 · 4 citations
- Neural DAG Scheduling via One-Shot Priority SamplingWonseok Jeon, Mukul Gagrani, Burak Bartan, Weiliang Will Zeng et al.ICLR 2023
