Deep Reinforcement Learning Guided Improvement Heuristic for Job Shop Scheduling
Cong Zhang, Zhiguang Cao, Wen Song, Yaoxin Wu, Jie Zhang
Abstract
Recent studies in using deep reinforcement learning (DRL) to solve Job-shop scheduling problems (JSSP) focus on construction heuristics. However, their performance is still far from optimality, mainly because the underlying graph representation scheme is unsuitable for modelling partial solutions at each construction step. This paper proposes a novel DRL-guided improvement heuristic for solving JSSP, where graph representation is employed to encode complete solutions. We design a Graph Neural-Network-based representation scheme, consisting of two modules to effectively capture the information of dynamic topology and different types of nodes in graphs encountered during the improvement process. To speed up solution evaluation during improvement, we present a novel message-passing mechanism that can evaluate multiple solutions simultaneously. We prove that the computational complexity of our method scales linearly with problem size. Experiments on classic benchmarks show that the improvement policy learned by our method outperforms state-of-the-art DRL-based methods by a large margin.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers10
- Self-Labeling the Job Shop Scheduling ProblemAndrea Corsini, Angelo Porrello, Simone Calderara, Mauro Dell'AmicoNeurIPS 2024 · 39 citations
- Towards Efficient Constraint Handling in Neural Solvers for Routing ProblemsJieyi Bi, Zhiguang Cao, Jianan Zhou, Wen Song et al.ICLR 2026 · 5 citations
- Towards Generalizable Multi-Policy Optimization with Self-Evolution for Job SchedulingInguk Choi, Woo-Jin Shin, Sang-Hyun Cho, Hyun-Jung KimNeurIPS 2025 · 4 citations
- Instance-wise Adaptive Scheduling via Derivative-Free Meta-LearningHefang Qing, Miao Zhang, Yaoxin Wu, Weinan Huang et al.ICLR 2026
- Learning-Guided Rolling Horizon Optimization for Long-Horizon Flexible Job-Shop SchedulingSirui Li, Wenbin Ouyang, Yining Ma, Cathy WuICLR 2025
Builds on8
- Learning to Dispatch for Job Shop Scheduling via Deep Reinforcement LearningCong Zhang, Wen Song, Zhiguang Cao, Jie Zhang et al.NeurIPS 2020 · 497 citations
- A Learning-based Iterative Method for Solving Vehicle Routing ProblemsHao Lu, Xingwen Zhang, Shuang YangICLR 2020 · 270 citations
- Multi-Decoder Attention Model with Embedding Glimpse for Solving Vehicle Routing ProblemsLiang Xin, Wen Song, Zhiguang Cao, Jie ZhangAAAI 2021 · 209 citations
- NeuroLKH: Combining Deep Learning Model with Lin-Kernighan-Helsgaun Heuristic for Solving the Traveling Salesman ProblemLiang Xin, Wen Song, Zhiguang Cao, Jie ZhangNeurIPS 2021 · 202 citations
- Matrix encoding networks for neural combinatorial optimizationYeong-Dae Kwon, Jinho Choo, Iljoo Yoon, Minah Park et al.NeurIPS 2021 · 172 citations
Related papers
- Learning Memory-Enhanced Improvement Heuristics for Flexible Job Shop SchedulingJiaqi Wang, Zhiguang Cao, Peng Zhao, Rui Cao et al.NeurIPS 2025
- RESCHED: Rethinking Flexible Job Shop Scheduling from a Transformer-based Architecture with Simplified StatesXiangjie Xiao, Cong Zhang, Wen Song, Zhiguang CaoICLR 2026 · 2 citations
- Fast Approximations for Job Shop Scheduling: A Lagrangian Dual Deep Learning MethodJames Kotary, Ferdinando Fioretto, Pascal Van HentenryckAAAI 2022 · 27 citations
- Dual Operation Aggregation Graph Neural Networks for Solving Flexible Job-Shop Scheduling Problem with Reinforcement LearningPeng Zhao, You Zhou, Di Wang, Zhiguang Cao et al.WWW 2025 · 3 citations
- Neural DAG Scheduling via One-Shot Priority SamplingWonseok Jeon, Mukul Gagrani, Burak Bartan, Weiliang Will Zeng et al.ICLR 2023
