Instance-wise Adaptive Scheduling via Derivative-Free Meta-Learning
Hefang Qing, Miao Zhang, Yaoxin Wu, Weinan Huang, Jianhao Yang, Wen Song, Gang Wang
摘要
Deep Reinforcement Learning has achieved remarkable progress in solving NP-hard scheduling problems. However, existing methods primarily focus on optimizing average performance over training instances, overlooking the core objective of solving each individual instance with high quality. While several instance-wise adaptation mechanisms have been proposed, they are test-time approaches only and cannot share knowledge across different adaptation tasks. Moreover, they largely rely on gradient-based optimization, which could be ineffective in dealing with combinatorial optimization problems. We address the above issues by proposing an instance-wise meta-learning framework. It trains a meta model to acquire a generalizable initialization that effectively guides per-instance adaptation during inference, and overcomes the limitations of gradient-based methods by leveraging a derivative-free optimization scheme that is fully GPU parallelizable. Experimental results on representative scheduling problems demonstrate that our method consistently outperforms existing learning-based scheduling methods and instance-wise adaptation mechanisms under various task sizes and distributions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper12
- Learning to Dispatch for Job Shop Scheduling via Deep Reinforcement LearningCong Zhang, Wen Song, Zhiguang Cao, Jie Zhang 等NeurIPS 2020 · 被引用 497 次
- DIMES: A Differentiable Meta Solver for Combinatorial Optimization ProblemsRuizhong Qiu, Zhiqing Sun, Yiming YangNeurIPS 2022 · 被引用 183 次
- ES-MAML: Simple Hessian-Free Meta LearningXingyou Song, Wenbo Gao, Yuxiang Yang, Krzysztof Choromanski 等ICLR 2020 · 被引用 128 次
- Efficient Active Search for Combinatorial Optimization ProblemsAndré Hottung, Yeong-Dae Kwon, Kevin TierneyICLR 2022 · 被引用 123 次
- Towards Omni-generalizable Neural Methods for Vehicle Routing ProblemsJianan Zhou, Yaoxin Wu, Wen Song, Zhiguang Cao 等ICML 2023 · 被引用 90 次
相关 Paper
- Unsupervised Learning for Combinatorial Optimization Needs Meta LearningHaoyu Peter Wang, Pan LiICLR 2023 · 被引用 2 次
- Meta-SAGE: Scale Meta-Learning Scheduled Adaptation with Guided Exploration for Mitigating Scale Shift on Combinatorial OptimizationJiwoo Son, Minsu Kim, Hyeonah Kim, Jinkyoo ParkICML 2023 · 被引用 33 次
- Towards Generalizable Multi-Policy Optimization with Self-Evolution for Job SchedulingInguk Choi, Woo-Jin Shin, Sang-Hyun Cho, Hyun-Jung KimNeurIPS 2025 · 被引用 4 次
- Architecture, Dataset and Model-Scale Agnostic Data-free Meta-LearningZixuan Hu, Li Shen, Zhenyi Wang, Tongliang Liu 等CVPR 2023
- ASAP: Exploiting the Satisficing Generalization Edge in Neural Combinatorial OptimizationHan Fang, Paul Weng, Yutong BanICML 2026 · 被引用 1 次
