USPR: Learning a Unified Solver for Profiled Routing
Chuanbo Hua, Federico Berto, Zhikai Zhao, Jiwoo Son, Changhyun Kwon, Jinkyoo Park
Abstract
The Profiled Vehicle Routing Problem (PVRP) extends the classical VRP by incorporating vehicle–client-specific preferences and constraints, reflecting real‑world requirements such as zone restrictions and service‑level preferences. While recent reinforcement‑learning solvers have shown promising performance, they require retraining for each new profile distribution, suffer from poor representation ability, and struggle to generalize to out‑of‑distribution instances. In this paper, we address these limitations by introducing Unified Solver for Profiled Routing (USPR), a novel framework that natively handles arbitrary profile types. USPR introduces on three key innovations: (i) Profile Embeddings (PE) to encode any combination of profile types; (ii) Multi‑Head Profiled Attention (MHPA), an attention mechanism that models rich interactions between vehicles and clients; (iii) Profile‑aware Score Reshaping (PSR), which dynamically adjusts decoder logits using profile scores to improve generalization. Empirical results on diverse PVRP benchmarks demonstrate that USPR achieves state‑of‑the‑art results among learning‑based methods while offering significant gains in flexibility and computational efficiency. We make our source code publicly available to foster future research.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8aab9fdb-c334-44fa-b3f5-dd3e7bbf19dfCited by top-tier papers1
Ask how each one uses itBuilds on22
- POMO: Policy Optimization with Multiple Optima for Reinforcement LearningYeong-Dae Kwon, Jinho Choo, Byoungjip Kim, Iljoo Yoon et al.NeurIPS 2020 · 731 citations
- ReEvo: Large Language Models as Hyper-Heuristics with Reflective EvolutionHaoran Ye, Jiarui Wang, Zhiguang Cao, Federico Berto et al.NeurIPS 2024 · 424 citations
- Evolution of Heuristics: Towards Efficient Automatic Algorithm Design Using Large Language ModelFei Liu, Xialiang Tong, Mingxuan Yuan, Xi Lin et al.ICML 2024 · 238 citations
- Learning to delegate for large-scale vehicle routingSirui Li, Zhongxia Yan, Cathy WuNeurIPS 2021 · 181 citations
- Learning to Search Feasible and Infeasible Regions of Routing Problems with Flexible Neural k-OptYining Ma, Zhiguang Cao, Yeow Meng CheeNeurIPS 2023 · 129 citations
Related papers
- Chain-of-Context Learning: Dynamic Constraint Understanding for Multi-Task VRPsShuangchun Gui, Suyu Liu, Xuehe Wang, Zhiguang CaoICLR 2026 · 3 citations
- PoMtVRS: Preference-Optimized Multi-Task Vehicle Routing Solver with Preference GatingDian Meng, Yaoxin Wu, Yaqing Hou, Zhiguang CaoICML 2026
- Elite Pattern Reinforcement for Vehicle Routing ProblemsNing Li, Peng Lin, Peng Zhang, Ruichen TianAAAI 2026
- Rethinking Light Decoder-based Solvers for Vehicle Routing ProblemsZiwei Huang, Jianan Zhou, Zhiguang Cao, Yixin XuICLR 2025
- URS: A Unified Neural Routing Solver for Cross-Problem Zero-Shot GeneralizationChangliang Zhou, Canhong Yu, Shunyu Yao, Xi Lin et al.ICML 2026 · 12 citations
