Trajectory-Aware Heuristic Learning for Combinatorial Search
Mustafa Seddiqi, Marta Kersten-Oertel, Tiberiu Popa
摘要
Learning effective value heuristics for combinatorial search is difficult, as prior methods rely on surrogate supervision or costly downstream search to assess progress. We introduce a trajectory-aware probabilistic framework that models uncertainty in cost-to-go labels instead of treating them as fixed targets. Heuristic learning is cast as inference over state trajectories using an HMM-style model, where estimated depth-change dynamics define transitions and forward-backward inference yields soft supervision. To evaluate heuristic quality without search, we propose a large-scale local ranking metric that measures a model's ability to order neighboring states. On the Rubik's Cube, our approach consistently improves local ranking accuracy and downstream search performance under matched computational budgets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper2
相关 Paper
- Learning Admissible Heuristics for A*: Theory and PracticeEhsan Futuhi, Nathan R. SturtevantICLR 2026 · 被引用 3 次
- Can We Learn Heuristics for Graphical Model Inference Using Reinforcement Learning?Safa Messaoud, Maghav Kumar, Alexander G. SchwingCVPR 2020
- Contrastive Representations for Temporal ReasoningAlicja Ziarko, Michal Bortkiewicz, Michal Zawalski, Benjamin Eysenbach 等NeurIPS 2025 · 被引用 8 次
- Goal Recognition as Reinforcement LearningLeonardo Amado, Reuth Mirsky, Felipe MeneguzziAAAI 2022 · 被引用 22 次
- Large-State Reinforcement Learning for Hyper-HeuristicsLucas Kletzander, Nysret MusliuAAAI 2023 · 被引用 12 次
