Game-Theoretic Co-Evolution for LLM-Based Heuristic Discovery
Xinyi Ke, Kai Li, Junliang Xing, Yifan Zhang, Jian Cheng
Abstract
Large language models (LLMs) have enabled rapid progress in automatic heuristic discovery (AHD), yet most existing methods are predominantly limited by static evaluation against fixed instance distributions, leading to potential overfitting and poor generalization under distributional shifts. We propose Algorithm Space Response Oracles (ASRO), a game-theoretic framework that reframes heuristic discovery as a program level co-evolution between solver and instance generator. ASRO models their interaction as a two-player zero-sum game, maintains growing strategy pools on both sides, and iteratively expands them via LLM-based best-response oracles against mixed opponent meta-strategies, thereby replacing static evaluation with an adaptive, self-generated curriculum. Across multiple combinatorial optimization domains, ASRO consistently outperforms static-training AHD baselines built on the same program search mechanisms, achieving substantially improved generalization and robustness on diverse and out-of-distribution instances.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b165253e-a515-46e4-b47a-77f15ff179e2Builds on17
- Reflexion: language agents with verbal reinforcement learningNoah Shinn, Federico Cassano, Ashwin Gopinath, Karthik Narasimhan et al.NeurIPS 2023 · 5,828 citations
- Self-Refine: Iterative Refinement with Self-FeedbackAman Madaan, Niket Tandon, Prakhar Gupta, Skyler Hallinan et al.NeurIPS 2023 · 4,972 citations
- Neural Combinatorial Optimization with Heavy Decoder: Toward Large Scale GeneralizationFu Luo, Xi Lin, Fei Liu, Qingfu Zhang et al.NeurIPS 2023 · 248 citations
- Evolution of Heuristics: Towards Efficient Automatic Algorithm Design Using Large Language ModelFei Liu, Xialiang Tong, Mingxuan Yuan, Xi Lin et al.ICML 2024 · 238 citations
- DIMES: A Differentiable Meta Solver for Combinatorial Optimization ProblemsRuizhong Qiu, Zhiqing Sun, Yiming YangNeurIPS 2022 · 183 citations
Related papers
- HiFo-Prompt: Prompting with Hindsight and Foresight for LLM-based Automatic Heuristic DesignChentongChen, Mengyuan Zhong, Jialong Shi, Jianyong Sun et al.ICLR 2026 · 17 citations
- DEPT: Large Language Model–Driven Automated Algorithm Design via Evolutionary Program TreesBin Chen, Shouliang Zhu, Beidan Liu, Yong Zhao et al.ICML 2026 · 3 citations
- EoH-S: Evolution of Heuristic Set Using LLMs for Automated Heuristic DesignFei Liu, Yilu Liu, Qingfu Zhang, Xialiang Tong et al.AAAI 2026 · 11 citations
- Hierarchical Representations for Cross-task Automated Heuristic Design using LLMsFei Liu, Rui Zhang, Shunyu Yao, Qinglong Hu et al.ICML 2026
- Generalizable Heuristic Generation Through LLMs with Meta-OptimizationYiding Shi, Jianan Zhou, Wen Song, Jieyi Bi et al.ICLR 2026 · 14 citations
