Evolution of Benchmark: Black-Box Optimization Benchmark Design through Large Language Model
Chen Wang, Sijie Ma, Zeyuan Ma, Yue-Jiao Gong
Abstract
Benchmark Design in Black-Box Optimization (BBO) is a fundamental yet open-ended topic. Early BBO benchmarks are predominantly human-crafted, introducing expert bias and constraining diversity. Automating this design process can relieve the human-in-the-loop burden while enhancing diversity and objectivity. We propose Evolution of Benchmark (EoB), an automated BBO benchmark designer empowered by the large language model (LLM) and its program evolution capability. Specifically, we formulate benchmark design as a bi-objective optimization problem towards maximizing (i) landscape similarity to target tasks and (ii) algorithm-differentiation ability across a portfolio of BBO solvers. Under this paradigm, EoB iteratively prompts LLM to evolve a population of benchmark programs and employs a reflection-based scheme to co-evolve the landscape and its corresponding program. Comprehensive experiments validate our EoB is a competitive candidate in multi-dimensional usages: 1) Benchmarking BBO algorithms; 2) Training and testing learning-assisted BBO algorithms; 3) Extending proxy for expensive real-world problems.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext da5efcb3-0e18-4927-bc81-741ec4fccbbaBuilds on12
- Connecting Large Language Models with Evolutionary Algorithms Yields Powerful Prompt OptimizersQingyan Guo, Rui Wang, Junliang Guo, Bei Li et al.ICLR 2024 · 257 citations
- Evolution of Heuristics: Towards Efficient Automatic Algorithm Design Using Large Language ModelFei Liu, Xialiang Tong, Mingxuan Yuan, Xi Lin et al.ICML 2024 · 238 citations
- EvoPrompting: Language Models for Code-Level Neural Architecture SearchAngelica Chen, David Dohan, David R. SoNeurIPS 2023 · 184 citations
- Multi-Objective Evolution of Heuristic Using Large Language ModelShunyu Yao, Fei Liu, Xi Lin, Zhichao Lu et al.AAAI 2025 · 48 citations
- Pretrained Optimization Model for Zero-Shot Black Box OptimizationXiaobin Li, Kai Wu, Yujian Betterest Li, Xiaoyu Zhang et al.NeurIPS 2024 · 23 citations
Related papers
- ReEvo: Large Language Models as Hyper-Heuristics with Reflective EvolutionHaoran Ye, Jiarui Wang, Zhiguang Cao, Federico Berto et al.NeurIPS 2024 · 424 citations
- On the Evaluation of Large Language Models in Unit Test Evolution (Experience Paper)Weichang Liu, Junwei Zhang, Yuqing Niu, Bo ZhouISSTA 2026
- Co-Evolution of Large Language Models and Configuration Strategies to Enhance Surrogate-Assisted Evolutionary AlgorithmLindong Xie, Yang Zhang, Zhixian Tang, Edward Chung et al.KDD 2025
- Learn to Relax with Large Language Models: Solving Constraint Optimization Problems via Bidirectional CoevolutionBeidan Liu, Zhengqiu Zhu, Chen Gao, Tianle Pu et al.ACL 2026
- HSEvo: Elevating Automatic Heuristic Design with Diversity-Driven Harmony Search and Genetic Algorithm Using LLMsPham Vu Tuan Dat, Long Doan, Huynh Thi Thanh BinhAAAI 2025 · 3 citations
