Evolution of Benchmark: Black-Box Optimization Benchmark Design through Large Language Model
Chen Wang, Sijie Ma, Zeyuan Ma, Yue-Jiao Gong
摘要
Benchmark Design in Black-Box Optimization (BBO) is a fundamental yet open-ended topic. Early BBO benchmarks are predominantly human-crafted, introducing expert bias and constraining diversity. Automating this design process can relieve the human-in-the-loop burden while enhancing diversity and objectivity. We propose Evolution of Benchmark (EoB), an automated BBO benchmark designer empowered by the large language model (LLM) and its program evolution capability. Specifically, we formulate benchmark design as a bi-objective optimization problem towards maximizing (i) landscape similarity to target tasks and (ii) algorithm-differentiation ability across a portfolio of BBO solvers. Under this paradigm, EoB iteratively prompts LLM to evolve a population of benchmark programs and employs a reflection-based scheme to co-evolve the landscape and its corresponding program. Comprehensive experiments validate our EoB is a competitive candidate in multi-dimensional usages: 1) Benchmarking BBO algorithms; 2) Training and testing learning-assisted BBO algorithms; 3) Extending proxy for expensive real-world problems.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper12
- Connecting Large Language Models with Evolutionary Algorithms Yields Powerful Prompt OptimizersQingyan Guo, Rui Wang, Junliang Guo, Bei Li 等ICLR 2024 · 被引用 257 次
- Evolution of Heuristics: Towards Efficient Automatic Algorithm Design Using Large Language ModelFei Liu, Xialiang Tong, Mingxuan Yuan, Xi Lin 等ICML 2024 · 被引用 238 次
- EvoPrompting: Language Models for Code-Level Neural Architecture SearchAngelica Chen, David Dohan, David R. SoNeurIPS 2023 · 被引用 184 次
- Multi-Objective Evolution of Heuristic Using Large Language ModelShunyu Yao, Fei Liu, Xi Lin, Zhichao Lu 等AAAI 2025 · 被引用 48 次
- Pretrained Optimization Model for Zero-Shot Black Box OptimizationXiaobin Li, Kai Wu, Yujian Betterest Li, Xiaoyu Zhang 等NeurIPS 2024 · 被引用 23 次
相关 Paper
- ReEvo: Large Language Models as Hyper-Heuristics with Reflective EvolutionHaoran Ye, Jiarui Wang, Zhiguang Cao, Federico Berto 等NeurIPS 2024 · 被引用 424 次
- On the Evaluation of Large Language Models in Unit Test Evolution (Experience Paper)Weichang Liu, Junwei Zhang, Yuqing Niu, Bo ZhouISSTA 2026
- Co-Evolution of Large Language Models and Configuration Strategies to Enhance Surrogate-Assisted Evolutionary AlgorithmLindong Xie, Yang Zhang, Zhixian Tang, Edward Chung 等KDD 2025
- Learn to Relax with Large Language Models: Solving Constraint Optimization Problems via Bidirectional CoevolutionBeidan Liu, Zhengqiu Zhu, Chen Gao, Tianle Pu 等ACL 2026
- HSEvo: Elevating Automatic Heuristic Design with Diversity-Driven Harmony Search and Genetic Algorithm Using LLMsPham Vu Tuan Dat, Long Doan, Huynh Thi Thanh BinhAAAI 2025 · 被引用 3 次
