Breaking Multi-Task Curse: Reward-Weighted Evolution for Black-Box Many-Task Optimization
Yanchi Li, Jiao Liu, Wenyin Gong, Qiong Gu, Yue Zhao, Yew Soon ONG
Abstract
Evolutionary multi-tasking accelerates black-box optimization via knowledge transfer but falters in scenarios involving many low-similarity tasks. We identify this scalability barrier as the Multi-Task Curse , driven by evaluation budget dispersion and negative transfer. To overcome this, we propose MES-RET ( M any-task E volution S trategy with R eward-weighted E valuation and T ransfer), which combats budget dispersion via a reward-weighted evaluation scheme that guarantees superior expected improvement, while simultaneously mitigating negative transfer through a robust reward-weighted aggregation of mean and covariance statistics, ensuring a safe fallback to independent evolution. Furthermore, to handle neural dimensional mismatches in many-task policy search, we introduce a semantic parameter alignment strategy that bridges heterogeneous state-action spaces. Extensive experiments on synthetic benchmarks, real-world engineering problems, and reinforcement learning tasks demonstrate that MES-RET consistently outperforms state-of-the-art methods, notably enabling skill transfer across morphologically distinct policies.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on1
Related papers
- EMT-NAS: Transferring architectural knowledge between tasks from different datasetsPeng Liao, Yaochu Jin, Wenli DuCVPR 2023
- Doubly Robust Augmented Transfer for Meta-Reinforcement LearningYuankun Jiang, Nuowen Kan, Chenglin Li, Wenrui Dai et al.NeurIPS 2023 · 3 citations
- HyMTRL: A Hybrid Multi-Task Reinforcement Learning Framework via Phased Policy EvolutionJinmin He, Kai Li, Xiaoyi Dong, Yifan Zang et al.ICML 2026
- Centralized Reward Agent for Knowledge Sharing and Transfer in Multi-Task Reinforcement LearningHaozhe Ma, Zhengding Luo, Thanh Vinh Vo, Kuankuan Sima et al.NeurIPS 2025 · 9 citations
- Aligned Multi Objective OptimizationYonathan Efroni, Ben Kretzu, Daniel Jiang, Jalaj Bhandari et al.ICML 2025
