AlphaAgentEvo: Evolution-Oriented Alpha Mining via Self-Evolving Agentic Reinforcement Learning
Ziyi Tang, Xuexiong Yin, Weixing Chen, Zechuan Chen, Yongsen Zheng, Wenxuan Ye, Keze Wang, Liang Lin
Abstract
Alpha mining seeks to identify predictive alpha factors that generate excess returns relative to the market from a vast and noisy search space; however, existing evolution-based approaches struggle to facilitate the systematic evolution of alphas. Traditional methods, such as Genetic Programming (GP), cannot interpret natural language instructions and often fail to extract valuable insights from unsuccessful attempts, leading to low interpretability and inefficient exploration. Analogously, without mechanisms for systematic evolution, e.g., long-term planning and reflection, existing multi-agent approaches may easily fall into repetitive evolutionary routines, resulting in inefficient evolution. To overcome these limitations, we introduce AlphaAgentEvo, a self-evolving Agentic Reinforcement Learning (ARL) framework for alpha mining, which moves alpha mining beyond the brittle "search-backtest-restart" cycle toward a continuous trajectory of evolution. Guided by a hierarchical reward function, our agent engages in selfexploration of the search space, progressively learning basic requirements (e.g., valid tool calls) and then more complex objectives (e.g., continuous performance improvements). Through this process, the agent acquires advanced behaviors such as long-horizon planning and reflective reasoning, which enable it to actively react to the underlying state (e.g., market regime shifts) and realize a self-evolving agent, marking a step toward more principled and scalable alpha mining. Extensive experiments demonstrate that AlphaAgentEvo achieves more efficient alpha evolution and generates diverse and transferable alphas, consistently surpassing a wide range of baselines. Notably, with only 4B parameters, it outperforms LLMdriven evolution methods configured with state-of-the-art closed-source reasoning models, highlighting the promise of ARL for next-generation alpha mining.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6839d879-decc-42b7-8adb-8b330ceff168Builds on9
- GEPA: Reflective Prompt Evolution Can Outperform Reinforcement LearningLakshya A. Agrawal, Shangyin Tan, Dilara Soylu, Noah Ziems et al.ICLR 2026 · 466 citations
- EvoPrompting: Language Models for Code-Level Neural Architecture SearchAngelica Chen, David Dohan, David R. SoNeurIPS 2023 · 184 citations
- ReSearch: Learning to Reason with Search for LLMs via Reinforcement LearningMingyang Chen, Linzhuang Sun, Tianpeng Li, Haoze Sun et al.NeurIPS 2025 · 125 citations
- HybridFlow: A Flexible and Efficient RLHF FrameworkGuangming Sheng, Chi Zhang, Zilingfeng Ye, Xibin Wu et al.EuroSys 2025 · 61 citations
- StockMixer: A Simple Yet Strong MLP-Based Architecture for Stock Price ForecastingJinyong Fan, Yanyan ShenAAAI 2024 · 43 citations
Related papers
- Cognitive Alpha Mining via LLM-Driven Code-Based EvolutionFengyuan Liu, Yi Huang, Sichun Luo, Yuqi Wang et al.ACL 2026 · 3 citations
- Navigating the Alpha Jungle: An LLM-Powered MCTS Framework for Formulaic Alpha Factor MiningYu Shi, Yitong Duan, Jian LiAAAI 2026 · 11 citations
- AlphaAgent: LLM-Driven Alpha Mining with Regularized Exploration to Counteract Alpha DecayZiyi Tang, Zechuan Chen, Jiarui Yang, Jiayao Mai et al.KDD 2025 · 3 citations
- AlphaEval: A Comprehensive and Efficient Evaluation Framework for Formula Alpha MiningHongjun Ding, Binqi Chen, Jinsheng Huang, Taian Guo et al.KDD 2026 · 11 citations
- AlphaMaster: Dual-Chain Feedback for Scalable and Diverse Alpha Factor DiscoveryHaozengran Wang, Shuo Yin, Rong Fu, Mengting Zhang et al.KDD 2026 · 2 citations
