AlphaQCM: Alpha Discovery in Finance with Distributional Reinforcement Learning
Zhoufan Zhu, Ke Zhu
摘要
For researchers and practitioners in finance, finding synergistic formulaic alphas is very important but challenging. In this paper, we reconsider the discovery of synergistic formulaic alphas from the viewpoint of sequential decision-making, and conceptualize the entire alpha discovery process as a non-stationary and reward-sparse Markov decision process. To overcome the challenges of non-stationarity and reward-sparsity, we propose the AlphaQCM method, a novel distributional reinforcement learning method designed to search for synergistic formulaic alphas efficiently. The AlphaQCM method first learns the Q function and quantiles via a Q network and a quantile network, respectively. Then, the AlphaQCM method applies the quantiled conditional moment method to learn unbiased variance from the potentially biased quantiles. Guided by the learned Q function and variance, the AlphaQCM method navigates the non-stationarity and reward-sparsity to explore the vast search space of formulaic alphas with high efficacy. Empirical applications to realworld datasets demonstrate that our AlphaQCM method significantly outperforms its competitors, particularly when dealing with large datasets comprising numerous stocks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- AlphaSAGE: Structure-Aware Alpha Mining via GFlowNets for Robust ExplorationBinqi Chen, Hongjun Ding, Ning Shen, Taian Guo 等ICLR 2026 · 被引用 18 次
- Conditional Quantile Adjusted Conformal Prediction for Time SeriesCheng Yu, Zhoufan Zhu, Ke ZhuICML 2026 · 被引用 11 次
- AlphaEval: A Comprehensive and Efficient Evaluation Framework for Formula Alpha MiningHongjun Ding, Binqi Chen, Jinsheng Huang, Taian Guo 等KDD 2026 · 被引用 11 次
- CTBench: Cryptocurrency Time Series Generation BenchmarkYihao Ang, Qiang Wang, Qiang Huang, Yifan Bao 等ICLR 2026 · 被引用 5 次
- AlphaAgentEvo: Evolution-Oriented Alpha Mining via Self-Evolving Agentic Reinforcement LearningZiyi Tang, Xuexiong Yin, Weixing Chen, Zechuan Chen 等ICLR 2026
它引用的顶会 Paper3
- Deep symbolic regression: Recovering mathematical expressions from data via risk-seeking policy gradientsBrenden K. Petersen, Mikel Landajuela, T. Nathan Mundhenk, Cláudio Prata Santiago 等ICLR 2021 · 被引用 444 次
- AutoML-Zero: Evolving Machine Learning Algorithms From ScratchEsteban Real, Chen Liang, David R. So, Quoc V. LeICML 2020 · 被引用 265 次
- AlphaEvolve: A Learning Framework to Discover Novel Alphas in Quantitative InvestmentCan Cui, Wei Wang, Meihui Zhang, Gang Chen 等SIGMOD 2021 · 被引用 25 次
相关 Paper
- AlphaForge: A Framework to Mine and Dynamically Combine Formulaic Alpha FactorsHao Shi, Weili Song, Xinting Zhang, Jiahe Shi 等AAAI 2025 · 被引用 23 次
- Navigating the Alpha Jungle: An LLM-Powered MCTS Framework for Formulaic Alpha Factor MiningYu Shi, Yitong Duan, Jian LiAAAI 2026 · 被引用 11 次
- Being Optimistic to Be Conservative: Quickly Learning a CVaR PolicyRamtin Keramati, Christoph Dann, Alex Tamkin, Emma BrunskillAAAI 2020 · 被引用 86 次
- Nonparametric Quantile Regression with ReLU-Activated Recurrent Neural NetworksHan Yu, Lyumin Wu, Wenxin Zhou, Zhao RenNeurIPS 2025 · 被引用 1 次
- Variance Control for Distributional Reinforcement LearningQi Kuang, Zhoufan Zhu, Liwen Zhang, Fan ZhouICML 2023 · 被引用 4 次
