Mastering Stock Markets with Efficient Mixture of Diversified Trading Experts
Shuo Sun, Xinrun Wang, Wanqi Xue, Xiaoxuan Lou, Bo An
Abstract
Quantitative stock investment is a fundamental financial task that highly relies on accurate prediction of market status and profitable investment decision making. Despite recent advances in deep learning (DL) have shown stellar performance on capturing trading opportunities in the stochastic stock market, the performance of existing DL methods is unstable with sensitivity to network initialization and hyperparameter selection. One major limitation of existing works is that investment decisions are made based on one individual neural network predictor with high uncertainty, which is inconsistent with the workflow in real-world trading firms. To tackle this limitation, we propose AlphaMix, a novel three-stage mixture-of-experts (MoE) framework for quantitative investment to mimic the efficient bottom-up hierarchical trading strategy design workflow of successful trading companies. In Stage one, we introduce an efficient ensemble learning method, whose computational and memory costs are significantly lower comparing to traditional ensemble methods, to train multiple groups of trading experts with personalised market understanding and trading styles. In Stage two, we collect diversified investment suggestions through building a pool of trading experts utilizing hyperparameter level and initialization level diversity of neural networks for post hoc ensemble construction. In Stage three, we design three different mechanisms, namely as-needed router, with-replacement selection and integrated expert soup, to dynamically pick experts from the expert pool, which takes the responsibility of a portfolio manager. Through extensive experiments on US and Chinese stock markets, we demonstrate that AlphaMix significantly outperforms many state-of-the-art baselines in terms of 7 popular financial criteria.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers9
- A Multimodal Foundation Agent for Financial Trading: Tool-Augmented, Diversified, and GeneralistWentao Zhang, Lingxuan Zhao, Haochong Xia, Shuo Sun et al.KDD 2024 · 50 citations
- EarnHFT: Efficient Hierarchical Reinforcement Learning for High Frequency TradingMolei Qin, Shuo Sun, Wentao Zhang, Haochong Xia et al.AAAI 2024 · 28 citations
- DHMoE: Diffusion Generated Hierarchical Multi-Granular Expertise for Stock PredictionWeijun Chen, Yanze WangAAAI 2025 · 6 citations
- Logic-Q: Improving Deep Reinforcement Learning-based Quantitative Trading via Program Sketch-based TuningZhiming Li, Junzhe Jiang, Yushi Cao, Aixin Cui et al.AAAI 2025 · 5 citations
- Two Heads Are Better than One: Distilling Large Language Model Features into Small Models with Feature Decomposition and MixtureTianhao Fu, Xinxin Xu, Weichen Xu, Jue Chen et al.AAAI 2026 · 2 citations
Builds on15
- Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference timeMitchell Wortsman, Gabriel Ilharco, Samir Yitzhak Gadre, Rebecca Roelofs et al.ICML 2022 · 1,464 citations
- GLaM: Efficient Scaling of Language Models with Mixture-of-ExpertsNan Du, Yanping Huang, Andrew M. Dai, Simon Tong et al.ICML 2022 · 1,173 citations
- What is being transferred in transfer learning?Behnam Neyshabur, Hanie Sedghi, Chiyuan ZhangNeurIPS 2020 · 654 citations
- BatchEnsemble: an Alternative Approach to Efficient Ensemble and Lifelong LearningYeming Wen, Dustin Tran, Jimmy BaICLR 2020 · 569 citations
- Long-tailed Recognition by Routing Diverse Distribution-Aware ExpertsXudong Wang, Long Lian, Zhongqi Miao, Ziwei Liu et al.ICLR 2021 · 481 citations
Related papers
- Towards Understanding the Mixture-of-Experts Layer in Deep LearningZixiang Chen, Yihe Deng, Yue Wu, Quanquan Gu et al.NeurIPS 2022 · 199 citations
- StockMixer: A Simple Yet Strong MLP-Based Architecture for Stock Price ForecastingJinyong Fan, Yanyan ShenAAAI 2024 · 43 citations
- HyperMoE: Towards Better Mixture of Experts via Transferring Among ExpertsHao Zhao, Zihan Qiu, Huijia Wu, Zili Wang et al.ACL 2024
- Disentangled Dual-Granularity Learning for Market-Adaptive Stock Return Forecasting under Non-Stationary EnvironmentsMinghui Su, Xiaobo Guo, Deyu Tian, Binfeng Wang et al.KDD 2026
- Harder Task Needs More Experts: Dynamic Routing in MoE ModelsQuzhe Huang, Zhenwei An, Nan Zhuang, Mingxu Tao et al.ACL 2024 · 11 citations
