Reinforcement Learning with Maskable Stock Representation for Portfolio Management in Customizable Stock Pools
Wentao Zhang, Yilei Zhao, Shuo Sun, Jie Ying, Yonggang Xie, Zitao Song, Xinrun Wang, Bo An
Abstract
Portfolio management (PM) is a fundamental financial trading task, which explores the optimal periodical reallocation of capitals into different stocks to pursue long-term profits. Reinforcement learning (RL) has recently shown its potential to train profitable agents for PM through interacting with financial markets. However, existing work mostly focuses on fixed stock pools, which is inconsistent with investors' practical demand. Specifically, the target stock pool of different investors varies dramatically due to their discrepancy on market states and individual investors may temporally adjust stocks they desire to trade (e.g., adding one popular stocks), which lead to customizable stock pools (CSPs). Existing RL methods require to retrain RL agents even with a tiny change of the stock pool, which leads to high computational cost and unstable performance. To tackle this challenge, we propose EarnMore, a rEinforcement leARNing framework with Maskable stOck REpresentation to handle PM with CSPs through one-shot training in a global stock pool (GSP). Specifically, we first introduce a mechanism to mask out the representation of the stocks outside the target pool. Second, we learn meaningful stock representations through a self-supervised masking and reconstruction process. Third, a re-weighting mechanism is designed to make the portfolio concentrate on favorable stocks and neglect the stocks outside the target pool. Through extensive experiments on 8 subset stock pools of the US stock market, we demonstrate that EarnMore significantly outperforms 14 stateof-the-art baselines in terms of 6 popular financial metrics with over 40% improvement on profit. Code is available in PyTorch 1 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 27e2509f-e090-4400-9af2-e32e9bc086c8Cited by top-tier papers4
- Pre-training Time Series Models with Stock Data CustomizationMengyu Wang, Tiejun Ma, Shay B. CohenKDD 2025 · 1 citation
- Regime-Adaptive Continual Learning for Portfolio ManagementChaofan Pan, Lingfei Ren, Linbo Xiong, Yonghao Li et al.KDD 2026
- MarS: a Financial Market Simulation Engine Powered by Generative Foundation ModelJunjie Li, Yang Liu, Weiqing Liu, Shikai Fang et al.ICLR 2025
- FineFT: Efficient and Risk-Aware Ensemble Reinforcement Learning for Futures TradingMolei Qin, Xinyu Cai, Yewen Li, Haochong Xia et al.KDD 2026
Builds on10
- A Time Series is Worth 64 Words: Long-term Forecasting with TransformersYuqi Nie, Nam H. Nguyen, Phanwadee Sinthong, Jayant KalagnanamICLR 2023 · 536 citations
- TimesNet: Temporal 2D-Variation Modeling for General Time Series AnalysisHaixu Wu, Tengge Hu, Yong Liu, Hang Zhou et al.ICLR 2023 · 423 citations
- SimMTM: A Simple Pre-Training Framework for Masked Time-Series ModelingJiaxiang Dong, Haixu Wu, Haoran Zhang, Li Zhang et al.NeurIPS 2023 · 225 citations
- Reinforcement-Learning Based Portfolio Management with Augmented Asset Movement Prediction StatesYunan Ye, Hengzhi Pei, Boxin Wang, Pin-Yu Chen et al.AAAI 2020 · 181 citations
- DeepTrader: A Deep Reinforcement Learning Approach for Risk-Return Balanced Portfolio Management with Market Conditions EmbeddingZhicheng Wang, Biwei Huang, Shikui Tu, Kun Zhang et al.AAAI 2021 · 154 citations
Related papers
- Commission Fee is not Enough: A Hierarchical Reinforced Framework for Portfolio ManagementRundong Wang, Hongxin Wei, Bo An, Zhouyan Feng et al.AAAI 2021 · 50 citations
- EarnHFT: Efficient Hierarchical Reinforcement Learning for High Frequency TradingMolei Qin, Shuo Sun, Wentao Zhang, Haochong Xia et al.AAAI 2024 · 28 citations
- MARS: A Meta-Adaptive Reinforcement Learning Framework for Risk-Aware Multi-Agent Portfolio ManagementJiayi Chen, Jing Li, Guiling WangAAAI 2026 · 2 citations
- Autoregressive Policy Optimization for Constrained Allocation TasksDavid Winkel, Niklas Strauß, Maximilian Bernhard, Zongyue Li et al.NeurIPS 2024 · 2 citations
- MetaTrader: Learning to Generalize RL Trading Policies Beyond Offline DataHaochen Yuan, Minting Pan, Yunbo Wang, Siyu Gao et al.AAAI 2026
