Pre-training Time Series Models with Stock Data Customization
Mengyu Wang, Tiejun Ma, Shay B. Cohen
Abstract
Stock selection, which aims to predict stock prices and identify the most profitable ones, is a crucial task in finance. While existing methods primarily focus on developing model structures and building graphs for improved selection, pre-training strategies remain underexplored in this domain. Current stock series pre-training follows methods from other areas without adapting to the unique characteristics of financial data, particularly overlooking stock-specific contextual information and the non-stationary nature of stock prices. Consequently, the latent statistical features inherent in stock data are underutilized. In this paper, we propose three novel pre-training tasks tailored to stock data characteristics: stock code classification, stock sector classification, and moving average prediction. We develop the Stock Specialized Pre-trained Transformer (SSPT) based on a two-layer transformer architecture. Extensive experimental results validate the effectiveness of our pre-training methods and provide detailed guidance on their application. Evaluations on five stock datasets, including four markets and two time periods, demonstrate that SSPT consistently outperforms the market and existing methods in terms of both cumulative investment return ratio and Sharpe ratio. Additionally, our experiments on simulated data investigate the underlying mechanisms of our methods, providing insights into understanding price series. Our code is publicly available at: https://github.com/astudentuser/Pre-training-Time-Series-Models-with-Stock-Data-Customization.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- FinKario: Event-Enhanced Automated Construction of Financial Knowledge GraphXiang Li, Penglei Sun, Wanyun Zhou, Zikai Wei et al.ACL 2026 · 4 citations
- The Label Horizon Paradox: Rethinking Supervision Targets in Financial ForecastingChen-Hui Song, Shuoling Liu, Liyuan ChenICML 2026 · 1 citation
- Terminal Dimension Reduction for Time Series with ApplicationsAlexander Munteanu, Matteo Russo, David Saulpic, Chris SchwiegelshohnICML 2026
Builds on18
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li et al.ICLR 2021 · 7,353 citations
- One Fits All: Power General Time Series Analysis by Pretrained LMTian Zhou, Peisong Niu, Xue Wang, Liang Sun et al.NeurIPS 2023 · 1,178 citations
- TimeMixer: Decomposable Multiscale Mixing for Time Series ForecastingShiyu Wang, Haixu Wu, Xiaoming Shi, Tengge Hu et al.ICLR 2024 · 573 citations
- Self-Supervised Contrastive Pre-Training For Time Series via Time-Frequency ConsistencyXiang Zhang, Ziyuan Zhao, Theodoros Tsiligkaridis, Marinka ZitnikNeurIPS 2022 · 558 citations
Related papers
- CI-STHPAN: Pre-trained Attention Network for Stock Selection with Channel-Independent Spatio-Temporal HypergraphHongjie Xia, Huijie Ao, Long Li, Yu Liu et al.AAAI 2024 · 48 citations
- StockMixer: A Simple Yet Strong MLP-Based Architecture for Stock Price ForecastingJinyong Fan, Yanyan ShenAAAI 2024 · 43 citations
- MASTER: Market-Guided Stock Transformer for Stock Price ForecastingTong Li, Zhaoyang Liu, Yanyan Shen, Xue Wang et al.AAAI 2024 · 80 citations
- Integrating Inductive Biases in Transformers via Distillation for Financial Time Series ForecastingYu-Chen Den, Kuan-Yu Chen, Kendro Vincent, Tien-Hao ChangKDD 2026 · 1 citation
- MDGNN: Multi-Relational Dynamic Graph Neural Network for Comprehensive and Dynamic Stock Investment PredictionHao Qian, Hongting Zhou, Qian Zhao, Hao Chen et al.AAAI 2024 · 65 citations
