FineFT: Efficient and Risk-Aware Ensemble Reinforcement Learning for Futures Trading
Molei Qin, Xinyu Cai, Yewen Li, Haochong Xia, Chuqiao Zong, Shuo Sun, Xinrun Wang, Bo An
摘要
Futures are contracts obligating the exchange of an asset at a predetermined date and price, notable for their high leverage (e.g., 5-fold) and liquidity (e.g., trillions of dollars) and, therefore, thrive in the Crypto market. Reinforcement learning (RL) has been widely applied in various quantitative tasks. However, most methods focus on the spot (e.g., stock) and could not be directly applied to the futures market with high leverage because of 2 key challenges. First, high leverage amplifies reward fluctuations, making RL training highly stochastic and difficult to converge. Second, prior works lacked self-awareness of capability boundaries, exposing them to the risk of significant capital loss when encountering previously unseen market state representations (e.g., during a black swan event like COVID-19). To tackle these challenges, we propose the eFficient and rIsk-aware eNsemble rEinforcement learning for Futures Trading (FineFT), a novel three-stage ensemble RL framework with stable training and proper risk management. In stage I, ensemble Q learners are selectively updated by ensemble temporal difference (TD) errors, i.e., TD errors across different learners, to improve convergence and performance. In stage II, we filter the Q-learners based on their profitabilities under different market dynamics and train variational autoencoders (VAEs) on market representations of each dynamic to identify the capability boundaries of the filtered learners. In stage III, we dynamically choose from the filtered ensemble and a conservative policy, guided by trained VAEs, to maintain profitability and mitigate risk with new market states. Through extensive experiments on crypto futures in a high-frequency trading environment with high fidelity and 5× leverage, we demonstrate that FineFT significantly outperforms 12 state-of-the-art baselines in 6 widely-used financial metrics, reducing risk by more than 40% while achieving superior profitability compared to the runner-up. Visualization of the selective update mechanism shows that different agents specialize in distinct market dynamics, and ablation studies * Corresponding authors. 2021-10-16 2021-11-07 CCS Concepts • Computing methodologies → Artificial intelligence; Dynamic programming for Markov decision processes.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper22
- Informer: Beyond Efficient Transformer for Long Sequence Time-Series ForecastingHaoyi Zhou, Shanghang Zhang, Jieqi Peng, Shuai Zhang 等AAAI 2021 · 被引用 7,289 次
- Are Transformers Effective for Time Series Forecasting?Ailing Zeng, Muxi Chen, Lei Zhang, Qiang XuAAAI 2023 · 被引用 3,619 次
- TimeMixer: Decomposable Multiscale Mixing for Time Series ForecastingShiyu Wang, Haixu Wu, Xiaoming Shi, Tengge Hu 等ICLR 2024 · 被引用 573 次
- TimesNet: Temporal 2D-Variation Modeling for General Time Series AnalysisHaixu Wu, Tengge Hu, Yong Liu, Hang Zhou 等ICLR 2023 · 被引用 423 次
- SUNRISE: A Simple Unified Framework for Ensemble Learning in Deep Reinforcement LearningKimin Lee, Michael Laskin, Aravind Srinivas, Pieter AbbeelICML 2021 · 被引用 239 次
相关 Paper
- EarnHFT: Efficient Hierarchical Reinforcement Learning for High Frequency TradingMolei Qin, Shuo Sun, Wentao Zhang, Haochong Xia 等AAAI 2024 · 被引用 28 次
- MacroHFT: Memory Augmented Context-aware Reinforcement Learning On High Frequency TradingChuqiao Zong, Chaojie Wang, Molei Qin, Lei Feng 等KDD 2024 · 被引用 1 次
- ArchetypeTrader: Reinforcement Learning for Selecting and Refining Learnable Strategic Archetypes in Quantitative TradingChuqiao Zong, Molei Qin, Haochong Xia, Bo AnAAAI 2026
- OPHR: Mastering Volatility Trading with Multi-Agent Deep Reinforcement LearningZeting Chen, Xinyu Cai, Molei Qin, Bo AnNeurIPS 2025 · 被引用 2 次
- A Multimodal Foundation Agent for Financial Trading: Tool-Augmented, Diversified, and GeneralistWentao Zhang, Lingxuan Zhao, Haochong Xia, Shuo Sun 等KDD 2024 · 被引用 50 次
