Wavelet Predictive Representations for Non-Stationary Reinforcement Learning
Min Wang, Xin Li, Ye He, Yao-Hui Li, Hasnaa Bennis, Riashat Islam, Mingzhong Wang
Abstract
The real world is inherently non-stationary, with ever-changing factors, such as weather conditions and traffic flows, making it challenging for agents to adapt to varying environmental dynamics. Non-Stationary Reinforcement Learning (NSRL) addresses this challenge by training agents to adapt rapidly to sequences of distinct Markov Decision Processes (MDPs). However, existing NSRL approaches often focus on tasks with regularly evolving patterns, leading to limited adaptability in highly dynamic settings. Inspired by the success of Wavelet analysis in time series modeling, specifically its ability to capture signal trends at multiple scales, we propose WISDOM to leverage wavelet-domain predictive task representations to enhance NSRL. WISDOM captures these multi-scale features in evolving MDP sequences by transforming task representation sequences into the wavelet domain, where wavelet coefficients represent both global trends and fine-grained variations of non-stationary changes. In addition to the auto-regressive modeling commonly employed in time series forecasting, we devise a wavelet temporal difference (TD) update operator to enhance tracking and prediction of MDP evolution. We theoretically prove the convergence of this operator and demonstrate policy improvement with wavelet task representations. Experiments on diverse benchmarks show that WISDOM significantly outperforms existing baselines in both sample efficiency and asymptotic performance, demonstrating its remarkable adaptability in complex environments characterized by non-stationary and stochastically evolving tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e4407fcd-9ee5-4e07-9c56-9ddcd7ded550Builds on13
- FEDformer: Frequency Enhanced Decomposed Transformer for Long-term Series ForecastingTian Zhou, Ziqing Ma, Qingsong Wen, Xue Wang et al.ICML 2022 · 2,912 citations
- WaveFill: A Wavelet-based Generation Network for Image InpaintingYingchen Yu, Fangneng Zhan, Shijian Lu, Jianxiong Pan et al.ICCV 2021 · 133 citations
- Look-ahead Meta Learning for Continual LearningGunshi Gupta, Karmesh Yadav, Liam PaullNeurIPS 2020 · 74 citations
- WHEN: A Wavelet-DTW Hybrid Attention Network for Heterogeneous Time Series AnalysisJingyuan Wang, Chen Yang, Xiaohan Jiang, Junjie WuKDD 2023 · 28 citations
- Sequence Modeling with Multiresolution Convolutional MemoryJiaxin Shi, Ke Alexander Wang, Emily B. FoxICML 2023 · 24 citations
Related papers
- Balancing Plasticity and Stability with Fast and Slow Successor FeaturesRaymond Chua, Doina Precup, Blake RichardsICML 2026
- Wavelet Policy: Lifting Scheme for Policy Learning in Long-Horizon TasksHao Huang, Shuaihang Yuan, Geeta Chandra Raju Bethala, Congcong Wen et al.ICCV 2025
- WaveForM: Graph Enhanced Wavelet Learning for Long Sequence Forecasting of Multivariate Time SeriesFuhao Yang, Xin Li, Min Wang, Hongyu Zang et al.AAAI 2023 · 35 citations
- Tackling Non-Stationarity in Reinforcement Learning via Causal-Origin RepresentationWanpeng Zhang, Yilin Li, Boyu Yang, Zongqing LuICML 2024 · 5 citations
- Robust Situational Reinforcement Learning in Face of Context DisturbancesJinpeng Zhang, Yufeng Zheng, Chuheng Zhang, Li Zhao et al.ICML 2023 · 5 citations
