Locality Sensitive Sparse Encoding for Learning World Models Online
Zichen Liu, Chao Du, Wee Sun Lee, Min Lin
Abstract
Acquiring an accurate world model online for model-based reinforcement learning (MBRL) is challenging due to data nonstationarity, which typically causes catastrophic forgetting for neural networks (NNs). From the online learning perspective, a Follow-The-Leader (FTL) world model is desirable, which optimally fits all previous experiences at each round. Unfortunately, NN-based models need re-training on all accumulated data at every interaction step to achieve FTL, which is computationally expensive for lifelong agents. In this paper, we revisit models that can achieve FTL with incremental updates. Specifically, our world model is a linear regression model supported by nonlinear random features. The linear part ensures efficient FTL update while the nonlinear random feature empowers the fitting of complex environments. To best trade off model capacity and computation efficiency, we introduce a locality sensitive sparse encoding, which allows us to conduct efficient sparse updates even with very high dimensional nonlinear features. We validate the representation power of our encoding and verify that it allows efficient online learning under data covariate shift. We also show, in the Dyna MBRL setting, that our world models learned online using a single pass of trajectory data either surpass or match the performance of deep world models trained with replay and other continual learning methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b65a775b-310c-43fb-b8ea-c1394890ca39Cited by top-tier papers6
- GACL: Exemplar-Free Generalized Analytic Continual LearningHuiping Zhuang, Yizhu Chen, Di Fang, Run He et al.NeurIPS 2024 · 36 citations
- F-OAL: Forward-only Online Analytic Learning with Fast Training and Low Memory Footprint in Class Incremental LearningHuiping Zhuang, Yuchen Liu, Run He, Kai Tong et al.NeurIPS 2024 · 17 citations
- Continual Reinforcement Learning by Planning with Online World ModelsZichen Liu, Guoji Fu, Chao Du, Wee Sun Lee et al.ICML 2025
- Knowledge Retention in Continual Model-Based Reinforcement LearningHaotian Fu, Yixiang Sun, Michael Littman, George KonidarisICML 2025
- AFL: A Single-Round Analytic Approach for Federated Learning with Pre-trained ModelsRun He, Kai Tong, Di Fang, Han Sun et al.CVPR 2025
Builds on8
- MOPO: Model-based Offline Policy OptimizationTianhe Yu, Garrett Thomas, Lantao Yu, Stefano Ermon et al.NeurIPS 2020 · 989 citations
- Model Based Reinforcement Learning for AtariLukasz Kaiser, Mohammad Babaeizadeh, Piotr Milos, Blazej Osinski et al.ICLR 2020 · 969 citations
- Temporal Difference Learning for Model Predictive ControlNicklas Hansen, Hao Su, Xiaolong WangICML 2022 · 388 citations
- ACIL: Analytic Class-Incremental Learning with Absolute Memorization and Privacy ProtectionHuiping Zhuang, Zhenyu Weng, Hongxin Wei, Renchunzi Xie et al.NeurIPS 2022 · 106 citations
- Offline Reinforcement Learning with Reverse Model-based ImaginationJianhao Wang, Wenzhe Li, Haozhe Jiang, Guangxiang Zhu et al.NeurIPS 2021 · 74 citations
Related papers
- Deep Reinforcement Learning amidst Continual Structured Non-StationarityAnnie Xie, James Harrison, Chelsea FinnICML 2021 · 43 citations
- When Online Learning Meets ODE: Learning without Forgetting on Variable Feature SpaceDiyang Li, Bin GuAAAI 2023 · 4 citations
- Behavior-Invariant Task Representation Learning with Transformer-based World Models for Offline Meta-Reinforcement LearningFuyuan Qian, Menglong Zhang, Song Wang, Quanying LiuICML 2026
- Kalman Filter for Online Classification of Non-Stationary DataMichalis K. Titsias, Alexandre Galashov, Amal Rannen-Triki, Razvan Pascanu et al.ICLR 2024 · 14 citations
- Learning Successor Features the Simple WayRaymond Chua, Arna Ghosh, Christos Kaplanis, Blake A. Richards et al.NeurIPS 2024 · 14 citations
