Locality Sensitive Sparse Encoding for Learning World Models Online
Zichen Liu, Chao Du, Wee Sun Lee, Min Lin
摘要
Acquiring an accurate world model online for model-based reinforcement learning (MBRL) is challenging due to data nonstationarity, which typically causes catastrophic forgetting for neural networks (NNs). From the online learning perspective, a Follow-The-Leader (FTL) world model is desirable, which optimally fits all previous experiences at each round. Unfortunately, NN-based models need re-training on all accumulated data at every interaction step to achieve FTL, which is computationally expensive for lifelong agents. In this paper, we revisit models that can achieve FTL with incremental updates. Specifically, our world model is a linear regression model supported by nonlinear random features. The linear part ensures efficient FTL update while the nonlinear random feature empowers the fitting of complex environments. To best trade off model capacity and computation efficiency, we introduce a locality sensitive sparse encoding, which allows us to conduct efficient sparse updates even with very high dimensional nonlinear features. We validate the representation power of our encoding and verify that it allows efficient online learning under data covariate shift. We also show, in the Dyna MBRL setting, that our world models learned online using a single pass of trajectory data either surpass or match the performance of deep world models trained with replay and other continual learning methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- GACL: Exemplar-Free Generalized Analytic Continual LearningHuiping Zhuang, Yizhu Chen, Di Fang, Run He 等NeurIPS 2024 · 被引用 36 次
- F-OAL: Forward-only Online Analytic Learning with Fast Training and Low Memory Footprint in Class Incremental LearningHuiping Zhuang, Yuchen Liu, Run He, Kai Tong 等NeurIPS 2024 · 被引用 17 次
- Continual Reinforcement Learning by Planning with Online World ModelsZichen Liu, Guoji Fu, Chao Du, Wee Sun Lee 等ICML 2025
- Knowledge Retention in Continual Model-Based Reinforcement LearningHaotian Fu, Yixiang Sun, Michael Littman, George KonidarisICML 2025
- AFL: A Single-Round Analytic Approach for Federated Learning with Pre-trained ModelsRun He, Kai Tong, Di Fang, Han Sun 等CVPR 2025
它引用的顶会 Paper8
- MOPO: Model-based Offline Policy OptimizationTianhe Yu, Garrett Thomas, Lantao Yu, Stefano Ermon 等NeurIPS 2020 · 被引用 989 次
- Model Based Reinforcement Learning for AtariLukasz Kaiser, Mohammad Babaeizadeh, Piotr Milos, Blazej Osinski 等ICLR 2020 · 被引用 969 次
- Temporal Difference Learning for Model Predictive ControlNicklas Hansen, Hao Su, Xiaolong WangICML 2022 · 被引用 388 次
- ACIL: Analytic Class-Incremental Learning with Absolute Memorization and Privacy ProtectionHuiping Zhuang, Zhenyu Weng, Hongxin Wei, Renchunzi Xie 等NeurIPS 2022 · 被引用 106 次
- Offline Reinforcement Learning with Reverse Model-based ImaginationJianhao Wang, Wenzhe Li, Haozhe Jiang, Guangxiang Zhu 等NeurIPS 2021 · 被引用 74 次
相关 Paper
- Deep Reinforcement Learning amidst Continual Structured Non-StationarityAnnie Xie, James Harrison, Chelsea FinnICML 2021 · 被引用 43 次
- When Online Learning Meets ODE: Learning without Forgetting on Variable Feature SpaceDiyang Li, Bin GuAAAI 2023 · 被引用 4 次
- Behavior-Invariant Task Representation Learning with Transformer-based World Models for Offline Meta-Reinforcement LearningFuyuan Qian, Menglong Zhang, Song Wang, Quanying LiuICML 2026
- Kalman Filter for Online Classification of Non-Stationary DataMichalis K. Titsias, Alexandre Galashov, Amal Rannen-Triki, Razvan Pascanu 等ICLR 2024 · 被引用 14 次
- Learning Successor Features the Simple WayRaymond Chua, Arna Ghosh, Christos Kaplanis, Blake A. Richards 等NeurIPS 2024 · 被引用 14 次
