Novelty Detection in Reinforcement Learning with World Models
Geigh Zollicoffer, Kenneth Eaton, Jonathan C. Balloch, Julia M. Kim, Wei Zhou, Robert Wright, Mark O. Riedl
摘要
Reinforcement learning (RL) using world models has found significant recent successes. However, when a sudden change to world mechanics or properties occurs then agent performance and reliability can dramatically decline. We refer to the sudden change in visual properties or state transitions as novelties. Implementing novelty detection within generated world model frameworks is a crucial task for protecting the agent when deployed. In this paper, we propose straightforward bounding approaches to incorporate novelty detection into world model RL agents by utilizing the misalignment of the world model's hallucinated states and the true observed states as a novelty score. We provide effective approaches to detecting novelties in a distribution of transitions learned by an agent in a world model. Finally, we show the advantage of our work in Mini-Grid, Atari, and DeepMind Control environments compared to traditional machine learning novelty detection methods as well as currently accepted RL-focused novelty detection algorithms. While RL agents are often trained and evaluated in environments with stationary transition functions, the real world
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper10
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Mastering Atari with Discrete World ModelsDanijar Hafner, Timothy P. Lillicrap, Mohammad Norouzi, Jimmy BaICLR 2021 · 被引用 1,170 次
- Leveraging Procedural Generation to Benchmark Reinforcement LearningKarl Cobbe, Christopher Hesse, Jacob Hilton, John SchulmanICML 2020 · 被引用 685 次
- Diffusion for World Modeling: Visual Details Matter in AtariEloi Alonso, Adam Jelley, Vincent Micheli, Anssi Kanervisto 等NeurIPS 2024 · 被引用 359 次
- Is Out-of-Distribution Detection Learnable?Zhen Fang, Yixuan Li, Jie Lu, Jiahua Dong 等NeurIPS 2022 · 被引用 188 次
相关 Paper
- Perceiving the Knowledge Boundary: Uncertainty-Guided Exploration and Imagination for World ModelsZhenxian Liu, Peixi Peng, Yangru Huang, Yonghong TianAAAI 2026
- From Word to World: Can Large Language Models be Implicit Text-based World Models?Yixia Li, Hongru Wang, Jiahao Qiu, Zhenfei Yin 等ACL 2026 · 被引用 27 次
- Guaranteeing Out-Of-Distribution Detection in Deep RL via Transition EstimationMohit Prashant, Arvind Easwaran, Suman Das, Michael YuhasAAAI 2025 · 被引用 6 次
- Latent World Models For Intrinsically Motivated ExplorationAleksandr Ermolov, Nicu SebeNeurIPS 2020 · 被引用 26 次
- AdaWorld: Learning Adaptable World Models with Latent ActionsShenyuan Gao, Siyuan Zhou, Yilun Du, Jun Zhang 等ICML 2025
