Active World Model Learning with Progress Curiosity
Kuno Kim, Megumi Sano, Julian De Freitas, Nick Haber, Daniel Yamins
Abstract
World models are self-supervised predictive models of how the world evolves. Humans learn world models by curiously exploring their environment, in the process acquiring compact abstractions of high bandwidth sensory inputs, the ability to plan across long temporal horizons, and an understanding of the behavioral patterns of other agents. In this work, we study how to design such a curiosity-driven Active World Model Learning (AWML) system. To do so, we construct a curious agent building world models while visually exploring a 3D physical environment rich with distillations of representative real-world agents. We propose an AWML system driven by -Progress: a scalable and effective learning progress-based curiosity signal. We show that -Progress naturally gives rise to an exploration policy that directs attention to complex but learnable dynamics in a balanced manner, thus overcoming the "white noise problem". As a result, our -Progress-driven controller achieves significantly higher AWML performance than baseline controllers equipped with state-of-the-art exploration strategies such as Random Network Distillation and Model Disagreement.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers12
- Dreamwalker: Mental Planning for Continuous Vision-Language NavigationHanqing Wang, Wei Liang, Luc Van Gool, Wenguan WangICCV 2023 · 98 citations
- Curious Exploration via Structured World Models Yields Zero-Shot Object ManipulationCansu Sancaktar, Sebastian Blaes, Georg MartiusNeurIPS 2022 · 43 citations
- ELIGN: Expectation Alignment as a Multi-Agent Intrinsic RewardZixian Ma, Rose E. Wang, Fei-Fei Li, Michael S. Bernstein et al.NeurIPS 2022 · 22 citations
- Curious Replay for Model-based AdaptationIsaac Kauvar, Chris Doyle, Linqi Zhou, Nick HaberICML 2023 · 18 citations
- SMiRL: Surprise Minimizing Reinforcement Learning in Unstable EnvironmentsGlen Berseth, Daniel Geng, Coline Manon Devin, Nicholas Rhinehart et al.ICLR 2021 · 12 citations
Related papers
- What You Think is What You See: Driving Exploration in VLM Agents via Visual-Linguistic CuriosityHaoxi Li, Qinglin Hou, Jianfei Ma, Jinxiang Lai et al.ICML 2026
- BYOL-Explore: Exploration by Bootstrapped PredictionZhaohan Guo, Shantanu Thakoor, Miruna Pislar, Bernardo Ávila Pires et al.NeurIPS 2022 · 104 citations
- Learning to Model the World With LanguageJessy Lin, Yuqing Du, Olivia Watkins, Danijar Hafner et al.ICML 2024 · 76 citations
- Perceiving the Knowledge Boundary: Uncertainty-Guided Exploration and Imagination for World ModelsZhenxian Liu, Peixi Peng, Yangru Huang, Yonghong TianAAAI 2026
- Beyond Single-Speed Reasoning: Coordinating Fast and Slow Dynamics for Efficient World ModelingHongwei Wang, Yangru Huang, Guangyao Chen, Xu Wang et al.AAAI 2026
