Latent World Models For Intrinsically Motivated Exploration
Aleksandr Ermolov, Nicu Sebe
Abstract
In this work we consider partially observable environments with sparse rewards. We present a self-supervised representation learning method for image-based observations, which arranges embeddings respecting temporal distance of observations. This representation is empirically robust to stochasticity and suitable for novelty detection from the error of a predictive forward model. We consider episodic and life-long uncertainties to guide the exploration. We propose to estimate the missing information about the environment with the world model, which operates in the learned latent space. As a motivation of the method, we analyse the exploration problem in a tabular Partially Observable Labyrinth. We demonstrate the method on image-based hard exploration environments from the Atari benchmark and report significant improvement with respect to prior work. The source code of the method and all the experiments is available at https://github.com/htdt/lwm .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- Flipping Coins to Estimate Pseudocounts for Exploration in Reinforcement LearningSam Lobel, Akhil Bagaria, George KonidarisICML 2023 · 29 citations
- Incremental Reinforcement Learning with Dual-Adaptive ε-Greedy ExplorationWei Ding, Siyang Jiang, Hsi-Wen Chen, Ming-Syan ChenAAAI 2023 · 11 citations
- Time to augment self-supervised visual representation learningArthur Aubret, Markus Roland Ernst, Céline Teulière, Jochen TrieschICLR 2023 · 1 citation
- Leveraging Skills from Unlabeled Prior Data for Efficient Online ExplorationMax Wilcoxson, Qiyang Li, Kevin Frans, Sergey LevineICML 2025
- What You Think is What You See: Driving Exploration in VLM Agents via Visual-Linguistic CuriosityHaoxi Li, Qinglin Hou, Jianfei Ma, Jinxiang Lai et al.ICML 2026
Builds on3
- Image Augmentation Is All You Need: Regularizing Deep Reinforcement Learning from PixelsDenis Yarats, Ilya Kostrikov, Rob FergusICLR 2021 · 911 citations
- Planning to Explore via Self-Supervised World ModelsRamanan Sekar, Oleh Rybkin, Kostas Daniilidis, Pieter Abbeel et al.ICML 2020 · 489 citations
- Whitening for Self-Supervised Representation LearningAleksandr Ermolov, Aliaksandr Siarohin, Enver Sangineto, Nicu SebeICML 2021 · 378 citations
Related papers
- Bootstrap Latent-Predictive Representations for Multitask Reinforcement LearningZhaohan Daniel Guo, Bernardo Ávila Pires, Bilal Piot, Jean-Bastien Grill et al.ICML 2020 · 153 citations
- Novelty Search in Representational Space for Sample Efficient ExplorationRuo Yu Tao, Vincent François-Lavet, Joelle PineauNeurIPS 2020 · 53 citations
- Task-Aware Exploration via a Predictive Bisimulation MetricDayang Liang, Ruihan LIU, Lipeng Wan, Yunlong Liu et al.ICML 2026 · 1 citation
- BYOL-Explore: Exploration by Bootstrapped PredictionZhaohan Guo, Shantanu Thakoor, Miruna Pislar, Bernardo Ávila Pires et al.NeurIPS 2022 · 104 citations
- Go Beyond Imagination: Maximizing Episodic Reachability with World ModelsYao Fu, Run Peng, Honglak LeeICML 2023 · 1 citation
