Learning Hierarchical World Models with Adaptive Temporal Abstractions from Discrete Latent Dynamics
Christian Gumbsch, Noor Sajid, Georg Martius, Martin V. Butz
摘要
Hierarchical world models can significantly improve model-based reinforcement learning (MBRL) and planning by enabling reasoning across multiple time scales. Nonetheless, the majority of state-of-the-art MBRL methods employ flat, nonhierarchical models. We propose Temporal Hierarchies from Invariant Context Kernels (THICK), an algorithm that learns a world model hierarchy via discrete latent dynamics. The lower level of THICK updates parts of its latent state sparsely in time, forming invariant contexts. The higher level exclusively predicts situations involving context changes. Our experiments demonstrate that THICK learns categorical, interpretable, temporal abstractions on the high level, while maintaining precise low-level predictions. Furthermore, we show that the emergent hierarchical predictive model seamlessly enhances the abilities of MBRL or planning methods. We believe that THICK contributes to the further development of hierarchical agents capable of more sophisticated planning and reasoning abilities. i t a t Level 2 ît+τ ît+1 Level 1 (a) (b)
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- Dream to Manipulate: Compositional World Models Empowering Robot Imitation Learning with ImaginationLeonardo Barcellona, Andrii Zadaianchuk, Davide Allegro, Samuele Papa 等ICLR 2025 · 被引用 2 次
- DyMoDreamer: World Modeling with Dynamic ModulationBoxuan Zhang, Runqing Wang, Wei Xiao, Weipu Zhang 等NeurIPS 2025 · 被引用 2 次
- Hierarchical World Models as Visual Whole-Body Humanoid ControllersNicklas Hansen, Jyothir S. V, Vlad Sobal, Yann LeCun 等ICLR 2025 · 被引用 1 次
- Ego-Foresight: Self-supervised Learning of Agent-Aware Representations for Improved RLManuel Serra Nunes, Atabak Dehban, Yiannis Demiris, José Santos-VictorICLR 2026
- Beyond Single-Speed Reasoning: Coordinating Fast and Slow Dynamics for Efficient World ModelingHongwei Wang, Yangru Huang, Guangyao Chen, Xu Wang 等AAAI 2026
它引用的顶会 Paper18
- Dream to Control: Learning Behaviors by Latent ImaginationDanijar Hafner, Timothy P. Lillicrap, Jimmy Ba, Mohammad NorouziICLR 2020 · 被引用 1,852 次
- Mastering Atari with Discrete World ModelsDanijar Hafner, Timothy P. Lillicrap, Mohammad Norouzi, Jimmy BaICLR 2021 · 被引用 1,170 次
- Planning to Explore via Self-Supervised World ModelsRamanan Sekar, Oleh Rybkin, Kostas Daniilidis, Pieter Abbeel 等ICML 2020 · 被引用 489 次
- Skew-Fit: State-Covering Self-Supervised Reinforcement LearningVitchyr Pong, Murtaza Dalal, Steven Lin, Ashvin Nair 等ICML 2020 · 被引用 303 次
- The NetHack Learning EnvironmentHeinrich Küttler, Nantas Nardelli, Alexander H. Miller, Roberta Raileanu 等NeurIPS 2020 · 被引用 251 次
相关 Paper
- Hieros: Hierarchical Imagination on Structured State Space Sequence World ModelsPaul Mattes, Rainer Schlosser, Ralf HerbrichICML 2024 · 被引用 8 次
- Learning Multi-Timescale Abstractions for Hierarchical Combinatorial PlanningVivienne Huiling Wang, Tinghuai Wang, Joni PajarinenICML 2026
- Planning with Abstract Learned Models While Learning Transferable SubtasksJohn Winder, Stephanie Milani, Matthew Landen, Erebus Oh 等AAAI 2020 · 被引用 10 次
- Deep Hierarchical Planning from PixelsDanijar Hafner, Kuang-Huei Lee, Ian Fischer, Pieter AbbeelNeurIPS 2022 · 被引用 153 次
- Learning transferable motor skills with hierarchical latent mixture policiesDushyant Rao, Fereshteh Sadeghi, Leonard Hasenclever, Markus Wulfmeier 等ICLR 2022 · 被引用 34 次
