Hyperbolic Deep Reinforcement Learning
Edoardo Cetin, Benjamin Paul Chamberlain, Michael M. Bronstein, Jonathan J. Hunt
摘要
We propose a new class of deep reinforcement learning (RL) algorithms that model latent representations in hyperbolic space. Sequential decision-making requires reasoning about the possible future consequences of current behavior. Consequently, capturing the relationship between key evolving features for a given task is conducive to recovering effective policies. To this end, hyperbolic geometry provides deep RL models with a natural basis to precisely encode this inherently hierarchical information. However, applying existing methodologies from the hyperbolic deep learning literature leads to fatal optimization instabilities due to the non-stationarity and variance characterizing RL gradient estimators. Hence, we design a new general method that counteracts such optimization challenges and enables stable end-to-end learning with deep hyperbolic representations. We empirically validate our framework by applying it to popular on-policy and offpolicy RL algorithms on the Procgen and Atari 100K benchmarks, attaining near universal performance and generalization benefits. Given its natural fit, we hope future RL research will consider hyperbolic representations as a standard tool. Project website: sites.google.com/view/hyperbolic-rl
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- EDGI: Equivariant Diffusion for Planning with Embodied AgentsJohann Brehmer, Joey Bose, Pim de Haan, Taco S. CohenNeurIPS 2023 · 被引用 52 次
- Hyperbolic Fine-Tuning for Large Language ModelsMenglin Yang, Ram Samarth B. B., Aosong Feng, Bo Xiong 等NeurIPS 2025 · 被引用 31 次
- Hyperbolic Diffusion Embedding and Distance for Hierarchical Representation LearningYa-Wei Eileen Lin, Ronald R. Coifman, Gal Mishne, Ronen TalmonICML 2023 · 被引用 26 次
- Hyperbolic VAE via Latent Gaussian DistributionsSeunghyuk Cho, Juyong Lee, Dongwoo KimNeurIPS 2023 · 被引用 16 次
- Riemannian SAM: Sharpness-Aware Minimization on Riemannian ManifoldsJihun Yun, Eunho YangNeurIPS 2023 · 被引用 10 次
它引用的顶会 Paper21
- CURL: Contrastive Unsupervised Representations for Reinforcement LearningMichael Laskin, Aravind Srinivas, Pieter AbbeelICML 2020 · 被引用 1,261 次
- Model Based Reinforcement Learning for AtariLukasz Kaiser, Mohammad Babaeizadeh, Piotr Milos, Blazej Osinski 等ICLR 2020 · 被引用 969 次
- Image Augmentation Is All You Need: Regularizing Deep Reinforcement Learning from PixelsDenis Yarats, Ilya Kostrikov, Rob FergusICLR 2021 · 被引用 911 次
- Reinforcement Learning with Augmented DataMichael Laskin, Kimin Lee, Adam Stooke, Lerrel Pinto 等NeurIPS 2020 · 被引用 833 次
- Hyperbolic Neural Networks++Ryohei Shimizu, Yusuke Mukuta, Tatsuya HaradaICLR 2021 · 被引用 791 次
相关 Paper
- Understanding and Improving Hyperbolic Deep Reinforcement LearningTimo Klein, Thomas Lang, Andrii Shkabrii, Alexander Sturm 等ICLR 2026 · 被引用 1 次
- Exploiting Geometric Structures for Modeling Multi-Agent Behaviors: A New ThinkingBohao Qu, Xiaofeng Cao, Bing Li, Menglin Zhang 等AAAI 2026
- HELM: Hyperbolic Large Language Models via Mixture-of-Curvature ExpertsNeil He, Rishabh Anand, Hiren Madhu, Ali Maatouk 等NeurIPS 2025 · 被引用 27 次
- Learning the Predictability of the FutureDidac Suris, Ruoshi Liu, Carl VondrickCVPR 2021
- Stochastic Latent Actor-Critic: Deep Reinforcement Learning with a Latent Variable ModelAlex X. Lee, Anusha Nagabandi, Pieter Abbeel, Sergey LevineNeurIPS 2020 · 被引用 437 次
