Towards Better Laplacian Representation in Reinforcement Learning with Generalized Graph Drawing
Kaixin Wang, Kuangqi Zhou, Qixin Zhang, Jie Shao, Bryan Hooi, Jiashi Feng
摘要
The Laplacian representation recently gains increasing attention for reinforcement learning as it provides succinct and informative representation for states, by taking the eigenvectors of the Laplacian matrix of the state-transition graph as state embeddings. Such representation captures the geometry of the underlying state space and is beneficial to RL tasks such as option discovery and reward shaping. To approximate the Laplacian representation in large (or even continuous) state spaces, recent works propose to minimize a spectral graph drawing objective, which however has infinitely many global minimizers other than the eigenvectors. As a result, their learned Laplacian representation may differ from the ground truth. To solve this problem, we reformulate the graph drawing objective into a generalized form and derive a new learning objective, which is proved to have eigenvectors as its unique global minimizer. It enables learning high-quality Laplacian representations that faithfully approximate the ground truth. We validate this via comprehensive experiments on a set of gridworld and continuous control environments. Moreover, we show that our learned Laplacian representations lead to more exploratory options and better reward shaping.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper19
- Learning from Reward-Free Offline Data: A Case for Planning with Latent Dynamics ModelsUladzislau Sobal, Wancong Zhang, Kyunghyun Cho, Randall Balestriero 等NeurIPS 2025 · 被引用 109 次
- Deep Laplacian-based Options for Temporally-Extended ExplorationMartin Klissarov, Marlos C. MachadoICML 2023 · 被引用 31 次
- Scalable Multi-agent Covering Option Discovery based on Kronecker GraphsJiayu Chen, Jingdi Chen, Tian Lan, Vaneet AggarwalNeurIPS 2022 · 被引用 16 次
- Reachability-Aware Laplacian Representation in Reinforcement LearningKaixin Wang, Kuangqi Zhou, Jiashi Feng, Bryan Hooi 等ICML 2023 · 被引用 10 次
- A Unified Algorithm Framework for Unsupervised Discovery of Skills based on Determinantal Point ProcessJiayu Chen, Vaneet Aggarwal, Tian LanNeurIPS 2023 · 被引用 8 次
它引用的顶会 Paper4
- Decoupling Representation Learning from Reinforcement LearningAdam Stooke, Kimin Lee, Pieter Abbeel, Michael LaskinICML 2021 · 被引用 389 次
- Count-Based Exploration with the Successor RepresentationMarlos C. Machado, Marc G. Bellemare, Michael BowlingAAAI 2020 · 被引用 206 次
- Exploration in Reinforcement Learning with Deep Covering OptionsYuu Jinnai, Jee Won Park, Marlos C. Machado, George Dimitri KonidarisICLR 2020 · 被引用 64 次
- Contrastive Behavioral Similarity Embeddings for Generalization in Reinforcement LearningRishabh Agarwal, Marlos C. Machado, Pablo Samuel Castro, Marc G. BellemareICLR 2021 · 被引用 27 次
相关 Paper
- Proper Laplacian Representation LearningDiego Gomez, Michael Bowling, Marlos C. MachadoICLR 2024 · 被引用 11 次
- Online Laplacian-Based Representation Learning in Reinforcement LearningMaheed H. Ahmed, Jayanth Bhargav, Mahsa GhasemiICML 2025
- Impact of Connectivity on Laplacian Representations in Reinforcement LearningTommaso Giorgi, Pierriccardo Olivieri, Keyue Jiang, Laura Toni 等ICML 2026 · 被引用 1 次
- Novel Exploration via OrthogonalityAndreas Theophilou, Özgür SimsekNeurIPS 2025 · 被引用 1 次
- Option Discovery in the Absence of Rewards with Manifold AnalysisAmitay Bar, Ronen Talmon, Ron MeirICML 2020 · 被引用 6 次
