SimSR: Simple Distance-Based State Representations for Deep Reinforcement Learning
Hongyu Zang, Xin Li, Mingzhong Wang
摘要
This work explores how to learn robust and generalizable state representation from image-based observations with deep reinforcement learning methods. Addressing the computational complexity, stringent assumptions and representation collapse challenges in existing work of bisimulation metric, we devise Simple State Representation (SimSR) operator. SimSR enables us to design a stochastic approximation method that can practically learn the mapping functions (encoders) from observations to latent representation space. In addition to the theoretical analysis and comparison with the existing work, we experimented and compared our work with recent state-of-the-art solutions in visual MuJoCo tasks. The results shows that our model generally achieves better performance and has better robustness and good generalization.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- What Effects the Generalization in Visual Reinforcement Learning: Policy Consistency with Truncated Return PredictionShuo Wang, Zhihao Wu, Xiaobo Hu, Jinwen Wang 等AAAI 2024 · 被引用 18 次
- Understanding and Addressing the Pitfalls of Bisimulation-based Representations in Offline Reinforcement LearningHongyu Zang, Xin Li, Leiji Zhang, Yang Liu 等NeurIPS 2023 · 被引用 15 次
- Reward-Aware Proto-Representations in Reinforcement LearningHon Tik Tse, Siddarth Chandrasekar, Marlos C. MachadoNeurIPS 2025 · 被引用 6 次
- State Chrono Representation for Enhancing Generalization in Reinforcement LearningJianda Chen, Wen Zheng Terence Ng, Zichen Chen, Sinno Jialin Pan 等NeurIPS 2024 · 被引用 5 次
- Behavior Prior Representation learning for Offline Reinforcement LearningHongyu Zang, Xin Li, Jie Yu, Chen Liu 等ICLR 2023 · 被引用 3 次
它引用的顶会 Paper11
- Understanding Contrastive Representation Learning through Alignment and Uniformity on the HypersphereTongzhou Wang, Phillip IsolaICML 2020 · 被引用 2,360 次
- CURL: Contrastive Unsupervised Representations for Reinforcement LearningMichael Laskin, Aravind Srinivas, Pieter AbbeelICML 2020 · 被引用 1,261 次
- Image Augmentation Is All You Need: Regularizing Deep Reinforcement Learning from PixelsDenis Yarats, Ilya Kostrikov, Rob FergusICLR 2021 · 被引用 911 次
- Stochastic Latent Actor-Critic: Deep Reinforcement Learning with a Latent Variable ModelAlex X. Lee, Anusha Nagabandi, Pieter Abbeel, Sergey LevineNeurIPS 2020 · 被引用 437 次
- Decoupling Representation Learning from Reinforcement LearningAdam Stooke, Kimin Lee, Pieter Abbeel, Michael LaskinICML 2021 · 被引用 389 次
相关 Paper
- Learning Invariant Representations for Reinforcement Learning without ReconstructionAmy Zhang, Rowan Thomas McAllister, Roberto Calandra, Yarin Gal 等ICLR 2021 · 被引用 77 次
- Towards Robust Bisimulation Metric LearningMete Kemertas, Tristan Aumentado-ArmstrongNeurIPS 2021 · 被引用 68 次
- Bisimulation Metric for Model Predictive ControlYutaka Shimizu, Masayoshi TomizukaICLR 2025
- Robust Representation Learning by Clustering with Bisimulation Metrics for Visual Reinforcement Learning with DistractionsQiyuan Liu, Qi Zhou, Rui Yang, Jie WangAAAI 2023 · 被引用 22 次
- Provably sample-efficient RL with side information about latent dynamicsYao Liu, Dipendra Misra, Miro Dudík, Robert E. SchapireNeurIPS 2022 · 被引用 2 次
