RD: Reward Decomposition with Representation Decomposition
Zichuan Lin, Derek Yang, Li Zhao, Tao Qin, Guangwen Yang, Tie-Yan Liu
Abstract
Reward decomposition, which aims to decompose the full reward into multiple sub-rewards, has been proven beneficial for improving sample efficiency in reinforcement learning. Existing works on discovering reward decomposition are mostly policy dependent, which constrains diversified or disentangled behavior between different policies induced by different sub-rewards. In this work, we propose a set of novel policy-independent reward decomposition principles by constraining uniqueness and compactness of different state representations relevant to different sub-rewards. Our principles encourage sub-rewards with minimal relevant features, while maintaining the uniqueness of each sub-reward. We derive a deep learning algorithm based on our principle, and refer to our method as RD 2 , since we learn reward decomposition and disentangled representation jointly. RD 2 is evaluated on a toy case, where we have the true reward structure, and chosen Atari environments where the reward structure exists but is unknown to the agent to demonstrate the effectiveness of RD 2 against existing reward decomposition methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1ba7f807-df77-4b70-93e3-67ca93a8c4ccCited by top-tier papers4
- Distributional Reinforcement Learning for Multi-Dimensional Reward FunctionsPushi Zhang, Xiaoyu Chen, Li Zhao, Wei Xiong et al.NeurIPS 2021 · 33 citations
- Dual-Objective Reinforcement Learning with Novel Hamilton-Jacobi-Bellman FormulationsWilliam Sharpless, Dylan Hirsch, Sander Tonkens, Nikhil Uday Shinde et al.ICLR 2026 · 12 citations
- Locality Matters: A Scalable Value Decomposition Approach for Cooperative Multi-Agent Reinforcement LearningRoy Zohar, Shie Mannor, Guy TennenholtzAAAI 2022 · 11 citations
- Distributional Reinforcement Learning with Regularized Wasserstein LossKe Sun, Yingnan Zhao, Wulong Liu, Bei Jiang et al.NeurIPS 2024 · 2 citations
Related papers
- Orchestrated Value Mapping for Reinforcement LearningMehdi Fatemi, Arash TavakoliICLR 2022 · 8 citations
- Catching Two Birds with One Stone: Reward Shaping with Dual Random Networks for Balancing Exploration and ExploitationHaozhe Ma, Fangling Li, Jing Yu Lim, Zhengding Luo et al.ICML 2025
- DeepSynth: Automata Synthesis for Automatic Task Segmentation in Deep Reinforcement LearningMohammadhosein Hasanbeig, Natasha Yogananda Jeppu, Alessandro Abate, Tom Melham et al.AAAI 2021 · 62 citations
- Deep PQR: Solving Inverse Reinforcement Learning using Anchor ActionsSinong Geng, Houssam Nassif, Carlos A. Manzanares, A. Max Reppen et al.ICML 2020 · 14 citations
- Learning World Models with Identifiable FactorizationYuren Liu, Biwei Huang, Zhengmao Zhu, Hong-Long Tian et al.NeurIPS 2023 · 32 citations
