A Unified Self-Regulating Training Framework for Federated Deep Reinforcement Learning
Meng Xu, Xinhong Chen, Zhongying Chen, Guanyi Zhao, Yang Jin, Jianping Wang
Abstract
Federated Deep Reinforcement Learning (FDRL) aims to enable distributed collaborative training of multiple DRL models while preserving privacy. Existing FDRL methods function in static client environments, but real-world scenarios often involve dynamic state transitions, such as noise, which render static model topologies inadequate and result in biased policy loss. This degrades client performance and leads to suboptimal global policies. To address this challenge, we develop a generic solution, referred to as the self-regulating training framework, which can be seamlessly integrated into existing FDRL approaches to address dynamic state transitions. Specifically, we propose a Sparse Training (ST) method that dynamically sparsifies and adjusts the topology of each model during training to maximize model performance and reduce model complexity. Additionally, we introduce an auxiliary model to adaptively regulate the policy loss of client models, mitigating loss bias and facilitating updates that yield improved returns. Experimental results demonstrate that our method enhances six state-of-the-art (SOTA) FDRL approaches across nine tasks in terms of return.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7d824fae-c65e-4ce0-9ba8-5aab7ad5a16cBuilds on8
- Context-aware Dynamics Model for Generalization in Model-Based Reinforcement LearningKimin Lee, Younggyo Seo, Seunghyun Lee, Honglak Lee et al.ICML 2020 · 158 citations
- Fault-Tolerant Federated Reinforcement Learning with Theoretical GuaranteeFlint Xiaofeng Fan, Yining Ma, Zhongxiang Dai, Wei Jing et al.NeurIPS 2021 · 102 citations
- State Regularized Policy Optimization on Data with Dynamics ShiftZhenghai Xue, Qingpeng Cai, Shuchang Liu, Dong Zheng et al.NeurIPS 2023 · 30 citations
- On Lottery Tickets and Minimal Task Representations in Deep Reinforcement LearningMarc Aurel Vischer, Robert Tjarko Lange, Henning SprekelerICLR 2022 · 27 citations
- Federated Ensemble-Directed Offline Reinforcement LearningDesik Rengarajan, Nitin Ragothaman, Dileep Kalathil, Srinivas ShakkottaiNeurIPS 2024 · 11 citations
Related papers
- Expected Returns and Policy Inconsistency-Aware Offline Federated Deep Reinforcement LearningMeng XU, Zhongying Chen, Weiwei Fu, Yan Li et al.ICML 2026
- Federated Dynamic Sparse Training: Computing Less, Communicating Less, Yet Learning BetterSameer Bibikar, Haris Vikalo, Zhangyang Wang, Xiaohan ChenAAAI 2022 · 133 citations
- Value-Based Deep Multi-Agent Reinforcement Learning with Dynamic Sparse TrainingPihe Hu, Shaolong Li, Zhuoran Li, Ling Pan et al.NeurIPS 2024 · 2 citations
- RLx2: Training a Sparse Deep Reinforcement Learning Model from ScratchYiqin Tan, Pihe Hu, Ling Pan, Jiatai Huang et al.ICLR 2023 · 1 citation
- On the Linear Speedup of Personalized Federated Reinforcement Learning with Shared RepresentationsGuojun Xiong, Shufan Wang, Daniel Jiang, Jian LiICLR 2025
