Knowledge Diversion for Efficient Morphology Control and Policy Transfer
Fu Feng, Ruixiao Shi, Yucheng Xie, Jianlu Shen, Jing Wang, Xin Geng
摘要
Universal morphology control aims to learn a universal policy that generalizes across heterogeneous robot morphologies, with Transformer-based controllers emerging as a dominant choice. However, such architectures incur substantial computational costs, resulting in high deployment overhead, and existing methods exhibit limited cross-task generalization, necessitating training from scratch for each new task. To this end, we propose DivMorph, a modular training paradigm that leverages knowledge diversion to learn decomposable controllers. DivMorph factorizes randomly initialized Transformer weights into basic knowledge units via SVD and employs dynamic soft gating, conditioned on task and morphology embeddings, to adaptively modulate these units into universal learngenes and morphology- and task-specific tailors during training, thereby achieving knowledge disentanglement. By selectively activating relevant components, DivMorph adaptively recomposes the controller, enabling efficient policy deployment and effective policy transfer to novel tasks. Extensive experiments demonstrate that DivMorph achieves state-of-the-art performance, improving sample efficiency for cross-task transfer by 3.3 and reducing model size for single-agent deployment by 16.7.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper13
- Mixture-of-Experts with Expert Choice RoutingYanqi Zhou, Tao Lei, Hanxiao Liu, Nan Du 等NeurIPS 2022 · 被引用 933 次
- Multi-Task Reinforcement Learning with Soft ModularizationRuihan Yang, Huazhe Xu, Yi Wu, Xiaolong WangNeurIPS 2020 · 被引用 247 次
- One Policy to Control Them All: Shared Modular Policies for Agent-Agnostic ControlWenlong Huang, Igor Mordatch, Deepak PathakICML 2020 · 被引用 214 次
- Orthogonalizing Convolutional Layers with the Cayley TransformAsher Trockman, J. Zico KolterICLR 2021 · 被引用 137 次
- MetaMorph: Learning Universal Controllers with TransformersAgrim Gupta, Linxi Fan, Surya Ganguli, Li Fei-FeiICLR 2022 · 被引用 130 次
相关 Paper
- Distilling Morphology-Conditioned Hypernetworks for Efficient Universal Morphology ControlZheng Xiong, Risto Vuorio, Jacob Beck, Matthieu Zimmer 等ICML 2024 · 被引用 8 次
- Universal Morphology Control via Contextual ModulationZheng Xiong, Jacob Beck, Shimon WhitesonICML 2023 · 被引用 27 次
- MeMo: Meaningful, Modular Controllers via Noise InjectionMegan Tjandrasuwita, Jie Xu, Armando Solar-Lezama, Wojciech MatusikNeurIPS 2024 · 被引用 1 次
- DivControl: Knowledge Diversion for Controllable Image GenerationYucheng Xie, Fu Feng, Ruixiao Shi, Jing Wang 等AAAI 2026 · 被引用 4 次
- Structure-Aware Transformer Policy for Inhomogeneous Multi-Task Reinforcement LearningSunghoon Hong, Deunsol Yoon, Kee-Eung KimICLR 2022 · 被引用 40 次
