Towards Adaptive Humanoid Control via Multi-Behavior Distillation and Reinforced Fine-Tuning
Yingnan Zhao, Xinmiao Wang, Dewei Wang, Xinzhe Liu, Dan Lu, Qilong Han, Peng Liu, Chenjia Bai
Abstract
Humanoid robots are promising to learn a diverse set of human-like locomotion behaviors, including standing up, walking, running, and jumping. However, existing methods predominantly require training independent policies for each skill, yielding behavior-specific controllers that exhibit limited generalization and brittle performance when deployed on irregular terrains and in diverse situations. To address this challenge, we propose Adaptive Humanoid Control (AHC) that adopts a two-stage framework to learn an adaptive humanoid locomotion controller across different skills and terrains. Specifically, we first train several primary locomotion policies and perform a multi-behavior distillation process to obtain a basic multi-behavior controller, facilitating adaptive behavior switching based on the environment. Then, we perform reinforced fine-tuning by collecting online feedback in performing adaptive behaviors on more diverse terrains, enhancing terrain adaptability for the controller. We conduct experiments in both simulation and real-world experiments in Unitree G1 robots. The results show that our method exhibits strong adaptability across various situations and terrains. Website - https://ahc-humanoid.github.io
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 50ae1e7f-b39a-4260-ad7e-8e0690524c7eBuilds on7
- Conflict-Averse Gradient Descent for Multi-task learningBo Liu, Xingchao Liu, Xiaojie Jin, Peter Stone et al.NeurIPS 2021 · 686 citations
- AMP: adversarial motion priors for stylized physics-based character controlXue Bin Peng, Ze Ma, Pieter Abbeel, Sergey Levine et al.SIGGRAPH 2021 · 392 citations
- Implementation Matters in Deep RL: A Case Study on PPO and TRPOLogan Engstrom, Andrew Ilyas, Shibani Santurkar, Dimitris Tsipras et al.ICLR 2020 · 305 citations
- Multi-Task Reinforcement Learning with Context-based RepresentationsShagun Sodhani, Amy Zhang, Joelle PineauICML 2021 · 241 citations
- Multi-Critic Actor Learning: Teaching RL Policies to Act with StyleSiddharth Mysore, George Cheng, Yunqi Zhao, Kate Saenko et al.ICLR 2022 · 40 citations
Related papers
- KungfuBot: Physics-Based Humanoid Whole-Body Control for Learning Highly-Dynamic SkillsWeiji Xie, Jinrui Han, Jiakun Zheng, Huanyu Li et al.NeurIPS 2025 · 120 citations
- Keep On Going: Learning Robust Humanoid Motion Skills via Selective Adversarial TrainingYang Zhang, Zhanxiang Cao, Buqing Nie, Haoyang Li et al.AAAI 2026 · 4 citations
- Adversarial Locomotion and Motion Imitation for Humanoid Policy LearningJiyuan Shi, Xinzhe Liu, Dewei Wang, Ouyang Lu et al.NeurIPS 2025 · 30 citations
- From Experts to a Generalist: Toward General Whole-Body Control for Humanoid RobotsYuxuan Wang, Ming Yang, Gang Ding, Yu Zhang et al.NeurIPS 2025 · 36 citations
- HWC-Loco: A Hierarchical Whole-Body Control Approach to Robust Humanoid LocomotionSixu Lin, Guanren Qiao, Yunxin Tai, Ang Li et al.ICLR 2026 · 7 citations
