Towards Adaptive Humanoid Control via Multi-Behavior Distillation and Reinforced Fine-Tuning
Yingnan Zhao, Xinmiao Wang, Dewei Wang, Xinzhe Liu, Dan Lu, Qilong Han, Peng Liu, Chenjia Bai
摘要
Humanoid robots are promising to learn a diverse set of human-like locomotion behaviors, including standing up, walking, running, and jumping. However, existing methods predominantly require training independent policies for each skill, yielding behavior-specific controllers that exhibit limited generalization and brittle performance when deployed on irregular terrains and in diverse situations. To address this challenge, we propose Adaptive Humanoid Control (AHC) that adopts a two-stage framework to learn an adaptive humanoid locomotion controller across different skills and terrains. Specifically, we first train several primary locomotion policies and perform a multi-behavior distillation process to obtain a basic multi-behavior controller, facilitating adaptive behavior switching based on the environment. Then, we perform reinforced fine-tuning by collecting online feedback in performing adaptive behaviors on more diverse terrains, enhancing terrain adaptability for the controller. We conduct experiments in both simulation and real-world experiments in Unitree G1 robots. The results show that our method exhibits strong adaptability across various situations and terrains. Website - https://ahc-humanoid.github.io
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper7
- Conflict-Averse Gradient Descent for Multi-task learningBo Liu, Xingchao Liu, Xiaojie Jin, Peter Stone 等NeurIPS 2021 · 被引用 686 次
- AMP: adversarial motion priors for stylized physics-based character controlXue Bin Peng, Ze Ma, Pieter Abbeel, Sergey Levine 等SIGGRAPH 2021 · 被引用 392 次
- Implementation Matters in Deep RL: A Case Study on PPO and TRPOLogan Engstrom, Andrew Ilyas, Shibani Santurkar, Dimitris Tsipras 等ICLR 2020 · 被引用 305 次
- Multi-Task Reinforcement Learning with Context-based RepresentationsShagun Sodhani, Amy Zhang, Joelle PineauICML 2021 · 被引用 241 次
- Multi-Critic Actor Learning: Teaching RL Policies to Act with StyleSiddharth Mysore, George Cheng, Yunqi Zhao, Kate Saenko 等ICLR 2022 · 被引用 40 次
相关 Paper
- KungfuBot: Physics-Based Humanoid Whole-Body Control for Learning Highly-Dynamic SkillsWeiji Xie, Jinrui Han, Jiakun Zheng, Huanyu Li 等NeurIPS 2025 · 被引用 120 次
- Keep On Going: Learning Robust Humanoid Motion Skills via Selective Adversarial TrainingYang Zhang, Zhanxiang Cao, Buqing Nie, Haoyang Li 等AAAI 2026 · 被引用 4 次
- Adversarial Locomotion and Motion Imitation for Humanoid Policy LearningJiyuan Shi, Xinzhe Liu, Dewei Wang, Ouyang Lu 等NeurIPS 2025 · 被引用 30 次
- From Experts to a Generalist: Toward General Whole-Body Control for Humanoid RobotsYuxuan Wang, Ming Yang, Gang Ding, Yu Zhang 等NeurIPS 2025 · 被引用 36 次
- HWC-Loco: A Hierarchical Whole-Body Control Approach to Robust Humanoid LocomotionSixu Lin, Guanren Qiao, Yunxin Tai, Ang Li 等ICLR 2026 · 被引用 7 次
