EASI: Evolutionary Adversarial Simulator Identification for Sim-to-Real Transfer
Haoyu Dong, Huiqiao Fu, Wentao Xu, Zhehao Zhou, Chunlin Chen
摘要
Reinforcement Learning (RL) controllers have demonstrated remarkable performance in complex robot control tasks. However, the presence of reality gap often leads to poor performance when deploying policies trained in simulation directly onto real robots. Previous sim-to-real algorithms like Domain Randomization (DR) requires domain-specific expertise and suffers from issues such as reduced control performance and high training costs. In this work, we introduce Evolutionary Adversarial Simulator Identification (EASI), a novel approach that combines Generative Adversarial Network (GAN) and Evolutionary Strategy (ES) to address sim-to-real challenges. Specifically, we consider the problem of sim-to-real as a search problem, where ES acts as a generator in adversarial competition with a neural network discriminator, aiming to find physical parameter distributions that make the state transitions between simulation and reality as similar as possible. The discriminator serves as the fitness function, guiding the evolution of the physical parameter distributions. EASI features simplicity, low cost, and high fidelity, enabling the construction of a more realistic simulator with minimal requirements for real-world data, thus aiding in transferring simulated-trained policies to the real world. We demonstrate the performance of EASI in both sim-to-sim and sim-to-real tasks, showing superior performance compared to existing sim-to-real algorithms.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper4
- AMP: adversarial motion priors for stylized physics-based character controlXue Bin Peng, Ze Ma, Pieter Abbeel, Sergey Levine 等SIGGRAPH 2021 · 被引用 392 次
- An Imitation from Observation Approach to Transfer Learning with Dynamics MismatchSiddharth Desai, Ishan Durugkar, Haresh Karnan, Garrett Warnell 等NeurIPS 2020 · 被引用 60 次
- ASID: Active Exploration for System Identification in Robotic ManipulationMarius Memmel, Andrew Wagenmaker, Chuning Zhu, Dieter Fox 等ICLR 2024 · 被引用 36 次
- Ess-InfoGAIL: Semi-supervised Imitation Learning from Imbalanced DemonstrationsHuiqiao Fu, Kaiqiang Tang, Yuanyang Lu, Yiming Qi 等NeurIPS 2023 · 被引用 15 次
相关 Paper
- Domain Randomization via Entropy MaximizationGabriele Tiboni, Pascal Klink, Jan Peters, Tatiana Tommasi 等ICLR 2024 · 被引用 24 次
- SPiDR: A Simple Approach for Zero-Shot Safety in Sim-to-Real TransferYarden As, Chengrui Qu, Benjamin Unger, Dongho Kang 等NeurIPS 2025 · 被引用 9 次
- Revisiting Domain Randomization via Relaxed State-Adversarial Policy OptimizationYun-Hsuan Lien, Ping-Chun Hsieh, Yu-Shuen WangICML 2023 · 被引用 1 次
- Understanding Domain Randomization for Sim-to-real TransferXiaoyu Chen, Jiachen Hu, Chi Jin, Lihong Li 等ICLR 2022 · 被引用 164 次
- UserSim: User Simulation via Supervised GenerativeAdversarial NetworkXiangyu Zhao, Long Xia, Lixin Zou, Hui Liu 等WWW 2021 · 被引用 31 次
