Cross-modal Domain Adaptation for Cost-Efficient Visual Reinforcement Learning
Xiong-Hui Chen, Shengyi Jiang, Feng Xu, Zongzhang Zhang, Yang Yu
摘要
In visual-input sim-to-real scenarios, to overcome the reality gap between images rendered in simulators and those from the real world, domain adaptation, i.e., learning an aligned representation space between simulators and the real world, then training and deploying policies in the aligned representation, is a promising direction. Previous methods focus on same-modal domain adaptation. However, those methods require building and running simulators that render high-quality images, which can be difficult and costly. In this paper, we consider a more costefficient setting of visual-input sim-to-real where only low-dimensional states are simulated. We first point out that the objective of learning mapping functions in previous methods that align the representation spaces is ill-posed, prone to yield an incorrect mapping. When the mapping crosses modalities, previous methods are easier to fail. Our algorithm, Cross-mOdal Domain Adaptation with Sequential structure (CODAS), mitigates the ill-posedness by utilizing the sequential nature of the data sampling process in RL tasks. Experiments on MuJoCo and Hand Manipulation Suite tasks show that the agents deployed with our method achieve similar performance as it has in the source domain, while those deployed with previous methods designed for same-modal domain adaptation suffer a larger performance gap.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Cooperative and Adversarial Learning: Co-enhancing Discriminability and Transferability in Domain AdaptationHui Sun, Zheng Xie, Xin-Ye Li, Ming LiAAAI 2023 · 被引用 5 次
- Online Prototype Alignment for Few-shot Policy TransferQi Yi, Rui Zhang, Shaohui Peng, Jiaming Guo 等ICML 2023 · 被引用 5 次
- Multi-Agent Domain Calibration with a Handful of Offline DataTao Jiang, Lei Yuan, Lihe Li, Cong Guan 等NeurIPS 2024 · 被引用 3 次
- FOUNDER: Grounding Foundation Models in World Models for Open-Ended Embodied Decision MakingYucen Wang, Rui Yu, Shenghua Wan, Le Gan 等ICML 2025
- Learning to Reuse Policies in State Evolvable EnvironmentsZiqian Zhang, Bohan Yang, Lihe Li, Yuqi Bian 等ICML 2025
它引用的顶会 Paper3
- MOPO: Model-based Offline Policy OptimizationTianhe Yu, Garrett Thomas, Lantao Yu, Stefano Ermon 等NeurIPS 2020 · 被引用 989 次
- Model Based Reinforcement Learning for AtariLukasz Kaiser, Mohammad Babaeizadeh, Piotr Milos, Blazej Osinski 等ICLR 2020 · 被引用 969 次
- RL-CycleGAN: Reinforcement Learning Aware Simulation-to-RealKanishka Rao, Chris Harris, Alex Irpan, Sergey Levine 等CVPR 2020
相关 Paper
- Latent Adaptation of Foundation Policies for Sim-to-Real TransferLongchao Da, Thirulogasankar Pranav Kutralingam, Lirong Xiang, Hua WeiICLR 2026
- Domain Adaptation In Reinforcement Learning Via Latent Unified State RepresentationJinwei Xing, Takashi Nagata, Kexin Chen, Xinyun Zou 等AAAI 2021 · 被引用 65 次
- Domain Adaptive Imitation Learning with Visual ObservationSungho Choi, Seungyul Han, Woojun Kim, Jongseong Chae 等NeurIPS 2023 · 被引用 15 次
- Synthetic-to-Real Camouflaged Object DetectionZhihao Luo, Luojun Lin, Zheng LinACM MM 2025 · 被引用 1 次
- Deep Head Pose Estimation Using Synthetic Images and Partial Adversarial Domain Adaption for Continuous Label SpacesFelix Kuhnke, Jörn OstermannICCV 2019 · 被引用 51 次
