Transform2Act: Learning a Transform-and-Control Policy for Efficient Agent Design
Ye Yuan, Yuda Song, Zhengyi Luo, Wen Sun, Kris M. Kitani
摘要
An agent's functionality is largely determined by its design, i.e., skeletal structure and joint attributes (e.g., length, size, strength). However, finding the optimal agent design for a given function is extremely challenging since the problem is inherently combinatorial and the design space is prohibitively large. Additionally, it can be costly to evaluate each candidate design which requires solving for its optimal controller. To tackle these problems, our key idea is to incorporate the design procedure of an agent into its decision-making process. Specifically, we learn a conditional policy that, in an episode, first applies a sequence of transform actions to modify an agent's skeletal structure and joint attributes, and then applies control actions under the new design. To handle a variable number of joints across designs, we use a graph-based policy where each graph node represents a joint and uses message passing with its neighbors to output joint-specific actions. Using policy gradient methods, our approach enables joint optimization of agent design and control as well as experience sharing across different designs, which improves sample efficiency substantially. Experiments show that our approach, Transform2Act, outperforms prior methods significantly in terms of convergence speed and final performance. Notably, Transform2Act can automatically discover plausible designs similar to giraffes, squids, and spiders. Code and videos are available at https://sites.google.com/view/transform2act.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper16
- Universal Morphology Control via Contextual ModulationZheng Xiong, Jacob Beck, Shimon WhitesonICML 2023 · 被引用 27 次
- Symmetry-Aware Robot Design with Structured SubgroupsHeng Dong, Junyu Zhang, Tonghan Wang, Chongjie ZhangICML 2023 · 被引用 19 次
- VLMgineer: Vision-Language Models as Robotic ToolsmithsGeorge Jiayuan Gao, Tianyu Li, Junyao Shi, Yihan Li 等ICLR 2026 · 被引用 14 次
- Accelerated co-design of robots through morphological pretrainingLuke Strgar, Sam KriegmanICLR 2026 · 被引用 13 次
- House Of Dextra : Cross-Embodied Co-Design for Dexterous HandsKehlani Fay, Darin Anthony Djapri, Anya Zorin, James Clinton 等ICLR 2026 · 被引用 9 次
它引用的顶会 Paper3
- One Policy to Control Them All: Shared Modular Policies for Agent-Agnostic ControlWenlong Huang, Igor Mordatch, Deepak PathakICML 2020 · 被引用 214 次
- My Body is a Cage: the Role of Morphology in Graph-Based Incompatible ControlVitaly Kurin, Maximilian Igl, Tim Rocktäschel, Wendelin Boehmer 等ICLR 2021 · 被引用 105 次
- Task-Agnostic Morphology EvolutionDonald Joseph Hejna III, Pieter Abbeel, Lerrel PintoICLR 2021 · 被引用 32 次
相关 Paper
- Structure-Aware Transformer Policy for Inhomogeneous Multi-Task Reinforcement LearningSunghoon Hong, Deunsol Yoon, Kee-Eung KimICLR 2022 · 被引用 40 次
- Efficient Morphology-Control Co-Design via Stackelberg Proximal Policy OptimizationYanning Dai, Yuhui Wang, Dylan R. Ashley, Jürgen SchmidhuberICLR 2026 · 被引用 2 次
- MetaMorph: Learning Universal Controllers with TransformersAgrim Gupta, Linxi Fan, Surya Ganguli, Li Fei-FeiICLR 2022 · 被引用 130 次
- Neurosymbolic Transformers for Multi-Agent CommunicationJeevana Priya Inala, Yichen Yang, James Paulos, Yewen Pu 等NeurIPS 2020 · 被引用 29 次
- BodyGen: Advancing Towards Efficient Embodiment Co-DesignHaofei Lu, Zhe Wu, Junliang Xing, Jianshu Li 等ICLR 2025
