Transform2Act: Learning a Transform-and-Control Policy for Efficient Agent Design
Ye Yuan, Yuda Song, Zhengyi Luo, Wen Sun, Kris M. Kitani
Abstract
An agent's functionality is largely determined by its design, i.e., skeletal structure and joint attributes (e.g., length, size, strength). However, finding the optimal agent design for a given function is extremely challenging since the problem is inherently combinatorial and the design space is prohibitively large. Additionally, it can be costly to evaluate each candidate design which requires solving for its optimal controller. To tackle these problems, our key idea is to incorporate the design procedure of an agent into its decision-making process. Specifically, we learn a conditional policy that, in an episode, first applies a sequence of transform actions to modify an agent's skeletal structure and joint attributes, and then applies control actions under the new design. To handle a variable number of joints across designs, we use a graph-based policy where each graph node represents a joint and uses message passing with its neighbors to output joint-specific actions. Using policy gradient methods, our approach enables joint optimization of agent design and control as well as experience sharing across different designs, which improves sample efficiency substantially. Experiments show that our approach, Transform2Act, outperforms prior methods significantly in terms of convergence speed and final performance. Notably, Transform2Act can automatically discover plausible designs similar to giraffes, squids, and spiders. Code and videos are available at https://sites.google.com/view/transform2act.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dc51f03e-0804-4598-90b9-16beb3d42fc3Cited by top-tier papers16
- Universal Morphology Control via Contextual ModulationZheng Xiong, Jacob Beck, Shimon WhitesonICML 2023 · 27 citations
- Symmetry-Aware Robot Design with Structured SubgroupsHeng Dong, Junyu Zhang, Tonghan Wang, Chongjie ZhangICML 2023 · 19 citations
- VLMgineer: Vision-Language Models as Robotic ToolsmithsGeorge Jiayuan Gao, Tianyu Li, Junyao Shi, Yihan Li et al.ICLR 2026 · 14 citations
- Accelerated co-design of robots through morphological pretrainingLuke Strgar, Sam KriegmanICLR 2026 · 13 citations
- House Of Dextra : Cross-Embodied Co-Design for Dexterous HandsKehlani Fay, Darin Anthony Djapri, Anya Zorin, James Clinton et al.ICLR 2026 · 9 citations
Builds on3
- One Policy to Control Them All: Shared Modular Policies for Agent-Agnostic ControlWenlong Huang, Igor Mordatch, Deepak PathakICML 2020 · 214 citations
- My Body is a Cage: the Role of Morphology in Graph-Based Incompatible ControlVitaly Kurin, Maximilian Igl, Tim Rocktäschel, Wendelin Boehmer et al.ICLR 2021 · 105 citations
- Task-Agnostic Morphology EvolutionDonald Joseph Hejna III, Pieter Abbeel, Lerrel PintoICLR 2021 · 32 citations
Related papers
- Structure-Aware Transformer Policy for Inhomogeneous Multi-Task Reinforcement LearningSunghoon Hong, Deunsol Yoon, Kee-Eung KimICLR 2022 · 40 citations
- Efficient Morphology-Control Co-Design via Stackelberg Proximal Policy OptimizationYanning Dai, Yuhui Wang, Dylan R. Ashley, Jürgen SchmidhuberICLR 2026 · 2 citations
- MetaMorph: Learning Universal Controllers with TransformersAgrim Gupta, Linxi Fan, Surya Ganguli, Li Fei-FeiICLR 2022 · 130 citations
- Neurosymbolic Transformers for Multi-Agent CommunicationJeevana Priya Inala, Yichen Yang, James Paulos, Yewen Pu et al.NeurIPS 2020 · 29 citations
- BodyGen: Advancing Towards Efficient Embodiment Co-DesignHaofei Lu, Zhe Wu, Junliang Xing, Jianshu Li et al.ICLR 2025
