Task-Agnostic Morphology Evolution
Donald Joseph Hejna III, Pieter Abbeel, Lerrel Pinto
摘要
Deep reinforcement learning primarily focuses on learning behavior, usually overlooking the fact that an agent's function is largely determined by form. So, how should one go about finding a morphology fit for solving tasks in a given environment? Current approaches that co-adapt morphology and behavior use a specific task's reward as a signal for morphology optimization. However, this often requires expensive policy optimization and results in task-dependent morphologies that are not built to generalize. In this work, we propose a new approach, Task-Agnostic Morphology Evolution (TAME), to alleviate both of these issues. Without any task or reward specification, TAME evolves morphologies by only applying randomly sampled action primitives on a population of agents. This is accomplished using an information-theoretic objective that efficiently ranks agents by their ability to reach diverse states in the environment and the causality of their actions. Finally, we empirically demonstrate that across 2D, 3D, and manipulation environments TAME can evolve morphologies that match the multi-task performance of those learned with task supervised algorithms. Our code and videos can be found at https://sites.google.com/view/task-agnostic-evolution .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- Evolution Gym: A Large-Scale Benchmark for Evolving Soft RobotsJagdeep Singh Bhatia, Holly Jackson, Yunsheng Tian, Jie Xu 等NeurIPS 2021 · 被引用 141 次
- MetaMorph: Learning Universal Controllers with TransformersAgrim Gupta, Linxi Fan, Surya Ganguli, Li Fei-FeiICLR 2022 · 被引用 130 次
- Transform2Act: Learning a Transform-and-Control Policy for Efficient Agent DesignYe Yuan, Yuda Song, Zhengyi Luo, Wen Sun 等ICLR 2022 · 被引用 51 次
- REvolveR: Continuous Evolutionary Models for Robot-to-robot Policy TransferXingyu Liu, Deepak Pathak, Kris KitaniICML 2022 · 被引用 24 次
- VLMgineer: Vision-Language Models as Robotic ToolsmithsGeorge Jiayuan Gao, Tianyu Li, Junyao Shi, Yihan Li 等ICLR 2026 · 被引用 14 次
它引用的顶会 Paper2
相关 Paper
- AnyMorph: Learning Transferable Polices By Inferring Agent MorphologyBrandon Trabucco, Mariano Phielipp, Glen BersethICML 2022 · 被引用 37 次
- A System for Morphology-Task Generalization via Unified Representation and Behavior DistillationHiroki Furuta, Yusuke Iwasawa, Yutaka Matsuo, Shixiang Shane GuICLR 2023
- ECo-MoE: Embodiment-Conditioned Mixture of Experts Increases the Evolvability of RobotsYibin Wang, Muhan Li, Zihan Guo, Sam KriegmanICML 2026
- Convergent Functions, Divergent FormsHyeonseong Jeon, Ainaz Eftekhar, Aaron Walsman, Kuo-Hao Zeng 等NeurIPS 2025 · 被引用 5 次
- Meta-Evolve: Continuous Robot Evolution for One-to-many Policy TransferXingyu Liu, Deepak Pathak, Ding ZhaoICLR 2024 · 被引用 6 次
