EDGI: Equivariant Diffusion for Planning with Embodied Agents
Johann Brehmer, Joey Bose, Pim de Haan, Taco S. Cohen
摘要
Embodied agents operate in a structured world, often solving tasks with spatial, temporal, and permutation symmetries. Most algorithms for planning and modelbased reinforcement learning (MBRL) do not take this rich geometric structure into account, leading to sample inefficiency and poor generalization. We introduce the Equivariant Diffuser for Generating Interactions (EDGI), an algorithm for MBRL and planning that is equivariant with respect to the product of the spatial symmetry group SE(3), the discrete-time translation group Z, and the object permutation group S n . EDGI follows the Diffuser framework by Janner et al. [2022] in treating both learning a world model and planning in it as a conditional generative modeling problem, training a diffusion model on an offline trajectory dataset. We introduce a new SE(3) × Z × S n -equivariant diffusion model that supports multiple representations. We integrate this model in a planning loop, where conditioning and classifier guidance let us softly break the symmetry for specific tasks as needed. On object manipulation and navigation tasks, EDGI is substantially more sample efficient and generalizes better across the symmetry group than non-equivariant models. * Equal contribution, order determined through a game of table tennis † Qualcomm AI Research is an initiative of Qualcomm Technologies, Inc. ‡ Work done during an internship at Qualcomm AI Research 4 This is true in the approximately flat spacetime on Earth, as long as all velocities are much smaller than the speed of light. A machine learning researcher who finds herself close to a black hole may disagree. 37th Conference on Neural Information Processing Systems (NeurIPS 2023).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper10
- SE(3)-Stochastic Flow Matching for Protein Backbone GenerationAvishek Joey Bose, Tara Akhound-Sadegh, Guillaume Huguet, Kilian Fatras 等ICLR 2024 · 被引用 162 次
- Iterated Denoising Energy Matching for Sampling from Boltzmann DensitiesTara Akhound-Sadegh, Jarrid Rector-Brooks, Avishek Joey Bose, Sarthak Mittal 等ICML 2024 · 被引用 109 次
- Geometric Algebra TransformerJohann Brehmer, Pim de Haan, Sönke Behrends, Taco S. CohenNeurIPS 2023 · 被引用 81 次
- Simple Hierarchical Planning with DiffusionChang Chen, Fei Deng, Kenji Kawaguchi, Caglar Gulcehre 等ICLR 2024 · 被引用 79 次
- Rigid Body Flows for Sampling Molecular Crystal StructuresJonas Köhler, Michele Invernizzi, Pim de Haan, Frank NoéICML 2023 · 被引用 42 次
它引用的顶会 Paper19
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Diffusion Models Beat GANs on Image SynthesisPrafulla Dhariwal, Alexander Quinn NicholNeurIPS 2021 · 被引用 13,211 次
- Conservative Q-Learning for Offline Reinforcement LearningAviral Kumar, Aurick Zhou, George Tucker, Sergey LevineNeurIPS 2020 · 被引用 2,881 次
- E(n) Equivariant Graph Neural NetworksVictor Garcia Satorras, Emiel Hoogeboom, Max WellingICML 2021 · 被引用 1,432 次
- Planning with Diffusion for Flexible Behavior SynthesisMichael Janner, Yilun Du, Joshua B. Tenenbaum, Sergey LevineICML 2022 · 被引用 1,115 次
相关 Paper
- Quotient-Space Diffusion ModelsYixian Xu, Yusong Wang, Shengjie Luo, Kaiyuan Gao 等ICLR 2026 · 被引用 1 次
- Space Group Equivariant Crystal DiffusionRees Chang, Angela Pak, Alex Guerra, Ni Zhan 等NeurIPS 2025 · 被引用 20 次
- Boosting Sample Efficiency and Generalization in Multi-agent Reinforcement Learning via EquivarianceJoshua McClellan, Naveed Haghani, John Winder, Furong Huang 等NeurIPS 2024 · 被引用 21 次
- Subequivariant Graph Reinforcement Learning in 3D EnvironmentsRunfa Chen, Jiaqi Han, Fuchun Sun, Wenbing HuangICML 2023 · 被引用 14 次
- Partially Equivariant Reinforcement Learning in Symmetry-Breaking EnvironmentsJunwoo Chang, Minwoo Park, Joohwan Seo, Roberto Horowitz 等ICLR 2026 · 被引用 3 次
