Blocks Assemble! Learning to Assemble with Large-Scale Structured Reinforcement Learning
Seyed Kamyar Seyed Ghasemipour, Satoshi Kataoka, Byron David, Daniel Freeman, Shixiang Shane Gu, Igor Mordatch
Abstract
Assembly of multi-part physical structures is both a valuable end product for autonomous robotics, as well as a valuable diagnostic task for open-ended training of embodied intelligent agents. We introduce a naturalistic physics-based environment with a set of connectable magnet blocks inspired by children's toy kits. The objective is to assemble blocks into a succession of target blueprints. Despite the simplicity of this objective, the compositional nature of building diverse blueprints from a set of blocks leads to an explosion of complexity in structures that agents encounter. Furthermore, assembly stresses agents' multi-step planning, physical reasoning, and bimanual coordination. We find that the combination of large-scale reinforcement learning and graph-based policies -surprisingly without any additional complexity -is an effective recipe for training agents that not only generalize to complex unseen blueprints in a zero-shot manner, but even operate in a reset-free setting without being trained to do so. Through extensive experiments, we highlight the importance of large-scale training, structured representations, contributions of multi-task vs. single-task learning, as well as the effects of curriculums, and discuss qualitative behaviors of trained agents. Our accompanying project webpage can be found at: sites.google.com/view/learning-direct-assembly
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b1dac019-b834-4db2-8adf-4297ed54d5d4Cited by top-tier papers5
- Simple Unsupervised Object-Centric Learning for Complex and Naturalistic VideosGautam Singh, Yi-Fu Wu, Sungjin AhnNeurIPS 2022 · 182 citations
- Unsupervised Object Interaction Learning with Counterfactual Dynamics ModelsJongwook Choi, Sungtae Lee, Xinyu Wang, Sungryull Sohn et al.AAAI 2024 · 6 citations
- BrickNet: Graph-Backed Generative Brick AssemblyPeter Kulits, Cordelia SchmidCVPR 2026 · 5 citations
- A3D: Adaptive Affordance Assembly with Dual-Arm ManipulationJiaqi Liang, Yue Chen, Qize Yu, Yan Shen et al.AAAI 2026 · 3 citations
- A System for Morphology-Task Generalization via Unified Representation and Behavior DistillationHiroki Furuta, Yusuke Iwasawa, Yutaka Matsuo, Shixiang Shane GuICLR 2023
Builds on5
- Emergent Tool Use From Multi-Agent AutocurriculaBowen Baker, Ingmar Kanitscheider, Todor M. Markov, Yi Wu et al.ICLR 2020 · 751 citations
- One Policy to Control Them All: Shared Modular Policies for Agent-Agnostic ControlWenlong Huang, Igor Mordatch, Deepak PathakICML 2020 · 214 citations
- My Body is a Cage: the Role of Morphology in Graph-Based Incompatible ControlVitaly Kurin, Maximilian Igl, Tim Rocktäschel, Wendelin Boehmer et al.ICLR 2021 · 105 citations
- What Matters for On-Policy Deep Actor-Critic Methods? A Large-Scale StudyMarcin Andrychowicz, Anton Raichuk, Piotr Stanczyk, Manu Orsini et al.ICLR 2021 · 52 citations
- Brick-by-Brick: Combinatorial Construction with Deep Reinforcement LearningHyunsoo Chung, Jungtaek Kim, Boris Knyazev, Jinhwi Lee et al.NeurIPS 2021 · 29 citations
Related papers
- Generalization to New Actions in Reinforcement LearningAyush Jain, Andrew Szot, Joseph J. LimICML 2020 · 39 citations
- Hierarchical Abstraction for Combinatorial Generalization in Object RearrangementMichael Chang, Alyssa L. Dayan, Franziska Meier, Thomas L. Griffiths et al.ICLR 2023
- Learning to Assemble with Alternative PlansZiqi Wang, Wenjun Liu, Jingwen Wang, Gabriel Vallat et al.SIGGRAPH 2025
- Learning Compositional Tasks from Language InstructionsLajanugen Logeswaran, Wilka Carvalho, Honglak LeeAAAI 2023 · 4 citations
- Ask Your Humans: Using Human Instructions to Improve Generalization in Reinforcement LearningValerie Chen, Abhinav Gupta, Kenneth MarinoICLR 2021 · 6 citations
