Powderworld: A Platform for Understanding Generalization via Rich Task Distributions
Kevin Frans, Phillip Isola
Abstract
One of the grand challenges of reinforcement learning is the ability to generalize to new tasks. However, general agents require a set of rich, diverse tasks to train on. Designing a 'foundation environment' for such tasks is tricky -the ideal environment would support a range of emergent phenomena, an expressive task space, and fast runtime. To take a step towards addressing this research bottleneck, this work presents Powderworld, a lightweight yet expressive simulation environment running directly on the GPU. Within Powderworld, two motivating challenges distributions are presented, one for world-modelling and one for reinforcement learning. Each contains hand-designed test tasks to examine generalization. Experiments indicate that increasing the environment's complexity improves generalization for world models and certain reinforcement learning agents, yet may inhibit learning in high-variance environments. Powderworld aims to support the study of generalization by providing a source of diverse tasks arising from the same core rules. Try an interactable demo at kvfrans.com/static/powder
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- The Generalization Gap in Offline Reinforcement LearningIshita Mediratta, Qingfei You, Minqi Jiang, Roberta RaileanuICLR 2024 · 24 citations
- Procedural Generation Of Algorithm Discovery Tasks in Machine LearningAlexander D. Goldie, Zilin Wang, Adrian Hayler, Deepak Nathani et al.ICML 2026 · 2 citations
- OGBench: Benchmarking Offline Goal-Conditioned RLSeohong Park, Kevin Frans, Benjamin Eysenbach, Sergey LevineICLR 2025
- Kinetix: Investigating the Training of General Agents through Open-Ended Physics-Based Control TasksMichael T. Matthews, Michael Beukman, Chris Lu, Jakob Nicolaus FoersterICLR 2025
- Provable Zero-Shot Generalization in Offline Reinforcement LearningZhiyong Wang, Chen Yang, John C. S. Lui, Dongruo ZhouICML 2025
Builds on11
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Zero-Shot Text-to-Image GenerationAditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray et al.ICML 2021 · 6,356 citations
- Dream to Control: Learning Behaviors by Latent ImaginationDanijar Hafner, Timothy P. Lillicrap, Jimmy Ba, Mohammad NorouziICLR 2020 · 1,852 citations
- Scaling Vision TransformersXiaohua Zhai, Alexander Kolesnikov, Neil Houlsby, Lucas BeyerCVPR 2022 · 767 citations
Related papers
- CausalWorld: A Robotic Manipulation Benchmark for Causal Structure and Transfer LearningOssama Ahmed, Frederik Träuble, Anirudh Goyal, Alexander Neitz et al.ICLR 2021 · 31 citations
- Improving Policy Optimization with Generalist-Specialist LearningZhiwei Jia, Xuanlin Li, Zhan Ling, Shuang Liu et al.ICML 2022 · 32 citations
- Toward Compositional Generalization in Object-Oriented World ModelingLinfeng Zhao, Lingzhi Kong, Robin Walters, Lawson L. S. WongICML 2022 · 28 citations
- Modeling Unseen Environments with Language-guided Composable Causal Components in Reinforcement LearningXinyue Wang, Biwei HuangICLR 2025
- Generalization to New Actions in Reinforcement LearningAyush Jain, Andrew Szot, Joseph J. LimICML 2020 · 39 citations
