Jelly Bean World: A Testbed for Never-Ending Learning
Emmanouil Antonios Platanios, Abulhair Saparov, Tom M. Mitchell
Abstract
Machine learning has shown growing success in recent years. However, current machine learning systems are highly specialized, trained for particular problems or domains, and typically on a single narrow dataset. Human learning, on the other hand, is highly general and adaptable. Never-ending learning is a machine learning paradigm that aims to bridge this gap, with the goal of encouraging researchers to design machine learning systems that can learn to perform a wider variety of inter-related tasks in more complex environments. To date, there is no environment or testbed to facilitate the development and evaluation of never-ending learning systems. To this end, we propose the Jelly Bean World testbed. The Jelly Bean World allows experimentation over two-dimensional grid worlds which are filled with items and in which agents can navigate. This testbed provides environments that are sufficiently complex and where more generally intelligent algorithms ought to perform better than current state-of-the-art reinforcement learning approaches. It does so by producing non-stationary environments and facilitating experimentation with multi-task, multi-agent, multi-modal, and curriculum learning settings. We hope that this new freely-available software will prompt new research and interest in the development and evaluation of never-ending learning systems and more broadly, general intelligence systems.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- A Definition of Continual Reinforcement LearningDavid Abel, André Barreto, Benjamin Van Roy, Doina Precup et al.NeurIPS 2023 · 167 citations
- Continual World: A Robotic Benchmark For Continual Reinforcement LearningMaciej Wolczyk, Michal Zajac, Razvan Pascanu, Lukasz Kucinski et al.NeurIPS 2021 · 152 citations
- Continuous Coordination As a Realistic Scenario for Lifelong LearningHadi Nekoei, Akilesh Badrinaaraayanan, Aaron C. Courville, Sarath ChandarICML 2021 · 51 citations
- Autonomous Reinforcement Learning: Formalism and BenchmarkingArchit Sharma, Kelvin Xu, Nikhil Sardana, Abhishek Gupta et al.ICLR 2022 · 39 citations
- Prediction and Control in Continual Reinforcement LearningNishanth Anand, Doina PrecupNeurIPS 2023 · 26 citations
Related papers
- Online Fast Adaptation and Knowledge Accumulation (OSAKA): a New Approach to Continual LearningMassimo Caccia, Pau Rodríguez, Oleksiy Ostapenko, Fabrice Normandin et al.NeurIPS 2020 · 83 citations
- Arena: A General Evaluation Platform and Building Toolkit for Multi-Agent IntelligenceYuhang Song, Andrzej Wojcicki, Thomas Lukasiewicz, Jianyi Wang et al.AAAI 2020 · 36 citations
- Online Continual Learning with Natural Distribution Shifts: An Empirical Study with Visual DataZhipeng Cai, Ozan Sener, Vladlen KoltunICCV 2021 · 101 citations
- Never-ending Learning of User InterfacesJason Wu, Rebecca Krosnick, Eldon Schoop, Amanda Swearngin et al.UIST 2023 · 17 citations
- Learning to Learn: How to Continuously Teach Humans and MachinesParantak Singh, You Li, Ankur Sikarwar, Weixian Lei et al.ICCV 2023 · 10 citations
