Jelly Bean World: A Testbed for Never-Ending Learning
Emmanouil Antonios Platanios, Abulhair Saparov, Tom M. Mitchell
摘要
Machine learning has shown growing success in recent years. However, current machine learning systems are highly specialized, trained for particular problems or domains, and typically on a single narrow dataset. Human learning, on the other hand, is highly general and adaptable. Never-ending learning is a machine learning paradigm that aims to bridge this gap, with the goal of encouraging researchers to design machine learning systems that can learn to perform a wider variety of inter-related tasks in more complex environments. To date, there is no environment or testbed to facilitate the development and evaluation of never-ending learning systems. To this end, we propose the Jelly Bean World testbed. The Jelly Bean World allows experimentation over two-dimensional grid worlds which are filled with items and in which agents can navigate. This testbed provides environments that are sufficiently complex and where more generally intelligent algorithms ought to perform better than current state-of-the-art reinforcement learning approaches. It does so by producing non-stationary environments and facilitating experimentation with multi-task, multi-agent, multi-modal, and curriculum learning settings. We hope that this new freely-available software will prompt new research and interest in the development and evaluation of never-ending learning systems and more broadly, general intelligence systems.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- A Definition of Continual Reinforcement LearningDavid Abel, André Barreto, Benjamin Van Roy, Doina Precup 等NeurIPS 2023 · 被引用 167 次
- Continual World: A Robotic Benchmark For Continual Reinforcement LearningMaciej Wolczyk, Michal Zajac, Razvan Pascanu, Lukasz Kucinski 等NeurIPS 2021 · 被引用 152 次
- Continuous Coordination As a Realistic Scenario for Lifelong LearningHadi Nekoei, Akilesh Badrinaaraayanan, Aaron C. Courville, Sarath ChandarICML 2021 · 被引用 51 次
- Autonomous Reinforcement Learning: Formalism and BenchmarkingArchit Sharma, Kelvin Xu, Nikhil Sardana, Abhishek Gupta 等ICLR 2022 · 被引用 39 次
- Prediction and Control in Continual Reinforcement LearningNishanth Anand, Doina PrecupNeurIPS 2023 · 被引用 26 次
相关 Paper
- Online Fast Adaptation and Knowledge Accumulation (OSAKA): a New Approach to Continual LearningMassimo Caccia, Pau Rodríguez, Oleksiy Ostapenko, Fabrice Normandin 等NeurIPS 2020 · 被引用 83 次
- Arena: A General Evaluation Platform and Building Toolkit for Multi-Agent IntelligenceYuhang Song, Andrzej Wojcicki, Thomas Lukasiewicz, Jianyi Wang 等AAAI 2020 · 被引用 36 次
- Online Continual Learning with Natural Distribution Shifts: An Empirical Study with Visual DataZhipeng Cai, Ozan Sener, Vladlen KoltunICCV 2021 · 被引用 101 次
- Never-ending Learning of User InterfacesJason Wu, Rebecca Krosnick, Eldon Schoop, Amanda Swearngin 等UIST 2023 · 被引用 17 次
- Learning to Learn: How to Continuously Teach Humans and MachinesParantak Singh, You Li, Ankur Sikarwar, Weixian Lei 等ICCV 2023 · 被引用 10 次
