Accelerating Reinforcement Learning through GPU Atari Emulation
Steven Dalton, Iuri Frosio
摘要
We introduce CuLE (CUDA Learning Environment), a CUDA port of the Atari Learning Environment (ALE) which is used for the development of deep reinforcement algorithms. CuLE overcomes many limitations of existing CPU-based emulators and scales naturally to multiple GPUs. It leverages GPU parallelization to run thousands of games simultaneously and it renders frames directly on the GPU, to avoid the bottleneck arising from the limited CPU-GPU communication bandwidth. CuLE generates up to 155M frames per hour on a single GPU, a finding previously achieved only through a cluster of CPUs. Beyond highlighting the differences between CPU and GPU emulators in the context of reinforcement learning, we show how to leverage the high throughput of CuLE by effective batching of the training data, and show accelerated convergence for A2C+V-trace. CuLE is available at https://github.com/NVlabs/cule .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- Habitat 2.0: Training Home Assistants to Rearrange their HabitatAndrew Szot, Alexander Clegg, Eric Undersander, Erik Wijmans 等NeurIPS 2021 · 被引用 826 次
- Jumanji: a Diverse Suite of Scalable Reinforcement Learning Environments in JAXClément Bonnet, Daniel Luo, Donal Byrne, Shikha Surana 等ICLR 2024 · 被引用 52 次
- Atari-5: Distilling the Arcade Learning Environment down to Five GamesMatthew Aitchison, Penny Sweetser, Marcus HutterICML 2023 · 被引用 40 次
- Large Batch Simulation for Deep Reinforcement LearningBrennan Shacklett, Erik Wijmans, Aleksei Petrenko, Manolis Savva 等ICLR 2021 · 被引用 29 次
- An Extensible, Data-Oriented Architecture for High-Performance, Many-World SimulationBrennan Shacklett, Luc Guy Rosenzweig, Zhiqiang Xie, Bidipta Sarkar 等SIGGRAPH 2023 · 被引用 13 次
它引用的顶会 Paper1
相关 Paper
- Octax: Accelerated CHIP-8 Arcade Environments for Reinforcement Learning in JAXWaris Radji, Thomas Michel, Hector PiteauICLR 2026 · 被引用 5 次
- Parallel Q-Learning: Scaling Off-policy Reinforcement Learning under Massively Parallel SimulationZechu Li, Tao Chen, Zhang-Wei Hong, Anurag Ajay 等ICML 2023 · 被引用 27 次
- SEED RL: Scalable and Efficient Deep-RL with Accelerated Central InferenceLasse Espeholt, Raphaël Marinier, Piotr Stanczyk, Ke Wang 等ICLR 2020 · 被引用 32 次
- GPEmu: A GPU Emulator for Faster and Cheaper Prototyping and Evaluation of Deep Learning System ResearchMeng Wang, Gus Waldspurger, Naufal Ananda, Yuyang Huang 等VLDB 2025 · 被引用 1 次
- High-Throughput Synchronous Deep RLIou-Jen Liu, Raymond A. Yeh, Alexander G. SchwingNeurIPS 2020 · 被引用 13 次
