SCC: an efficient deep reinforcement learning agent mastering the game of StarCraft II
Xiangjun Wang, Junxiao Song, Penghui Qi, Peng Peng, Zhenkun Tang, Wei Zhang, Weimin Li, Xiongjun Pi, Jujie He, Chao Gao, Haitao Long, Quan Yuan
Abstract
AlphaStar, the AI that reaches GrandMaster level in StarCraft II, is a remarkable milestone demonstrating what deep reinforcement learning can achieve in complex Real-Time Strategy (RTS) games. However, the complexities of the game, algorithms and systems, and especially the tremendous amount of computation needed are big obstacles for the community to conduct further research in this direction. We propose a deep reinforcement learning agent, StarCraft Commander (SCC). With order of magnitude less computation, it demonstrates top human performance defeating GrandMaster players in test matches and top professional players in a live event. Moreover, it shows strong robustness to various human strategies and discovers novel strategies unseen from human plays. In this paper, we will share the key insights and optimizations on efficient imitation learning and reinforcement learning for StarCraft II full game.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- Large Language Models Play StarCraft II: Benchmarks and A Chain of Summarization ApproachWeiyu Ma, Qirui Mi, Yongcheng Zeng, Xue Yan et al.NeurIPS 2024 · 122 citations
- A Robust and Opponent-Aware League Training Method for StarCraft IIRuozi Huang, Xipeng Wu, Hongsheng Yu, Zhong Fan et al.NeurIPS 2023 · 13 citations
- Iterative Regularized Policy Optimization with Imperfect DemonstrationsXudong Gong, Dawei Feng, Kele Xu, Yuanzhao Zhai et al.ICML 2024 · 5 citations
- DMR: Decomposed Multi-Modality Representations for Frames and Events Fusion in Visual Reinforcement LearningHaoran Xu, Peixi Peng, Guang Tan, Yuan Li et al.CVPR 2024 · 5 citations
- In-Context Compositional Q-Learning for Offline Reinforcement LearningQiushui Xu, Yuhao Huang, Yushu Jiang, Wenliang Zheng et al.ICLR 2026
Related papers
- Mastering Complex Control in MOBA Games with Deep Reinforcement LearningDeheng Ye, Zhao Liu, Mingfei Sun, Bei Shi et al.AAAI 2020 · 395 citations
- Incorporating Pragmatic Reasoning Communication into Emergent LanguageYipeng Kang, Tonghan Wang, Gerard de MeloNeurIPS 2020 · 26 citations
- Mimicking To Dominate: Imitation Learning Strategies for Success in Multiagent GamesThe Viet Bui, Tien Mai, Thanh Hong NguyenNeurIPS 2024 · 5 citations
- HiMacMic: Hierarchical Multi-Agent Deep Reinforcement Learning with Dynamic Asynchronous Macro StrategyHancheng Zhang, Guozheng Li, Chi Harold Liu, Guoren Wang et al.KDD 2023 · 1 citation
- LLM-PySC2: Starcraft II learning environment for Large Language ModelsZongyuan Li, Yanan Ni, Runnan Qi, Chang Lu et al.NeurIPS 2025 · 15 citations
