Speculative Monte-Carlo Tree Search
Scott Cheng, Mahmut T. Kandemir, Ding-Yong Hong
摘要
Monte-Carlo tree search (MCTS) is an influential sequential decision-making algo-rithm notably employed in AlphaZero. Despite its success, the primary challenge in AlphaZero training lies in its prolonged time-to-solution due to the high latency imposed by the sequential MCTS process. To address this challenge, this paper proposes and evaluates an inter-decision parallelization strategy called speculative MCTS , a new type of parallelism in AlphaZero which implements speculative execution. This approach allows for the parallel execution of future moves before the current MCTS computations are completed, thus reducing the latency. Additionally, we analyze factors contributing to the overall speedup by studying the synergistic effects of speculation and neural network caching in MCTS. We also provide an analytical model that can be used to evaluate the potential of different speculation strategies before they are implemented and deployed. Our empirical findings indicate that the proposed speculative MCTS can reduce training latency by 5.81 × in 9x9 Go games. Moreover, our study shows that speculative execution can enhance the NN cache hit rate by 26% during midgame. Overall, our end-to-end evaluation indicates 1.91 × speedup in 19x19 Go training time, compared to the state-of-the-art KataGo program.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Uncertainty-Guided Exploration for Efficient AlphaZero TrainingScott Cheng, Meng-Yu Tsai, Ding-Yong Hong, Mahmut T. KandemirNeurIPS 2025
- Breaking the Reward Barrier: Accelerating Tree-of-Thought Reasoning via Speculative ExplorationShuzhang Zhong, Haochen Huang, Shengxuan Qiu, Pengfei Zuo 等OSDI 2026
- ToC: Tree-of-Claims Search with Multi-Agent Language ModelsShuyang Yu, Jianan Liang, Hui HuAAAI 2026
它引用的顶会 Paper1
相关 Paper
- Learning to Stop: Dynamic Simulation Monte-Carlo Tree SearchLi-Cheng Lan, Ti-Rong Wu, I-Chen Wu, Cho-Jui HsiehAAAI 2021 · 被引用 7 次
- Speculative Actions: A Lossless Framework for Faster AI AgentsNaimeng Ye, Arnav Ahuja, Georgios Liargkovas, Yunan Lu 等ICLR 2026 · 被引用 9 次
- SpecFL: An Efficient Speculative Federated Learning System for Tree-based Model TrainingYuhui Zhang, Lutan Zhao, Cheng Che, XiaoFeng Wang 等HPCA 2024 · 被引用 3 次
- Are AlphaZero-like Agents Robust to Adversarial Perturbations?Li-Cheng Lan, Huan Zhang, Ti-Rong Wu, Meng-Yu Tsai 等NeurIPS 2022 · 被引用 15 次
- EasySpec: Layer-Parallel Speculative Decoding for Efficient Multi-GPU UtilizationYize Wu, Ke Gao, Ling Li, Yanjun WuNeurIPS 2025 · 被引用 3 次
