Alzo: Auto-Tuning with Reinforcement Learning for DAG-based Blockchains
Qiuyu Ding, Rongkai Zhang, Qinnan Zhang, Zhen Xiao, Jieyi Long, Mingchao Wan, Sen Liu, Jin Dong
Abstract
As critical infrastructure for Web 3.0, DAG-based blockchains promise high throughput for DeFi, IoT, and DApps. However, realizing this potential is challenging, as system performance is dictated by a multitude of interdependent parameters across network, node, and consensus layers. Manual configuration fails to adapt to dynamic workloads, leading to suboptimal performance. We introduce Alzo, a novel auto-tuner that employs hierarchical reinforcement learning (HRL) to navigate this complex configuration space. By decomposing the DAG blockchain's workflow into distinct stages, Alzo's HRL policy learns from stage-level performance metrics to control critical parameters governing consensus, execution, and graph topology in real-time. Furthermore, we employ a shadow-control loop to ensure the safety of all parameter adjustments. Our experiments show that Alzo significantly outperforms other configurations, achieving higher throughput and lower latency under variable workloads with minimal overhead.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 146b3d20-d33b-42ea-9c81-84904bdc7ed3Builds on7
- A Decentralized Blockchain with High Throughput and Fast ConfirmationChenxing Li, Peilun Li, Dong Zhou, Zhe Yang et al.USENIX ATC 2020 · 172 citations
- OHIE: Blockchain Scaling Made SimpleHaifeng Yu, Ivica Nikolic, Ruomu Hou, Prateek SaxenaS&P 2020 · 166 citations
- ResTune: Resource Oriented Tuning Boosted by Meta-Learning for Cloud DatabasesXinyi Zhang, Hong Wu, Zhuo Chang, Shuowei Jin et al.SIGMOD 2021 · 113 citations
- CGPTuner: a Contextual Gaussian Process Bandit Approach for the Automatic Tuning of IT Configurations Under Varying Workload ConditionsStefano Cereda, Stefano Valladares, Paolo Cremonesi, Stefano DoniVLDB 2021 · 74 citations
- SPRING: Improving the Throughput of Sharding Blockchain via Deep Reinforcement Learning Based State PlacementPengze Li, Mingxuan Song, Mingzhe Xing, Zhen Xiao et al.WWW 2024 · 32 citations
Related papers
- Auto-Tuning with Reinforcement Learning for Permissioned Blockchain SystemsMingxuan Li, Yazhe Wang, Shuai Ma, Chao Liu et al.VLDB 2023 · 32 citations
- AdaChain: A Learned Adaptive BlockchainChenyuan Wu, Bhavana Mehta, Mohammad Javad Amiri, Ryan Marcus et al.VLDB 2023 · 20 citations
- Scarf: Self-Adaptive Tuning via Multi-Objective Reinforcement Learning for Apache FlinkLiu Liu, Shenghao Gong, Ziquan Fang, Yunjun GaoVLDB 2026
- Sequential Multi-Agent Dynamic Algorithm ConfigurationChen Lu, Ke Xue, Lei Yuan, Yao Wang et al.NeurIPS 2025 · 8 citations
- Multi-agent Dynamic Algorithm ConfigurationKe Xue, Jiacheng Xu, Lei Yuan, Miqing Li et al.NeurIPS 2022 · 65 citations
