A Hierarchical Adaptive Multi-Task Reinforcement Learning Framework for Multiplier Circuit Design
Zhihai Wang, Jie Wang, Dongsheng Zuo, Yunjie Ji, Xilin Xia, Yuzhe Ma, Jianye Hao, Mingxuan Yuan, Yongdong Zhang, Feng Wu
Abstract
Multiplier design-which aims to explore a large combinatorial design space to simultaneously optimize multiple conflicting objectives-is a fundamental problem in the integrated circuits industry. Although traditional approaches tackle the multi-objective multiplier optimization problem by manually designed heuristics, reinforcement learning (RL) offers a promising approach to discover high-speed and area-efficient multipliers. However, the existing RL-based methods struggle to find Pareto-optimal circuit designs for all possible preferences, i.e., weights over objectives, in a sample-efficient manner. To address this challenge, we propose a novel hierarchical adaptive (HAVE) multi-task reinforcement learning framework. The hierarchical framework consists of a meta-agent to generate diverse multiplier preferences, and an adaptive multi-task agent to collaboratively optimize multipliers conditioned on the dynamic preferences given by the meta-agent. To the best of our knowledge, HAVE is the first to well approximate Pareto-optimal circuit designs for the entire preference space with high sample efficiency. Experiments on multipliers across a wide range of input widths demonstrate that HAVE significantly Pareto-dominates state-of-theart approaches, achieving up to 28% larger hypervolume. Moreover, experiments demonstrate that multipliers designed by HAVE can well generalize to large-scale computation-intensive circuits.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 56a243e2-b98c-481a-92ce-ea5f5e831514Cited by top-tier papers18
- Reinforcement Learning within Tree Search for Fast Macro PlacementZijie Geng, Jie Wang, Ziyan Liu, Siyuan Xu et al.ICML 2024 · 23 citations
- Towards Next-Generation Logic Synthesis: A Scalable Neural Circuit Generation FrameworkZhihai Wang, Jie Wang, Qingyue Yang, Yinqi Bai et al.NeurIPS 2024 · 17 citations
- Scalable and Effective Arithmetic Tree Generation for Adder and Multiplier DesignsYao Lai, Jinxin Liu, David Z. Pan, Ping LuoNeurIPS 2024 · 14 citations
- Uncertainty-based Offline Variational Bayesian Reinforcement Learning for Robustness under Diverse Data CorruptionsRui Yang, Jie Wang, Guoping Wu, Bin LiNeurIPS 2024 · 11 citations
- Fast and Interpretable Mixed-Integer Linear Program Solving by Learning Model ReductionYixuan Li, Can Chen, Jiajun Li, Jiahui Duan et al.AAAI 2025 · 3 citations
Builds on13
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- ChiPFormer: Transferable Chip Placement via Offline Decision TransformerYao Lai, Jinxin Liu, Zhentao Tang, Bin Wang et al.ICML 2023 · 69 citations
- PrefixRL: Optimization of Parallel Prefix Circuits using Deep Reinforcement LearningRajarshi Roy, Jonathan Raiman, Neel Kant, Ilyas Elkin et al.DAC 2021 · 53 citations
- Regret Minimization Experience Replay in Off-Policy Reinforcement LearningXu-Hui Liu, Zhenghai Xue, Jing-Cheng Pang, Shengyi Jiang et al.NeurIPS 2021 · 51 citations
- Sample-Efficient Reinforcement Learning via Conservative Model-Based Actor-CriticZhihai Wang, Jie Wang, Qi Zhou, Bin Li et al.AAAI 2022 · 38 citations
Related papers
- RL-MUL: Multiplier Design Optimization with Deep Reinforcement LearningDongsheng Zuo, Yikang Ouyang, Yuzhe MaDAC 2023 · 17 citations
- Automated Design of Complex Analog Circuits with Multiagent based Reinforcement LearningJinxin Zhang, Jiarui Bao, Zhangcheng Huang, Xuan Zeng et al.DAC 2023 · 27 citations
- Computing Circuits Optimization via Model-Based Circuit Genetic EvolutionZhihai Wang, Jie Wang, Xilin Xia, Dongsheng Zuo et al.ICLR 2025
- Towards Automated RISC-V Microarchitecture Design with Reinforcement LearningChen Bai, Jianwang Zhai, Yuzhe Ma, Bei Yu et al.AAAI 2024 · 20 citations
- PD-MORL: Preference-Driven Multi-Objective Reinforcement Learning AlgorithmToygun Basaklar, Suat Gumussoy, Ümit Y. OgrasICLR 2023 · 8 citations
