A3C-S: Automated Agent Accelerator Co-Search towards Efficient Deep Reinforcement Learning
Yonggan Fu, Yongan Zhang, Chaojian Li, Zhongzhi Yu, Yingyan Lin
Abstract
Driven by the explosive interest in applying deep reinforcement learning (DRL) agents to numerous real-time control and decision-making applications, there has been a growing demand to deploy DRL agents to empower daily-life intelligent devices, while the prohibitive complexity of DRL stands at odds with limited on-device resources. In this work, we propose an Automated Agent Accelerator Co-Search (A3C-S) framework, which to our best knowledge is the first to automatically co-search the optimally matched DRL agents and accelerators that maximize both test scores and hardware efficiency. Extensive experiments consistently validate the superiority of our A3C-S over state-of-the-art techniques.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 774712ab-da39-4a77-834e-be51c0eec3f5Cited by top-tier papers1
Ask how each one uses itBuilds on5
- AutoGAN-Distiller: Searching to Compress Generative Adversarial NetworksYonggan Fu, Wuyang Chen, Haotao Wang, Haoran Li et al.ICML 2020 · 91 citations
- Timely: Pushing Data Movements And Interfaces In Pim Accelerators Towards Local And In Time DomainWeitao Li, Pengfei Xu, Yang Zhao, Haitong Li et al.ISCA 2020 · 86 citations
- EDD: Efficient Differentiable DNN Architecture and Implementation Co-search for Embedded AI SolutionsYuhong Li, Cong Hao, Xiaofan Zhang, Xinheng Liu et al.DAC 2020 · 79 citations
- SmartExchange: Trading Higher-cost Memory Storage/Access for Lower-cost ComputationYang Zhao, Xiaohan Chen, Yue Wang, Chaojian Li et al.ISCA 2020 · 44 citations
- MiLeNAS: Efficient Neural Architecture Search via Mixed-Level ReformulationChaoyang He, Haishan Ye, Li Shen, Tong ZhangCVPR 2020
Related papers
- DeepWiERL: Bringing Deep Reinforcement Learning to the Internet of Self-Adaptive ThingsFrancesco Restuccia, Tommaso MelodiaINFOCOM 2020 · 34 citations
- : On-Device Real-Time Deep Reinforcement Learning for Autonomous RoboticsZexin Li, Aritra Samanta, Yufei Li, Andrea Soltoggio et al.RTSS 2023 · 9 citations
- Decentralized Application-Level Adaptive Scheduling for Multi-Instance DNNs on Open Mobile DevicesHsin-Hsuan Sung, Jou-An Chen, Wei Niu, Jiexiong Guan et al.USENIX ATC 2023 · 9 citations
- DETERRENT: detecting trojans using reinforcement learningVasudev Gohil, Satwik Patnaik, Hao Guo, Dileep Kalathil et al.DAC 2022 · 26 citations
- Auto-NBA: Efficient and Effective Search Over the Joint Space of Networks, Bitwidths, and AcceleratorsYonggan Fu, Yongan Zhang, Yang Zhang, David D. Cox et al.ICML 2021 · 23 citations
