HDPG: hyperdimensional policy-based reinforcement learning for continuous control
Yang Ni, Mariam Issa, Danny Abraham, Mahdi Imani, Xunzhao Yin, Mohsen Imani
Abstract
Traditional robot control or more general continuous control tasks often rely on carefully hand-crafted classic control methods. These models often lack the self-learning adaptability and intelligence to achieve human-level control. On the other hand, recent advancements in Reinforcement Learning (RL) present algorithms that have the capability of human-like learning. The integration of Deep Neural Networks (DNN) and RL thereby enables autonomous learning in robot control tasks. However, DNN-based RL brings both highquality learning and high computation cost, which is no longer ideal for currently fast-growing edge computing scenarios.
In this paper, we introduce HDPG, a highly-efficient policy-based RL algorithm using Hyperdimensional Computing. Hyperdimensional computing is a lightweight brain-inspired learning methodology; its holistic representation of information leads to a well-defined set of hardware-friendly high-dimensional operations. Our HDPG fully exploits the efficient HDC for high-quality state value approximation and policy gradient update. In our experiments, we use HDPG for robotics tasks with continuous action space and achieve significantly higher rewards than DNN-based RL. Our evaluation also shows that HDPG achieves 4.7× faster and 5.3× higher energy efficiency than DNN-based RL running on embedded FPGA.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5b9eb177-066d-4e4e-bb15-0672ba959eb5Cited by top-tier papers1
Ask how each one uses itBuilds on3
- Cognitive Correlative Encoding for Genome Sequence Matching in Hyperdimensional SystemPrathyush Poduval, Zhuowen Zou, Xunzhao Yin, Elaheh Sadredini et al.DAC 2021 · 39 citations
- StocHD: Stochastic Hyperdimensional System for Efficient and Robust Learning from Raw DataPrathyush Poduval, Zhuowen Zou, M. Hassan Najafi, Houman Homayoun et al.DAC 2021 · 34 citations
- PRID: Model Inversion Privacy Attacks in Hyperdimensional Learning SystemsAlejandro Hernández-Cano, Rosario Cammarota, Mohsen ImaniDAC 2021 · 30 citations
Related papers
- Scalable edge-based hyperdimensional learning system with brain-like neural adaptationZhuowen Zou, Yeseong Kim, Farhad Imani, Haleh Alimohamadi et al.SC 2021 · 70 citations
- DistHD: A Learner-Aware Dynamic Encoding Method for Hyperdimensional ClassificationJunyao Wang, Sitao Huang, Mohsen ImaniDAC 2023 · 16 citations
- Revisiting HyperDimensional Learning for FPGA and Low-Power ArchitecturesMohsen Imani, Zhuowen Zou, Samuel Bosch, Sanjay Anantha Rao et al.HPCA 2021 · 90 citations
- FATE: Boosting the Performance of Hyper-Dimensional Computing Intelligence with Flexible Numerical DAta TypEHaomin Li, Fangxin Liu, Yichi Chen, Zongwu Wang et al.ISCA 2025 · 4 citations
- Prive-HD: Privacy-Preserved Hyperdimensional ComputingBehnam Khaleghi, Mohsen Imani, Tajana RosingDAC 2020 · 35 citations
