HDPG: hyperdimensional policy-based reinforcement learning for continuous control
Yang Ni, Mariam Issa, Danny Abraham, Mahdi Imani, Xunzhao Yin, Mohsen Imani
摘要
Traditional robot control or more general continuous control tasks often rely on carefully hand-crafted classic control methods. These models often lack the self-learning adaptability and intelligence to achieve human-level control. On the other hand, recent advancements in Reinforcement Learning (RL) present algorithms that have the capability of human-like learning. The integration of Deep Neural Networks (DNN) and RL thereby enables autonomous learning in robot control tasks. However, DNN-based RL brings both highquality learning and high computation cost, which is no longer ideal for currently fast-growing edge computing scenarios.
In this paper, we introduce HDPG, a highly-efficient policy-based RL algorithm using Hyperdimensional Computing. Hyperdimensional computing is a lightweight brain-inspired learning methodology; its holistic representation of information leads to a well-defined set of hardware-friendly high-dimensional operations. Our HDPG fully exploits the efficient HDC for high-quality state value approximation and policy gradient update. In our experiments, we use HDPG for robotics tasks with continuous action space and achieve significantly higher rewards than DNN-based RL. Our evaluation also shows that HDPG achieves 4.7× faster and 5.3× higher energy efficiency than DNN-based RL running on embedded FPGA.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper3
- Cognitive Correlative Encoding for Genome Sequence Matching in Hyperdimensional SystemPrathyush Poduval, Zhuowen Zou, Xunzhao Yin, Elaheh Sadredini 等DAC 2021 · 被引用 39 次
- StocHD: Stochastic Hyperdimensional System for Efficient and Robust Learning from Raw DataPrathyush Poduval, Zhuowen Zou, M. Hassan Najafi, Houman Homayoun 等DAC 2021 · 被引用 34 次
- PRID: Model Inversion Privacy Attacks in Hyperdimensional Learning SystemsAlejandro Hernández-Cano, Rosario Cammarota, Mohsen ImaniDAC 2021 · 被引用 30 次
相关 Paper
- Scalable edge-based hyperdimensional learning system with brain-like neural adaptationZhuowen Zou, Yeseong Kim, Farhad Imani, Haleh Alimohamadi 等SC 2021 · 被引用 70 次
- DistHD: A Learner-Aware Dynamic Encoding Method for Hyperdimensional ClassificationJunyao Wang, Sitao Huang, Mohsen ImaniDAC 2023 · 被引用 16 次
- Revisiting HyperDimensional Learning for FPGA and Low-Power ArchitecturesMohsen Imani, Zhuowen Zou, Samuel Bosch, Sanjay Anantha Rao 等HPCA 2021 · 被引用 90 次
- FATE: Boosting the Performance of Hyper-Dimensional Computing Intelligence with Flexible Numerical DAta TypEHaomin Li, Fangxin Liu, Yichi Chen, Zongwu Wang 等ISCA 2025 · 被引用 4 次
- Prive-HD: Privacy-Preserved Hyperdimensional ComputingBehnam Khaleghi, Mohsen Imani, Tajana RosingDAC 2020 · 被引用 35 次
