FLNA: An Energy-Efficient Point Cloud Feature Learning Accelerator with Dataflow Decoupling
Dongxu Lyu, Zhenyu Li, Yuzhou Chen, Ningyi Xu, Guanghui He
摘要
Grid-based feature learning network plays a key role in recent point-cloud based 3D perception. However, high point sparsity and special operators lead to large memory footprint and long processing latency, posing great challenges to hardware acceleration. We propose FLNA, a novel feature learning accelerator with algorithm-architecture co-design. At algorithm level, the dataflow-decoupled graph is adopted to reduce 86% computation by exploiting inherent sparsity and concat redundancy. At hardware design level, we customize a pipelined architecture with block-wise processing, and introduce transposed SRAM strategy to save 82.1% access power. Implemented on a 40nm technology, FLNA achieves 13.4 − 43.3× speedup over RTX 2080Ti GPU. It rivals the state-of-the-art accelerator by 1.21× energy-efficiency improvement with 50.8% latency reduction.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper1
问问它们各自怎么用它相关 Paper
- High-throughput Point-Cloud Accelerator with Sparsity-aware Hierarchical Neighbor Voxel Search and SkippingYun-Chia Yu, Suraj Pn Reddy, Aryan Devrani, Anirudh Srinivasan 等DAC 2025
- PointAcc: Efficient Point Cloud AcceleratorYujun Lin, Zhekai Zhang, Haotian Tang, Hanrui Wang 等MICRO 2021 · 被引用 90 次
- FractalCloud: A Fractal-Inspired Architecture for Efficient Large-Scale Point Cloud ProcessingYuzhe Fu, Changchun Zhou, Hancheng Ye, Bowen Duan 等HPCA 2026 · 被引用 1 次
- PointShuffler: Accelerating Point Cloud Neural Networks on General-Purpose GPUsYangfan Li, Zhengjie Jin, Yue Tian, Mengquan Li 等EuroSys 2026
- Grid-GCN for Fast and Scalable Point Cloud LearningQiangeng Xu, Xudong Sun, Cho-Ying Wu, Panqu Wang 等CVPR 2020
