FractalCloud: A Fractal-Inspired Architecture for Efficient Large-Scale Point Cloud Processing
Yuzhe Fu, Changchun Zhou, Hancheng Ye, Bowen Duan, Qiyu Huang, Chiyue Wei, Cong Guo, Hai Helen Li, Yiran Chen
Abstract
Three-dimensional (3D) point clouds are increasingly used in applications such as autonomous driving, robotics, and virtual reality (VR). Point-based neural networks (PNNs) have demonstrated strong performance in point cloud analysis, originally targeting small-scale inputs. However, as PNNs evolve to process large-scale point clouds with hundreds of thousands of points, all-to-all computation and global memory access in point cloud processing introduce substantial overhead, causingcomputational complexity and memory traffic whereis the number of points. Existing accelerators, primarily optimized for small-scale workloads, overlook this challenge and scale poorly due to inefficient partitioning and non-parallel architectures. To address these issues, we propose FractalCloud, a fractal-inspired hardware architecture for efficient large-scale 3D point cloud processing. FractalCloud introduces two key optimizations: (1) a co-designed Fractal method for shape-aware and hardware-friendly partitioning, and (2) block-parallel point operations that decompose and parallelize all point operations. A dedicated hardware design with on-chip fractal and flexible parallelism further enables fully parallel processing within limited memory resources. Implemented in 28 nm technology as a chip layout with a core area of, FractalCloud achievesspeedup andenergy reduction over state-of-the-art accelerators while maintaining network accuracy, demonstrating its scalability and efficiency for PNN inference. The code for FractalCloud is available at https://github.com/Yuzhe-Fu/FractalCloud.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d806dafa-84e6-425a-88ab-2e7784a8ab3aCited by top-tier papers1
Ask how each one uses itBuilds on15
- PointNeXt: Revisiting PointNet++ with Improved Training and Scaling StrategiesGuocheng Qian, Yuchen Li, Houwen Peng, Jinjie Mai et al.NeurIPS 2022 · 1,270 citations
- OliVe: Accelerating Large Language Models via Hardware-friendly Outlier-Victim Pair QuantizationCong Guo, Jiaming Tang, Weiming Hu, Jingwen Leng et al.ISCA 2023 · 151 citations
- Object DGCNN: 3D Object Detection using Dynamic GraphsYue Wang, Justin M. SolomonNeurIPS 2021 · 127 citations
- PointAcc: Efficient Point Cloud AcceleratorYujun Lin, Zhekai Zhang, Haotian Tang, Hanrui Wang et al.MICRO 2021 · 90 citations
- Instant-3D: Instant Neural Radiance Field Training Towards On-Device AR/VR 3D ReconstructionSixu Li, Chaojian Li, Wenbo Zhu, Boyang Tony Yu et al.ISCA 2023 · 79 citations
Related papers
- An Efficient Accelerator for Point-based and Voxel-based Point Cloud Neural NetworksXinhao Yang, Tianyu Fu, Guohao Dai, Shulin Zeng et al.DAC 2023 · 25 citations
- Point Cloud Acceleration by Exploiting Geometric SimilarityCen Chen, Xiaofeng Zou, Hongen Shao, Yangfan Li et al.MICRO 2023 · 19 citations
- PointShuffler: Accelerating Point Cloud Neural Networks on General-Purpose GPUsYangfan Li, Zhengjie Jin, Yue Tian, Mengquan Li et al.EuroSys 2026
- PointCIM: A Computing-in-Memory Architecture for Accelerating Deep Point Cloud AnalyticsXuan-Jun Chen, Han-Ping Chen, Chia-Lin YangMICRO 2024 · 4 citations
- BitNN: A Bit-Serial Accelerator for K-Nearest Neighbor Search in Point CloudsMeng Han, Liang Wang, Limin Xiao, Hao Zhang et al.ISCA 2024 · 14 citations
