Mesorasi: Architecture Support for Point Cloud Analytics via Delayed-Aggregation
Yu Feng, Boyuan Tian, Tiancheng Xu, Paul N. Whatmough, Yuhao Zhu
摘要
Point cloud analytics is poised to become a key workload on battery-powered embedded and mobile platforms in a wide range of emerging application domains, such as autonomous driving, robotics, and augmented reality, where efficiency is paramount. This paper proposes Mesorasi, an algorithm-architecture co-designed system that simultaneously improves the performance and energy efficiency of point cloud analytics while retaining its accuracy.
Our extensive characterizations of state-of-the-art point cloud algorithms show that, while structurally reminiscent of convolutional neural networks (CNNs), point cloud algorithms exhibit inherent compute and memory inefficiencies due to the unique characteristics of point cloud data. We propose delayed-aggregation, a new algorithmic primitive for building efficient point cloud algorithms. Delayed-aggregation hides the performance bottlenecks and reduces the compute and memory redundancies by exploiting the approximately distributive property of key operations in point cloud algorithms. Delayed-aggregation let point cloud algorithms achieve 1.6× speedup and 51.1% energy reduction on a mobile GPU while retaining the accuracy (-0.9% loss to 1.2% gains). To maximize the algorithmic benefits, we propose minor extensions to contemporary CNN accelerators, which can be integrated into a mobile Systems-on-a-Chip (SoC) without modifying other SoC components. With additional hardware support, Mesorasi achieves up to 3.6× speedup.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper18
- SpAtten: Efficient Sparse Attention Architecture with Cascade Token and Head PruningHanrui Wang, Zhekai Zhang, Song HanHPCA 2021 · 被引用 412 次
- PointAcc: Efficient Point Cloud AcceleratorYujun Lin, Zhekai Zhang, Haotian Tang, Hanrui Wang 等MICRO 2021 · 被引用 90 次
- NeuRex: A Case for Neural Rendering AccelerationJunseo Lee, Kwanseok Choi, Jungi Lee, Seokwon Lee 等ISCA 2023 · 被引用 49 次
- Crescent: taming memory irregularities for accelerating deep point cloud analyticsYu Feng, Gunnar Hammonds, Yiming Gan, Yuhao ZhuISCA 2022 · 被引用 44 次
- TorchSparse++: Efficient Training and Inference Framework for Sparse Convolution on GPUsHaotian Tang, Shang Yang, Zhijian Liu, Ke Hong 等MICRO 2023 · 被引用 32 次
它引用的顶会 Paper3
- SemanticKITTI: A Dataset for Semantic Scene Understanding of LiDAR SequencesJens Behley, Martin Garbade, Andres Milioto, Jan Quenzel 等ICCV 2019 · 被引用 2,345 次
- HyGCN: A GCN Accelerator with Hybrid ArchitectureMingyu Yan, Lei Deng, Xing Hu, Ling Liang 等HPCA 2020 · 被引用 338 次
- DensePoint: Learning Densely Contextual Representation for Efficient Point Cloud ProcessingYongcheng Liu, Bin Fan, Gaofeng Meng, Jiwen Lu 等ICCV 2019 · 被引用 295 次
相关 Paper
- PointISA: ISA-Extensions for Efficient Point Cloud Analytics via Architecture and Algorithm Co-DesignMeng Han, Liang Wang, Limin Xiao, Hao Zhang 等MICRO 2025 · 被引用 3 次
- HgPCN: A Heterogeneous Architecture for E2E Embedded Point Cloud InferenceYiming Gao, Chao Jiang, Wesley Piard, Xiangru Chen 等MICRO 2024 · 被引用 6 次
- PointCIM: A Computing-in-Memory Architecture for Accelerating Deep Point Cloud AnalyticsXuan-Jun Chen, Han-Ping Chen, Chia-Lin YangMICRO 2024 · 被引用 4 次
- Point Cloud Acceleration by Exploiting Geometric SimilarityCen Chen, Xiaofeng Zou, Hongen Shao, Yangfan Li 等MICRO 2023 · 被引用 19 次
- PointShuffler: Accelerating Point Cloud Neural Networks on General-Purpose GPUsYangfan Li, Zhengjie Jin, Yue Tian, Mengquan Li 等EuroSys 2026
