Accelerating and pruning CNNs for semantic segmentation on FPGA
Pierpaolo Morì, Manoj Rohit Vemparala, Nael Fasfous, Saptarshi Mitra, Sreetama Sarkar, Alexander Frickenstein, Lukas Frickenstein, Domenik Helms, Naveen Shankar Nagaraja, Walter Stechele, Claudio Passerone
Abstract
Semantic segmentation is one of the popular tasks in computer vision, providing pixel-wise annotations for scene understanding. However, segmentation-based convolutional neural networks require tremendous computational power. In this work, a fully-pipelined hardware accelerator with support for dilated convolution is introduced, which cuts down the redundant zero multiplications. Furthermore, we propose a genetic algorithm based automated channel pruning technique to jointly optimize computational complexity and model accuracy. Finally, hardware heuristics and an accurate model of the custom accelerator design enable a hardware-aware pruning framework. We achieve 2.44X lower latency with minimal degradation in semantic prediction quality (−1.98 pp lower mean intersection over union) compared to the baseline DeepLabV3+ model, evaluated on an Arria-10 FPGA. The binary files of the FPGA design, baseline and pruned models can be found in github.com/pierpaolomori/SemanticSegmentationFPGA
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on1
Related papers
- 3D CNN Acceleration on FPGA using Hardware-Aware PruningMengshu Sun, Pu Zhao, Mehmet Güngör, Massoud Pedram et al.DAC 2020 · 39 citations
- DEPrune: Depth-wise Separable Convolution Pruning for Maximizing GPU ParallelismCheonjun Park, Mincheol Park, Hyunchan Moon, Myung Kuk Yoon et al.NeurIPS 2024 · 10 citations
- Pruning In Time (PIT): A Lightweight Network Architecture Optimizer for Temporal Convolutional NetworksMatteo Risso, Alessio Burrello, Daniele Jahier Pagliari, Francesco Conti et al.DAC 2021 · 12 citations
- DeepBurning-SEG: Generating DNN Accelerators of Segment-Grained Pipeline ArchitectureXuyi Cai, Ying Wang, Xiaohan Ma, Yinhe Han et al.MICRO 2022 · 25 citations
- AutoGO: Automated Computation Graph Optimization for Neural Network EvolutionMohammad Salameh, Keith G. Mills, Negar Hassanpour, Fred X. Han et al.NeurIPS 2023 · 7 citations
