PIDNet: A Real-time Semantic Segmentation Network Inspired by PID Controllers
Jiacong Xu, Zixiang Xiong, Shankar P. Bhattacharyya
Abstract
Two-branch network architecture has shown its efficiency and effectiveness in real-time semantic segmentation tasks. However, direct fusion of high-resolution details and low-frequency context has the drawback of detailed features being easily overwhelmed by surrounding contextual information. This overshoot phenomenon limits the improvement of the segmentation accuracy of existing two-branch models. In this paper, we make a connection between Convolutional Neural Networks (CNN) and Proportional-Integral-Derivative (PID) controllers and reveal that a two-branch network is equivalent to a Proportional-Integral (PI) controller, which inherently suffers from similar overshoot issues. To alleviate this problem, we propose a novel threebranch network architecture: PIDNet, which contains three branches to parse detailed, context and boundary information, respectively, and employs boundary attention to guide the fusion of detailed and context branches. Our family of PIDNets achieve the best trade-off between inference speed and accuracy and their accuracy surpasses all the existing models with similar inference speed on the Cityscapes and CamVid datasets. Specifically, PIDNet-S achieves 78.6% mIOU with inference speed of 93.2 FPS on Cityscapes and 80.1% mIOU with speed of 153.7 FPS on CamVid.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0e804a19-edbc-4396-9650-e9f5226d0140Cited by top-tier papers20
- SCTNet: Single-Branch CNN with Transformer Semantic Information for Real-Time SegmentationZhengze Xu, Dongyue Wu, Changqian Yu, Xiangxiang Chu et al.AAAI 2024 · 166 citations
- Contextrast: Contextual Contrastive Learning for Semantic SegmentationChangki Sung, Wanhee Kim, Jungho An, Wooju Lee et al.CVPR 2024 · 29 citations
- AMDANet: Attention-Driven Multi-Perspective Discrepancy Alignment for RGB-Infrared Image Fusion and SegmentationHaifeng Zhong, Fan Tang, Zhuo Chen, Hyung Jin Chang et al.ICCV 2025 · 9 citations
- An Embedding-Unleashing Video Polyp Segmentation Framework via Region Linking and Scale AlignmentZhixue Fang, Xinrong Guo, Jingyin Lin, Huisi Wu et al.AAAI 2024 · 8 citations
- SGFormer: Semantic-Geometry Fusion Transformer for Multi-modal 3D Panoptic SegmentationHongqi Yu, Sixian Chan, Xiaolong Zhou, Xiaoqin ZhangAAAI 2025 · 3 citations
Builds on7
- CCNet: Criss-Cross Attention for Semantic SegmentationZilong Huang, Xinggang Wang, Lichao Huang, Chang Huang et al.ICCV 2019 · 2,972 citations
- Gated-SCNN: Gated Shape CNNs for Semantic SegmentationTowaki Takikawa, David Acuna, Varun Jampani, Sanja FidlerICCV 2019 · 710 citations
- FasterSeg: Searching for Faster Real-time Semantic SegmentationWuyang Chen, Xinyu Gong, Xianming Liu, Qian Zhang et al.ICLR 2020 · 206 citations
- HyperSeg: Patch-Wise Hypernetwork for Real-Time Semantic SegmentationYuval Nirkin, Lior Wolf, Tal HassnerCVPR 2021
- Rethinking BiSeNet for Real-Time Semantic SegmentationMingyuan Fan, Shenqi Lai, Junshi Huang, Xiaoming Wei et al.CVPR 2021
Related papers
- Pixel Difference Networks for Efficient Edge DetectionZhuo Su, Wenzhe Liu, Zitong Yu, Dewen Hu et al.ICCV 2021 · 488 citations
- RTFormer: Efficient Design for Real-Time Semantic Segmentation with TransformerJian Wang, Chenhui Gou, Qiman Wu, Haocheng Feng et al.NeurIPS 2022 · 207 citations
- AttaNet: Attention-Augmented Network for Fast and Accurate Scene ParsingQi Song, Kangfu Mei, Rui HuangAAAI 2021 · 89 citations
- Efficient Parallel Multi-Scale Detail and Semantic Encoding Network for Lightweight Semantic SegmentationXiao Liu, Xiuya Shi, Lufei Chen, Linbo Qing et al.ACM MM 2023 · 7 citations
- Squeeze-and-Attention Networks for Semantic SegmentationZilong Zhong, Zhong Qiu Lin, Rene Bidart, Xiaodan Hu et al.CVPR 2020
