Efficient Multitask Dense Predictor via Binarization
Yuzhang Shang, Dan Xu, Gaowen Liu, Ramana Rao Kompella, Yan Yan
Abstract
Multi-task learning for dense prediction has emerged as a pivotal area in computer vision, enabling simultaneous processing of diverse yet interrelated pixel-wise prediction tasks. However, the substantial computational demands of state-of-the-art (SoTA) models often limit their widespread deployment. This paper addresses this challenge by introducing network binarization to compress resource-intensive multi-task dense predictors. Specifically, our goal is to significantly accelerate multi-task dense prediction models via Binary Neural Networks (BNNs) while maintaining and even improving model performance at the same time. To reach this goal, we propose a Binary Multi-task Dense Predictor, Bi -MTPD, and several variants of Bi -MTPD, in which a multi-task dense predictor is constructed via specified binarized modules. Our systematical analysis of this predictor reveals that performance drop from binarization is primarily caused by severe information degradation. To address this issue, we introduce a deep information bottleneck layer that enforces representations for downstream tasks satisfying Gaussian distribution in forward propagation. Moreover, we introduce a knowledge distillation mechanism to correct the direction of information flow in backward propagation. Intriguingly, one variant of Bi -MTPD outperforms full-precision (FP) multi-task dense prediction SoTAs, ARTC [2] (CNN-based) and InvPT [50] (ViT-Based). This result indicates that Bi -MTPD is not merely a naive trade-off between performance and efficiency, but is rather a benefit of the redundant information flow thanks to the multi-task architecture. Code is available at BiMTDP.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 45ae05e6-4ff9-4024-8e9c-7613d2aa7369Cited by top-tier papers1
Ask how each one uses itBuilds on13
- Do Vision Transformers See Like Convolutional Neural Networks?Maithra Raghu, Thomas Unterthiner, Simon Kornblith, Chiyuan Zhang et al.NeurIPS 2021 · 1,553 citations
- BiBERT: Accurate Fully Binarized BERTHaotong Qin, Yifu Ding, Mingyuan Zhang, Qinghua Yan et al.ICLR 2022 · 121 citations
- Exploring Relational Context for Multi-Task Dense PredictionDavid Brüggemann, Menelaos Kanakis, Anton Obukhov, Stamatios Georgoulis et al.ICCV 2021 · 110 citations
- How Do Adam and Training Strategies Help BNNs OptimizationZechun Liu, Zhiqiang Shen, Shichao Li, Koen Helwegen et al.ICML 2021 · 100 citations
- BiPointNet: Binary Neural Network for Point CloudsHaotong Qin, Zhongang Cai, Mingyuan Zhang, Yifu Ding et al.ICLR 2021 · 54 citations
Related papers
- BiDet: An Efficient Binarized Object DetectorZiwei Wang, Ziyi Wu, Jiwen Lu, Jie ZhouCVPR 2020
- Fast and Accurate Binary Neural Networks Based on Depth-Width ReshapingPing Xue, Yang Lu, Jingfei Chang, Xing Wei et al.AAAI 2023 · 3 citations
- Contrastive Multi-Task Dense PredictionSiwei Yang, Hanrong Ye, Dan XuAAAI 2023 · 13 citations
- MST-compression: Compressing and Accelerating Binary Neural Networks with Minimum Spanning TreeQuang Hieu Vo, Linh-Tam Tran, Sung-Ho Bae, Lok-Won Kim et al.ICCV 2023 · 2 citations
- Training Binary Neural Network without Batch Normalization for Image Super-ResolutionXinrui Jiang, Nannan Wang, Jingwei Xin, Keyu Li et al.AAAI 2021 · 52 citations
