DiSparse: Disentangled Sparsification for Multitask Model Compression
Xinglong Sun, Ali Hassani, Zhangyang Wang, Gao Huang, Humphrey Shi
Abstract
Despite the popularity of Model Compression and Mul-titask Learning, how to effectively compress a multitask model has been less thoroughly analyzed due to the chal-lenging entanglement of tasks in the parameter space. In this paper, we propose DiSparse, a simple, effective, and first-of-its-kind multitask pruning and sparse training scheme. We consider each task independently by disentangling the importance measurement and take the unani-mous decisions among all tasks when performing parame-ter pruning and selection. Our experimental results demon-strate superior performance on various configurations and settings compared to popular sparse training and pruning methods. Besides the effectiveness in compression, DiS-parse also provides a powerful tool to the multitask learning community. Surprisingly, we even observed better per-formance than some dedicated multitask learning methods in several cases despite the high model sparsity enforced by DiSparse. We analyzed the pruning masks generated with DiSparse and observed strikingly similar sparse net-work architecture identified by each task even before the training starts. We also observe the existence of a “water-shed” layer where the task relatedness sharply drops, implying no benefits in continued parameters sharing. Our code and models will be available at: https://github.com/SHI-Labs/DiSparse-Multitask-Model-Compression.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a1dedc09-2266-4be4-b5bd-9169527fcdbbCited by top-tier papers7
- Split-Ensemble: Efficient OOD-aware Ensemble via Task and Model SplittingAnthony Chen, Huanrui Yang, Yulu Gan, Denis A. Gudovskiy et al.ICML 2024 · 5 citations
- AdapMTL: Adaptive Pruning Framework for Multitask Learning ModelMingcan Xiang, Jiaxun Tang, Qizheng Yang, Hui Guan et al.ACM MM 2024 · 4 citations
- LoGIC: Multi-LoRA Guided Importance Consensus for Multi-Task Pruning in Vision TransformersYu-Hong Chou, Rui Fang, Hsi-Wen Chen, Ming-Syan ChenAAAI 2026
- CALM: Consensus-Aware Localized Merging for Multi-Task LearningKunda Yan, Min Zhang, Sen Cui, Zikun Qu et al.ICML 2025
- MDP: Multidimensional Vision Model Pruning with Latency ConstraintXinglong Sun, Barath Lakshmanan, Maying Shen, Shiyi Lan et al.CVPR 2025
Builds on9
- Rigging the Lottery: Making All Tickets WinnersUtku Evci, Trevor Gale, Jacob Menick, Pablo Samuel Castro et al.ICML 2020 · 723 citations
- Which Tasks Should Be Learned Together in Multi-task Learning?Trevor Standley, Amir Zamir, Dawn Chen, Leonidas J. Guibas et al.ICML 2020 · 651 citations
- AdaShare: Learning What To Share For Efficient Deep Multi-Task LearningXimeng Sun, Rameswar Panda, Rogério Feris, Kate SaenkoNeurIPS 2020 · 337 citations
- ResRep: Lossless CNN Pruning via Decoupling Remembering and ForgettingXiaohan Ding, Tianxiang Hao, Jianchao Tan, Ji Liu et al.ICCV 2021 · 202 citations
- Any-Precision Deep Neural NetworksHaichao Yu, Haoxiang Li, Humphrey Shi, Thomas S. Huang et al.AAAI 2021 · 79 citations
Related papers
- Pruning-Aware Merging for Efficient Multitask InferenceXiaoxi He, Dawei Gao, Zimu Zhou, Yongxin Tong et al.KDD 2021 · 8 citations
- Learning Sparse Sharing Architectures for Multiple TasksTianxiang Sun, Yunfan Shao, Xiaonan Li, Pengfei Liu et al.AAAI 2020 · 155 citations
- CABS: Conflict-Aware and Balanced Sparsification for Enhancing Model MergingZongzhen Yang, Binhang Qi, Hailong Sun, Wenrui Long et al.ICML 2025
- Controllable Dynamic Multi-Task ArchitecturesDripta S. Raychaudhuri, Yumin Suh, Samuel Schulter, Xiang Yu et al.CVPR 2022 · 24 citations
- Efficiently Identifying Task Groupings for Multi-Task LearningChris Fifty, Ehsan Amid, Zhe Zhao, Tianhe Yu et al.NeurIPS 2021 · 352 citations
