DiSparse: Disentangled Sparsification for Multitask Model Compression
Xinglong Sun, Ali Hassani, Zhangyang Wang, Gao Huang, Humphrey Shi
摘要
Despite the popularity of Model Compression and Mul-titask Learning, how to effectively compress a multitask model has been less thoroughly analyzed due to the chal-lenging entanglement of tasks in the parameter space. In this paper, we propose DiSparse, a simple, effective, and first-of-its-kind multitask pruning and sparse training scheme. We consider each task independently by disentangling the importance measurement and take the unani-mous decisions among all tasks when performing parame-ter pruning and selection. Our experimental results demon-strate superior performance on various configurations and settings compared to popular sparse training and pruning methods. Besides the effectiveness in compression, DiS-parse also provides a powerful tool to the multitask learning community. Surprisingly, we even observed better per-formance than some dedicated multitask learning methods in several cases despite the high model sparsity enforced by DiSparse. We analyzed the pruning masks generated with DiSparse and observed strikingly similar sparse net-work architecture identified by each task even before the training starts. We also observe the existence of a “water-shed” layer where the task relatedness sharply drops, implying no benefits in continued parameters sharing. Our code and models will be available at: https://github.com/SHI-Labs/DiSparse-Multitask-Model-Compression.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Split-Ensemble: Efficient OOD-aware Ensemble via Task and Model SplittingAnthony Chen, Huanrui Yang, Yulu Gan, Denis A. Gudovskiy 等ICML 2024 · 被引用 5 次
- AdapMTL: Adaptive Pruning Framework for Multitask Learning ModelMingcan Xiang, Jiaxun Tang, Qizheng Yang, Hui Guan 等ACM MM 2024 · 被引用 4 次
- LoGIC: Multi-LoRA Guided Importance Consensus for Multi-Task Pruning in Vision TransformersYu-Hong Chou, Rui Fang, Hsi-Wen Chen, Ming-Syan ChenAAAI 2026
- CALM: Consensus-Aware Localized Merging for Multi-Task LearningKunda Yan, Min Zhang, Sen Cui, Zikun Qu 等ICML 2025
- MDP: Multidimensional Vision Model Pruning with Latency ConstraintXinglong Sun, Barath Lakshmanan, Maying Shen, Shiyi Lan 等CVPR 2025
它引用的顶会 Paper9
- Rigging the Lottery: Making All Tickets WinnersUtku Evci, Trevor Gale, Jacob Menick, Pablo Samuel Castro 等ICML 2020 · 被引用 723 次
- Which Tasks Should Be Learned Together in Multi-task Learning?Trevor Standley, Amir Zamir, Dawn Chen, Leonidas J. Guibas 等ICML 2020 · 被引用 651 次
- AdaShare: Learning What To Share For Efficient Deep Multi-Task LearningXimeng Sun, Rameswar Panda, Rogério Feris, Kate SaenkoNeurIPS 2020 · 被引用 337 次
- ResRep: Lossless CNN Pruning via Decoupling Remembering and ForgettingXiaohan Ding, Tianxiang Hao, Jianchao Tan, Ji Liu 等ICCV 2021 · 被引用 202 次
- Any-Precision Deep Neural NetworksHaichao Yu, Haoxiang Li, Humphrey Shi, Thomas S. Huang 等AAAI 2021 · 被引用 79 次
相关 Paper
- Pruning-Aware Merging for Efficient Multitask InferenceXiaoxi He, Dawei Gao, Zimu Zhou, Yongxin Tong 等KDD 2021 · 被引用 8 次
- Learning Sparse Sharing Architectures for Multiple TasksTianxiang Sun, Yunfan Shao, Xiaonan Li, Pengfei Liu 等AAAI 2020 · 被引用 155 次
- CABS: Conflict-Aware and Balanced Sparsification for Enhancing Model MergingZongzhen Yang, Binhang Qi, Hailong Sun, Wenrui Long 等ICML 2025
- Controllable Dynamic Multi-Task ArchitecturesDripta S. Raychaudhuri, Yumin Suh, Samuel Schulter, Xiang Yu 等CVPR 2022 · 被引用 24 次
- Efficiently Identifying Task Groupings for Multi-Task LearningChris Fifty, Ehsan Amid, Zhe Zhao, Tianhe Yu 等NeurIPS 2021 · 被引用 352 次
