Learning Sparse Sharing Architectures for Multiple Tasks
Tianxiang Sun, Yunfan Shao, Xiaonan Li, Pengfei Liu, Hang Yan, Xipeng Qiu, Xuanjing Huang
摘要
Most existing deep multi-task learning models are based on parameter sharing, such as hard sharing, hierarchical sharing, and soft sharing. How choosing a suitable sharing mechanism depends on the relations among the tasks, which is not easy since it is difficult to understand the underlying shared factors among these tasks. In this paper, we propose a novel parameter sharing mechanism, named Sparse Sharing. Given multiple tasks, our approach automatically finds a sparse sharing structure. We start with an over-parameterized base network, from which each task extracts a subnetwork. The subnetworks of multiple tasks are partially overlapped and trained in parallel. We show that both hard sharing and hierarchical sharing can be formulated as particular instances of the sparse sharing framework. We conduct extensive experiments on three sequence labeling tasks. Compared with single-task models and three typical multi-task learning baselines, our proposed approach achieves consistent improvement while requiring fewer parameters.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper18
- Progressive Multi-task Learning with Controlled Information Flow for Joint Entity and Relation ExtractionKai Sun, Richong Zhang, Samuel Mensah, Yongyi Mao 等AAAI 2021 · 被引用 50 次
- Finding Sparse Structures for Domain Specific Neural Machine TranslationJianze Liang, Chengqi Zhao, Mingxuan Wang, Xipeng Qiu 等AAAI 2021 · 被引用 33 次
- A Contrastive Sharing Model for Multi-Task RecommendationTing Bai, Yudong Xiao, Bin Wu, Guojun Yang 等WWW 2022 · 被引用 30 次
- Multi-task Graph Neural Architecture Search with Task-aware Collaboration and CurriculumYijian Qin, Xin Wang, Ziwei Zhang, Hong Chen 等NeurIPS 2023 · 被引用 27 次
- Unsupervised Natural Language Inference via Decoupled Multimodal Contrastive LearningWanyun Cui, Guangyu Zheng, Wei WangEMNLP 2020 · 被引用 17 次
相关 Paper
- Learning Multi-Task Sparse Representation Based on Fisher InformationYayu Zhang, Yuhua Qian, Guoshuai Ma, Keyin Zheng 等AAAI 2024 · 被引用 4 次
- AdaShare: Learning What To Share For Efficient Deep Multi-Task LearningXimeng Sun, Rameswar Panda, Rogério Feris, Kate SaenkoNeurIPS 2020 · 被引用 337 次
- DiSparse: Disentangled Sparsification for Multitask Model CompressionXinglong Sun, Ali Hassani, Zhangyang Wang, Gao Huang 等CVPR 2022 · 被引用 22 次
- Multi-Task Recurrent Modular NetworksDongkuan Xu, Wei Cheng, Xin Dong, Bo Zong 等AAAI 2021 · 被引用 2 次
- Can Small Heads Help? Understanding and Improving Multi-Task GeneralizationYuyan Wang, Zhe Zhao, Bo Dai, Christopher Fifty 等WWW 2022 · 被引用 15 次
