Transfer Vision Patterns for Multi-Task Pixel Learning
Xiaoya Zhang, Ling Zhou, Yong Li, Zhen Cui, Jin Xie, Jian Yang
摘要
Multi-task pixel perception is one of the most important topics in the field of machine intelligence. Inspired by the observation of cross-task interdependencies of visual patterns, we propose a multi-task vision pattern transformation (VPT) method to adaptively correlate and transfer cross-task visual patterns by leveraging the powerful transformer mechanism. To better transfer visual patterns, specifically, we build two types of pattern transformation based on the statistic prior that the affinity relations across tasks are correlated. One aims to transfer feature patterns for the integration of different task features; the other aims to exchange structure patterns for mining and leveraging the latent interaction cues. These two types of transformations are encapsulated into two VPT units, which provide universal matching interfaces for multi-task learning, complement each other to guide the transmission of feature/structure patterns, and finally realize an adaptive selection of important patterns across tasks. Extensive experiments on the joint learning of semantic segmentation, depth prediction and surface normal estimation demonstrate that our proposed method is more effective than those baselines and achieve the state-of-that-art performance in three pixel-level visual tasks.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
引用它的顶会 Paper6
- MmAP: Multi-Modal Alignment Prompt for Cross-Domain Multi-Task LearningYi Xin, Junlong Du, Qiang Wang, Ke Yan 等AAAI 2024 · 被引用 102 次
- VMT-Adapter: Parameter-Efficient Transfer Learning for Multi-Task Dense Scene UnderstandingYi Xin, Junlong Du, Qiang Wang, Zhiwen Lin 等AAAI 2024 · 被引用 94 次
- TaskExpert: Dynamically Assembling Multi-Task Representations with Memorial Mixture-of-ExpertsHanrong Ye, Dan XuICCV 2023 · 被引用 60 次
- Fedhca2: Towards Hetero-Client Federated Multi-Task LearningYuxiang Lu, Suizhi Huang, Yuwen Yang, Shalayiding Sirejiding 等CVPR 2024 · 被引用 13 次
- Efficient Computation Sharing for Multi-Task Visual Scene UnderstandingSara Shoouri, Mingyu Yang, Zichen Fan, Hun-Seok KimICCV 2023 · 被引用 9 次
相关 Paper
- MuIT: An End-to-End Multitask Learning TransformerDeblina Bhattacharjee, Tong Zhang, Sabine Süsstrunk, Mathieu SalzmannCVPR 2022 · 被引用 64 次
- Multi-task Learning with 3D-Aware RegularizationWei-Hong Li, Steven McDonagh, Ales Leonardis, Hakan BilenICLR 2024 · 被引用 10 次
- Pattern-Structure Diffusion for Multi-Task LearningLing Zhou, Zhen Cui, Chunyan Xu, Zhenyu Zhang 等CVPR 2020
- Going Beyond Multi-Task Dense Prediction with Synergy Embedding ModelsHuimin Huang, Yawen Huang, Lanfen Lin, Ruofeng Tong 等CVPR 2024
- Any Resolution Any Geometry: From Multi-View To Multi-PatchWenqing Cui, Zhenyu Li, Mykola Lavreniuk, Jian Shi 等CVPR 2026 · 被引用 2 次
