Exploring Relational Context for Multi-Task Dense Prediction
David Brüggemann, Menelaos Kanakis, Anton Obukhov, Stamatios Georgoulis, Luc Van Gool
Abstract
The timeline of computer vision research is marked with advances in learning and utilizing efficient contextual representations. Most of them, however, are targeted at improving model performance on a single downstream task. We consider a multi-task environment for dense prediction tasks, represented by a common backbone and independent task-specific heads. Our goal is to find the most efficient way to refine each task prediction by capturing cross-task contexts dependent on tasks’ relations. We explore various attention-based contexts, such as global and local, in the multi-task setting and analyze their behavior when applied to refine each task independently. Empirical findings confirm that different source-target task pairs benefit from different context types. To automate the selection process, we propose an Adaptive Task-Relational Context (ATRC) module, which samples the pool of all available contexts for each task pair using neural architecture search and outputs the optimal configuration for deployment. Our method achieves state-of-the-art performance on two important multi-task benchmarks, namely NYUD-v2 and PASCAL-Context. The proposed ATRC has a low computational toll and can be used as a drop-in refinement module for any supervised multi-task architecture.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b8226cc2-6141-4c32-b4e5-1e3033c39d01Cited by top-tier papers31
- VMT-Adapter: Parameter-Efficient Transfer Learning for Multi-Task Dense Scene UnderstandingYi Xin, Junlong Du, Qiang Wang, Zhiwen Lin et al.AAAI 2024 · 94 citations
- DeMT: Deformable Mixer Transformer for Multi-Task Learning of Dense PredictionYangyang Xu, Yibo Yang, Lefei ZhangAAAI 2023 · 81 citations
- TaskExpert: Dynamically Assembling Multi-Task Representations with Memorial Mixture-of-ExpertsHanrong Ye, Dan XuICCV 2023 · 60 citations
- Multi-Task Dense Prediction via Mixture of Low-Rank ExpertsYuqi Yang, Peng-Tao Jiang, Qibin Hou, Hao Zhang et al.CVPR 2024 · 32 citations
- Vision Transformer Adapters for Generalizable Multitask LearningDeblina Bhattacharjee, Sabine Süsstrunk, Mathieu SalzmannICCV 2023 · 18 citations
Builds on6
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Attention Augmented Convolutional NetworksIrwan Bello, Barret Zoph, Quoc Le, Ashish Vaswani et al.ICCV 2019 · 1,149 citations
- ACFNet: Attentional Class Feature Network for Semantic SegmentationFan Zhang, Yanqin Chen, Zhihang Li, Zhibin Hong et al.ICCV 2019 · 297 citations
- DSNAS: Direct Neural Architecture Search Without Parameter RetrainingShoukang Hu, Sirui Xie, Hehui Zheng, Chunxiao Liu et al.CVPR 2020
- MTL-NAS: Task-Agnostic Neural Architecture Search Towards General-Purpose Multi-Task LearningYuan Gao, Haoping Bai, Zequn Jie, Jiayi Ma et al.CVPR 2020
Related papers
- Task-Conditional Adapter for Multi-Task Dense PredictionFengze Jiang, Shuling Wang, Xiaojin GongACM MM 2024 · 6 citations
- Contrastive Multi-Task Dense PredictionSiwei Yang, Hanrong Ye, Dan XuAAAI 2023 · 13 citations
- HR-NAS: Searching Efficient High-Resolution Neural Architectures With Lightweight TransformersMingyu Ding, Xiaochen Lian, Linjie Yang, Peng Wang et al.CVPR 2021
- Multi-Task Label Discovery via Hierarchical Task Tokens for Partially Annotated Dense PredictionsJingdong Zhang, Hanrong Ye, Xin Li, Wenping Wang et al.ACM MM 2025 · 1 citation
- Going Beyond Multi-Task Dense Prediction with Synergy Embedding ModelsHuimin Huang, Yawen Huang, Lanfen Lin, Ruofeng Tong et al.CVPR 2024
