Distribution Matching for Multi-Task Learning of Classification Tasks: A Large-Scale Study on Faces & Beyond
Dimitrios Kollias, Viktoriia Sharmanska, Stefanos Zafeiriou
Abstract
Multi-Task Learning (MTL) is a framework, where multiple related tasks are learned jointly and benefit from a shared representation space, or parameter transfer. To provide sufficient learning support, modern MTL uses annotated data with full, or sufficiently large overlap across tasks, i.e., each input sample is annotated for all, or most of the tasks. However, collecting such annotations is prohibitive in many real applications, and cannot benefit from datasets available for individual tasks. In this work, we challenge this setup and show that MTL can be successful with classification tasks with little, or non-overlapping annotations, or when there is big discrepancy in the size of labeled data per task. We explore task-relatedness for co-annotation and co-training, and propose a novel approach, where knowledge exchange is enabled between the tasks via distribution matching. To demonstrate the general applicability of our method, we conducted diverse case studies in the domains of affective computing, face recognition, species recognition, and shopping item classification using nine datasets. Our large-scale study of affective tasks for basic expression recognition and facial action unit detection illustrates that our approach is network agnostic and brings large performance improvements compared to the state-of-the-art in both tasks and across all studied databases. In all case studies, we show that co-training via task-relatedness is advantageous and prevents negative transfer (which occurs when MT model's performance is worse than that of at least one single-task model).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9504bcf1-c81d-478e-a408-b636708cf00bCited by top-tier papers4
- BD-Merging: Bias-Aware Dynamic Model Merging with Evidence-Guided Contrastive LearningYuhan Xie, Chen LyuCVPR 2026
- Fair Facial Attribute Recognition via Group-Decoupled Vision Transformer with Mask-Guided Correlation SuppressionHuichang Huang, Kunchi Li, Si Chen, Da-Han WangAAAI 2026
- In-Context Adaptation to Concept Drift for Learned Database OperationsJiaqi Zhu, Shaofeng Cai, Yanyan Shen, Gang Chen et al.ICML 2025
- FairMerging: Rethinking Model Merging through the Lens of FairnessBing Liu, Xinrui Shan, Boyu Zhang, Qiankun Zhang et al.ICML 2026
Builds on4
- AdaShare: Learning What To Share For Efficient Deep Multi-Task LearningXimeng Sun, Rameswar Panda, Rogério Feris, Kate SaenkoNeurIPS 2020 · 337 citations
- Understanding and Improving Information Transfer in Multi-Task LearningSen Wu, Hongyang R. Zhang, Christopher RéICLR 2020 · 183 citations
- Many Task Learning With Task RoutingGjorgji Strezoski, Nanne van Noord, Marcel WorringICCV 2019 · 112 citations
- RetinaFace: Single-Shot Multi-Level Face Localisation in the WildJiankang Deng, Jia Guo, Evangelos Ververas, Irene Kotsia et al.CVPR 2020
Related papers
- Multi-Label Compound Expression Recognition: C-EXPR Database & NetworkDimitrios KolliasCVPR 2023
- Selective Task Group Updates for Multi-Task OptimizationWooseong Jeong, Kuk-Jin YoonICLR 2025
- Identifying and Mitigating Spurious Correlation in Multi-Task LearningJunyi Chai, Shenyu Lu, Xiaoqian WangCVPR 2025
- Data-Scarce Animal Face Alignment via Bi-Directional Cross-Species Knowledge TransferDan Zeng, Shanchuan Hong, Shuiwang Li, Qiaomu Shen et al.ACM MM 2023 · 4 citations
- When Does Aggregating Multiple Skills with Multi-Task Learning Work? A Case Study in Financial NLPJingwei Ni, Zhijing Jin, Qian Wang, Mrinmaya Sachan et al.ACL 2023 · 2 citations
