More Reliable Pseudo-Labels, Better Performance: A Generalized Approach to Single Positive Multi-Label Learning
Luong Tran, Thieu Vo, Anh Nguyen, Sang Dinh, Van Nguyen
Abstract
Multi-label learning is a challenging computer vision task that requires assigning multiple categories to each image. However, fully annotating large-scale datasets is often impractical due to high costs and effort, motivating the study of learning from partially annotated data. In the extreme case of Single Positive Multi-Label Learning (SPML), each image is provided with only one positive label, while all other labels remain unannotated. Traditional SPML methods that treat missing labels as unknown or negative tend to yield inaccuracies and false negatives, and integrating various pseudo-labeling strategies can introduce additional noise. To address these challenges, we propose the Generalized Pseudo-Label Robust Loss (GPR Loss), a novel loss function that effectively learns from diverse pseudo-labels while mitigating noise. Complementing this, we introduce a simple yet effective Dynamic Augmented Multi-focus Pseudo-labeling (DAMP) technique. Together, these contributions form the Adaptive and Efficient Vision-Language Pseudo-Labeling (AEVLP) framework. Extensive experiments on four benchmark datasets demonstrate that our framework significantly advances multi-label classification, achieving state-of-the-art results.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d1fa9289-4e46-4a58-a435-6b414bdab09bCited by top-tier papers1
Ask how each one uses itBuilds on14
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer et al.CVPR 2022 · 6,782 citations
- FreeLB: Enhanced Adversarial Training for Natural Language UnderstandingChen Zhu, Yu Cheng, Zhe Gan, Siqi Sun et al.ICLR 2020 · 502 citations
- NEFTune: Noisy Embeddings Improve Instruction FinetuningNeel Jain, Ping-yeh Chiang, Yuxin Wen, John Kirchenbauer et al.ICLR 2024 · 120 citations
Related papers
- Revisiting Pseudo-Label for Single-Positive Multi-Label LearningBiao Liu, Ning Xu, Jiaqi Lv, Xin GengICML 2023 · 26 citations
- Label-Aware Global Consistency for Multi-Label Learning with Single Positive LabelsMing-Kun Xie, Jiahao Xiao, Sheng-Jun HuangNeurIPS 2022 · 43 citations
- Partially View-Aligned Representation Learning With Noise-Robust Contrastive LossMouxing Yang, Yunfan Li, Zhenyu Huang, Zitao Liu et al.CVPR 2021
- Partial Multi-Label Learning with Meta DisambiguationMing-Kun Xie, Feng Sun, Sheng-Jun HuangKDD 2021 · 25 citations
- Neighbor-aware Label Refinement: Enhancing Unreliable Instance-Dependent Partial LabelsXijia Tang, Yuhua Qian, Chao Xu, Chenping HouAAAI 2026
