Content Diversity-guided Ambiguity Mitigation for Open-Set Noisy Label Learning
Zhihao Zhou, Rui Li, Xueying Li
Abstract
Open-set noisy label learning faces a critical challenge in maintaining robust DNN performance when training data contain both in-distribution noisy (IDN) and out-of-distribution (OOD) samples. These noisy samples induce overconfident but erroneous predictions due to their ambiguous positions relative to category boundaries. Current methods address this by filtering noisy samples based on visual features alone, they fail to resolve the semantic ambiguity near decision boundaries, where limited visual cues lead to unreliable sample purification. To this end, we propose Content Diversity-guided Ambiguity Mitigation (CDgAM), a novel framework that leverages diverse contents to mitigate visual ambiguity in open-set noisy label learning. CDgAM leverages textual descriptions of intra-class commonality and inter-class disparity to dynamically refine semantic boundaries, reducing bias in prototype learning. To further suppress early-stage uncertainty in visual representations, we design a region-sensitive distillation regularization that transfers boundary-aware knowledge from a multimodal large language model to the target DNN. Extensive experiments conducted on various datasets with different noise levels demonstrate the effectiveness of our CDgAM, outperforming state-of-the-art methods for open-set noisy label learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext bbfeec0c-7214-44d2-8ef7-9f5b3c229eceBuilds on19
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and GenerationJunnan Li, Dongxu Li, Caiming Xiong, Steven C. H. HoiICML 2022 · 6,549 citations
- DivideMix: Learning with Noisy Labels as Semi-supervised LearningJunnan Li, Richard Socher, Steven C. H. HoiICLR 2020 · 1,326 citations
- Learning from Noisy Data with Robust Representation LearningJunnan Li, Caiming Xiong, Steven C. H. HoiICCV 2021 · 140 citations
- GLaMM: Pixel Grounding Large Multimodal ModelHanoona Abdul Rasheed, Muhammad Maaz, Sahal Shaji Mullappilly, Abdelrahman M. Shaker et al.CVPR 2024 · 113 citations
Related papers
- Open-set Label Noise Can Improve Robustness Against Inherent Label NoiseHongxin Wei, Lue Tao, Renchunzi Xie, Bo AnNeurIPS 2021 · 113 citations
- ROG_PL: Robust Open-Set Graph Learning via Region-Based Prototype LearningQin Zhang, Xiaowei Li, Jiexin Lu, Liping Qiu et al.AAAI 2024 · 2 citations
- Pre-Trained Vision-Language Models as Noisy Partial AnnotatorsQian-Wei Wang, Yuqiu Xie, Letian Zhang, Zimo Liu et al.AAAI 2025 · 3 citations
- Bridging the Unseen Gap: Label-Enhanced Information Bottleneck Distillation for Multimodal Named Entity RecognitionBo Xu, Jie Wei, Hongya Wang, Ming Du et al.ACM MM 2025 · 2 citations
- Incremental Label Distribution LearningChao Xu, Xijia Tang, Hong Tao, Chenping HouKDD 2025
