MarginMatch: Improving Semi-Supervised Learning with Pseudo-Margins
Tiberiu Sosea, Cornelia Caragea
Abstract
We introduce MarginMatch, a new SSL approach combining consistency regularization and pseudo-labeling, with its main novelty arising from the use of unlabeled data training dynamics to measure pseudo-label quality. Instead of using only the model's confidence on an unlabeled example at an arbitrary iteration to decide if the example should be masked or not, MarginMatch also analyzes the behavior of the model on the pseudo-labeled examples as the training progresses, to ensure low quality predictions are masked out. MarginMatch brings substantial improvements on four vision benchmarks in low data regimes and on two large-scale datasets, emphasizing the importance of enforcing highquality pseudo-labels. Notably, we obtain an improvement in error rate over the state-of-the-art of 3.25% on CIFAR-100 with only 25 labels per class and of 3.78% on STL-10 using as few as 4 labels per class. We make our code available at https://github.com/tsosea2/MarginMatch .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f2f64563-6bb4-4427-ae99-fb9333bab092Cited by top-tier papers8
- Self-Reinforcing Prototype Evolution with Dual-Knowledge Cooperation for Semi-Supervised Lifelong Person Re-IdentificationKunlun Xu, Fan Zhuo, Jiangmeng Li, Xu Zou et al.ICCV 2025 · 2 citations
- Bi-Level Optimization for Semi-Supervised Learning with Pseudo-LabelingMarzi Heidari, Yuhong GuoAAAI 2025 · 1 citation
- MultiMatch: Multihead Consistency Regularization Matching for Semi-Supervised Text ClassificationIustin Sirbu, Robert-Adrian Popovici, Cornelia Caragea, Stefan Trausan-Matu et al.EMNLP 2025 · 1 citation
- CGMatch: A Different Perspective of Semi-supervised LearningBo Cheng, Jueqing Lu, Yuan Tian, Haifeng Zhao et al.CVPR 2025
- Human-Corrected Labels Learning: Enhancing Labels Quality via Human Correction of VLMs DiscrepanciesZhongnian Li, Lan Chen, Yixin Xu, Shi Xu et al.AAAI 2026
Builds on9
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang et al.NeurIPS 2020 · 5,129 citations
- RandAugment: Practical Automated Data Augmentation with a Reduced Search SpaceEkin Dogus Cubuk, Barret Zoph, Jonathon Shlens, Quoc LeNeurIPS 2020 · 4,453 citations
- Unsupervised Data Augmentation for Consistency TrainingQizhe Xie, Zihang Dai, Eduard H. Hovy, Thang Luong et al.NeurIPS 2020 · 2,774 citations
- FlexMatch: Boosting Semi-Supervised Learning with Curriculum Pseudo LabelingBowen Zhang, Yidong Wang, Wenxin Hou, Hao Wu et al.NeurIPS 2021 · 1,389 citations
Related papers
- RegMixMatch: Optimizing Mixup Utilization in Semi-Supervised LearningHaorong Han, Jidong Yuan, Chixuan Wei, Zhongyang YuAAAI 2025 · 7 citations
- Boosting Semi-Supervised Learning by Exploiting All Unlabeled DataYuhao Chen, Xin Tan, Borui Zhao, Zhaowei Chen et al.CVPR 2023
- HyperMatch: Noise-Tolerant Semi-Supervised Learning via Relaxed Contrastive ConstraintBeitong Zhou, Jing Lu, Kerui Liu, Yunlu Xu et al.CVPR 2023
- PseudoSeg: Designing Pseudo Labels for Semantic SegmentationYuliang Zou, Zizhao Zhang, Han Zhang, Chun-Liang Li et al.ICLR 2021 · 364 citations
- SoftMatch: Addressing the Quantity-Quality Tradeoff in Semi-supervised LearningHao Chen, Ran Tao, Yue Fan, Yidong Wang et al.ICLR 2023
