Rethinking Guidance Information to Utilize Unlabeled Samples: A Label Encoding Perspective
Yulong Zhang, Yuan Yao, Shuhao Chen, Pengrong Jin, Yu Zhang, Jian Jin, Jiangang Lu
Abstract
Empirical Risk Minimization (ERM) is fragile in scenarios with insufficient labeled samples. A vanilla extension of ERM to unlabeled samples is Entropy Minimization (EntMin), which employs the soft-labels of unlabeled samples to guide their learning. However, EntMin emphasizes prediction discriminability while neglecting prediction diversity. To alleviate this issue, in this paper, we rethink the guidance information to utilize unlabeled samples. By analyzing the learning objective of ERM, we find that the guidance information for labeled samples in a specific category is the corresponding label encoding. Inspired by this finding, we propose a Label-Encoding Risk Minimization (LERM). It first estimates the label encodings through prediction means of unlabeled samples and then aligns them with their corresponding ground-truth label encodings. As a result, the LERM ensures both prediction discriminability and diversity, and it can be integrated into existing methods as a plugin. Theoretically, we analyze the relationships between LERM and ERM as well as EntMin. Empirically, we verify the superiority of the LERM under several label insufficient scenarios. The codes are available at https://github.com/zhangyl660/LERM .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 49bdd3c5-1abe-4fa2-a8d7-adb6911643a5Cited by top-tier papers4
- Time-Varying LoRA: Towards Effective Cross-Domain Fine-Tuning of Diffusion ModelsZhan Zhuang, Yulong Zhang, Xuehao Wang, Jiangang Lu et al.NeurIPS 2024 · 18 citations
- Generalized Category Discovery via Reciprocal Learning and Class-Wise Distribution RegularizationDuo Liu, Zhiquan Tan, Linglan Zhao, Zhongqiang Zhang et al.ICML 2025
- Rethinking the Flow-based Gradual Domain Adaptation: A Semi-Dual Optimal Transport PerspectiveZhichao Chen, Zhan Zhuang, Yunfei Teng, Hao Wang et al.ICML 2026
- Semi-Supervised Noise Adaptation: Transferring Knowledge from Noise DomainYuan Yao, Jin Song, Huixia Li, Tongtong Yuan et al.ICML 2026
Builds on13
- RandAugment: Practical Automated Data Augmentation with a Reduced Search SpaceEkin Dogus Cubuk, Barret Zoph, Jonathon Shlens, Quoc LeNeurIPS 2020 · 4,453 citations
- The Many Faces of Robustness: A Critical Analysis of Out-of-Distribution GeneralizationDan Hendrycks, Steven Basart, Norman Mu, Saurav Kadavath et al.ICCV 2021 · 2,294 citations
- Do We Really Need to Access the Source Data? Source Hypothesis Transfer for Unsupervised Domain AdaptationJian Liang, Dapeng Hu, Jiashi FengICML 2020 · 1,624 citations
- FlexMatch: Boosting Semi-Supervised Learning with Curriculum Pseudo LabelingBowen Zhang, Yidong Wang, Wenxin Hou, Hao Wu et al.NeurIPS 2021 · 1,389 citations
- Larger Norm More Transferable: An Adaptive Feature Norm Approach for Unsupervised Domain AdaptationRuijia Xu, Guanbin Li, Jihan Yang, Liang LinICCV 2019 · 563 citations
Related papers
- Heterogeneous Risk MinimizationJiashuo Liu, Zheyuan Hu, Peng Cui, Bo Li et al.ICML 2021 · 170 citations
- A Generalized Unbiased Risk Estimator for Learning with Augmented ClassesSenlin Shu, Shuo He, Haobo Wang, Hongxin Wei et al.AAAI 2023 · 4 citations
- Entropy-based Optimization on Individual and Global Predictions for Semi-Supervised LearningZhen Zhao, Meng Zhao, Ye Liu, Di Yin et al.ACM MM 2023 · 4 citations
- Dist-PU: Positive-Unlabeled Learning from a Label Distribution PerspectiveYunrui Zhao, Qianqian Xu, Yangbangyan Jiang, Peisong Wen et al.CVPR 2022 · 47 citations
- A Closer Look to Positive-Unlabeled Learning from Fine-grained Perspectives: An Empirical StudyYuanchao Dai, Zhengzhang Hou, Changchun Li, Yuanbo Xu et al.NeurIPS 2025 · 2 citations
