Learning from Noisy Labels via Conditional Distributionally Robust Optimization
Hui Guo, Grace Y. Yi, Boyu Wang
摘要
While crowdsourcing has emerged as a practical solution for labeling large datasets, it presents a significant challenge in learning accurate models due to noisy labels from annotators with varying levels of expertise. Existing methods typically estimate the true label posterior, conditioned on the instance and noisy annotations, to infer true labels or adjust loss functions. These estimates, however, often overlook potential misspecification in the true label posterior, which can degrade model performances, especially in high-noise scenarios. To address this issue, we investigate learning from noisy annotations with an estimated true label posterior through the framework of conditional distributionally robust optimization (CDRO). We propose formulating the problem as minimizing the worst-case risk within a distance-based ambiguity set centered around a reference distribution. By examining the strong duality of the formulation, we derive upper bounds for the worst-case risk and develop an analytical solution for the dual robust risk for each data point. This leads to a novel robust pseudo-labeling algorithm that leverages the likelihood ratio test to construct a pseudo-empirical distribution, providing a robust reference probability distribution in CDRO. Moreover, to devise an efficient algorithm for CDRO, we derive a closed-form expression for the empirical robust risk and the optimal Lagrange multiplier of the dual problem, facilitating a principled balance between robustness and model fitting. Our experimental results on both synthetic and real-world datasets demonstrate the superiority of our method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- When Priors Backfire: On the Vulnerability of Unlearnable Examples to PretrainingZhihao Li, Gezheng Xu, Jiale Cai, Ruiyi Fang 等ICLR 2026 · 被引用 5 次
- FUSE: Full‑spectrum Unlearnable Examples via Spectral EqualizationJiale Cai, Gezheng Xu, Zhihao Li, Ruiyi Fang 等ICML 2026 · 被引用 1 次
- Discretized Density-Guided Source-Free Adaptation for Continuous TargetsGezheng Xu, Qi CHEN, QIUHAO Zeng, Charles X. Ling 等ICML 2026
- Revisiting Source-Free Domain Adaptation: a New Perspective via Uncertainty ControlGezheng Xu, Hui Guo, Li Yi, Charles Ling 等ICLR 2025
- Optimal Fair Aggregation of Crowdsourced Noisy Labels using Demographic Parity ConstraintsGabriel Singer, Samuel Gruffaz, Olivier VO VAN, Nicolas Vayatis 等ICML 2026
它引用的顶会 Paper11
- Learning with Noisy Labels Revisited: A Study Using Real-World Human AnnotationsJiaheng Wei, Zhaowei Zhu, Hao Cheng, Tongliang Liu 等ICLR 2022 · 被引用 338 次
- Part-dependent Label Noise: Towards Instance-dependent Label NoiseXiaobo Xia, Tongliang Liu, Bo Han, Nannan Wang 等NeurIPS 2020 · 被引用 329 次
- Error-Bounded Correction of Noisy LabelsSongzhu Zheng, Pengxiang Wu, Aman Goswami, Mayank Goswami 等ICML 2020 · 被引用 153 次
- Combating Noisy Labels with Sample Selection by Mining High-Discrepancy ExamplesXiaobo Xia, Bo Han, Yibing Zhan, Jun Yu 等ICCV 2023 · 被引用 72 次
- Learning from Crowds by Modeling Common ConfusionsZhendong Chu, Jing Ma, Hongning WangAAAI 2021 · 被引用 60 次
相关 Paper
- Distributionally Robust Optimization with Probabilistic GroupSoumya Suvra Ghosal, Yixuan LiAAAI 2023 · 被引用 14 次
- Outlier-Robust Distributionally Robust Optimization via Unbalanced Optimal TransportZifan Wang, Yi Shen, Michael M. Zavlanos, Karl Henrik JohanssonNeurIPS 2024 · 被引用 16 次
- Soft-Label Integration for Robust Toxicity ClassificationZelei Cheng, Xian Wu, Jiahao Yu, Shuo Han 等NeurIPS 2024 · 被引用 7 次
- Distributed Distributionally Robust Optimization with Non-Convex ObjectivesYang Jiao, Kai Yang, Dongjin SongNeurIPS 2022 · 被引用 21 次
- Toward Robust Neural Reconstruction from Sparse Point SetsAmine Ouasfi, Shubhendu Jena, Éric Marchand, Adnane BoukhaymaCVPR 2025
