Heteroskedastic and Imbalanced Deep Learning with Adaptive Regularization
Kaidi Cao, Yining Chen, Junwei Lu, Nikos Aréchiga, Adrien Gaidon, Tengyu Ma
摘要
Real-world large-scale datasets are heteroskedastic and imbalanced -- labels have varying levels of uncertainty and label distributions are long-tailed. Heteroskedasticity and imbalance challenge deep learning algorithms due to the difficulty of distinguishing among mislabeled, ambiguous, and rare examples. Addressing heteroskedasticity and imbalance simultaneously is under-explored. We propose a data-dependent regularization technique for heteroskedastic datasets that regularizes different regions of the input space differently. Inspired by the theoretical derivation of the optimal regularization strength in a one-dimensional nonparametric classification setting, our approach adaptively regularizes the data points in higher-uncertainty, lower-density regions more heavily. We test our method on several benchmark tasks, including a real-world heteroskedastic and imbalanced dataset, WebVision. Our experiments corroborate our theory and demonstrate a significant improvement over other methods in noise-robust deep learning.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper29
- WILDS: A Benchmark of in-the-Wild Distribution ShiftsPang Wei Koh, Shiori Sagawa, Henrik Marklund, Sang Michael Xie 等ICML 2021 · 被引用 1,773 次
- Just Train Twice: Improving Group Robustness without Training Group InformationEvan Zheran Liu, Behzad Haghgoo, Annie S. Chen, Aditi Raghunathan 等ICML 2021 · 被引用 683 次
- Improving Out-of-Distribution Robustness via Selective AugmentationHuaxiu Yao, Yu Wang, Sai Li, Linjun Zhang 等ICML 2022 · 被引用 275 次
- Open-World Semi-Supervised LearningKaidi Cao, Maria Brbic, Jure LeskovecICLR 2022 · 被引用 246 次
- Self-supervised Learning is More Robust to Dataset ImbalanceHong Liu, Jeff Z. HaoChen, Adrien Gaidon, Tengyu MaICLR 2022 · 被引用 190 次
它引用的顶会 Paper4
- Dynamic Curriculum Learning for Imbalanced Data ClassificationYiru Wang, Weihao Gan, Jie Yang, Wei Wu 等ICCV 2019 · 被引用 263 次
- Learning with Bounded Instance and Label-dependent Label NoiseJiacheng Cheng, Tongliang Liu, Kotagiri Ramamohanarao, Dacheng TaoICML 2020 · 被引用 162 次
- Simple and Effective Regularization Methods for Training on Noisily Labeled Data with Generalization GuaranteeWei Hu, Zhiyuan Li, Dingli YuICLR 2020 · 被引用 140 次
- BBN: Bilateral-Branch Network With Cumulative Learning for Long-Tailed Visual RecognitionBoyan Zhou, Quan Cui, Xiu-Shen Wei, Zhao-Min ChenCVPR 2020
相关 Paper
- Correlated Input-Dependent Label Noise in Large-Scale Image ClassificationMark Collier, Basil Mustafa, Efi Kokiopoulou, Rodolphe Jenatton 等CVPR 2021
- Targeted Data-driven Regularization for Out-of-Distribution GeneralizationMohammad Mahdi Kamani, Sadegh Farhang, Mehrdad Mahdavi, James Z. WangKDD 2020 · 被引用 10 次
- Robust Learning from Noisily Labeled Long-Tailed Data via Fairness RegularizerJiaheng Wei, Zhaowei Zhu, Gang Niu, Tongliang Liu 等AAAI 2026
- Effective Bayesian Heteroscedastic Regression with Deep Neural NetworksAlexander Immer, Emanuele Palumbo, Alexander Marx, Julia E. VogtNeurIPS 2023 · 被引用 34 次
- Adaptive Logit Adjustment Loss for Long-Tailed Visual RecognitionYan Zhao, Weicong Chen, Xu Tan, Kai Huang 等AAAI 2022 · 被引用 83 次
