Learning from Noisy Labels via Self-Taught On-the-Fly Meta Loss Rescaling
Michael Heck, Christian Geishauser, Nurul Lubis, Carel van Niekerk, Shutong Feng, Hsien-Chin Lin, Benjamin Matthias Ruppik, Renato Vukovic, Milica Gasic
Abstract
Correct labels are indispensable for training effective machine learning models. However, creating high-quality labels is expensive, and even professionally labeled data contains errors and ambiguities. Filtering and denoising can be applied to curate labeled data prior to training, at the cost of additional processing and loss of information. An alternative is on-the-fly sample reweighting during the training process to decrease the negative impact of incorrect or ambiguous labels, but this typically requires clean seed data. In this work we propose unsupervised on-the-fly meta loss rescaling to reweight training samples. Crucially, we rely only on features provided by the model being trained, to learn a rescaling function in real time without knowledge of the true clean data distribution. We achieve this via a novel meta learning setup that samples validation data for the meta update directly from the noisy training corpus by employing the rescaling function being trained. Our proposed method consistently improves performance across various NLP tasks with minimal computational overhead. Further, we are among the first to attempt on-the-fly training data reweighting on the challenging task of dialogue modeling, where noisy and ambiguous labels are common. Our strategy is robust in the face of noisy and clean data, handles class imbalance, and prevents overfitting to noisy labels. Our self-taught loss rescaling improves as the model trains, showing the ability to keep learning from the model's own signals. As training progresses, the impact of correctly labeled data is scaled up, while the impact of wrongly labeled data is suppressed.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e1ba46f9-c41d-4724-bdcf-8ad9b36387c8Builds on2
Related papers
- Meta Self-training for Few-shot Neural Sequence LabelingYaqing Wang, Subhabrata Mukherjee, Haoda Chu, Yuancheng Tu et al.KDD 2021 · 56 citations
- A Model-Agnostic Approach for Learning with Noisy Labels of Arbitrary DistributionsShuang Hao, Peng Li, Renzhi Wu, Xu ChuICDE 2022 · 2 citations
- Meta Label Correction for Noisy Label LearningGuoqing Zheng, Ahmed Hassan Awadallah, Susan T. DumaisAAAI 2021 · 239 citations
- Learning to Purify Noisy Labels via Meta Soft Label CorrectorYichen Wu, Jun Shu, Qi Xie, Qian Zhao et al.AAAI 2021 · 86 citations
- Data Manipulation: Towards Effective Instance Learning for Neural Dialogue Generation via Learning to Augment and ReweightHengyi Cai, Hongshen Chen, Yonghao Song, Cheng Zhang et al.ACL 2020 · 57 citations
