Being Properly Improper
Tyler Sypherd, Richard Nock, Lalitha Sankar
Abstract
Properness for supervised losses stipulates that the loss function shapes the learning algorithm towards the true posterior of the data generating distribution. Unfortunately, data in modern machine learning can be corrupted or twisted in many ways. Hence, optimizing a proper loss function on twisted data could perilously lead the learning algorithm towards the twisted posterior, rather than to the desired clean posterior. Many papers cope with specific twists (e.g., label/feature/adversarial noise), but there is a growing need for a unified and actionable understanding atop properness. Our chief theoretical contribution is a generalization of the properness framework with a notion called twist-properness, which delineates loss functions with the ability to"untwist"the twisted posterior into the clean posterior. Notably, we show that a nontrivial extension of a loss function called -loss, which was first introduced in information theory, is twist-proper. We study the twist-proper -loss under a novel boosting algorithm, called PILBoost, and provide formal and experimental results for this algorithm. Our overarching practical conclusion is that the twist-proper -loss outperforms the proper -loss on several variants of twisted data.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers7
- Boosting with Tempered Exponential MeasuresRichard Nock, Ehsan Amid, Manfred K. WarmuthNeurIPS 2023 · 10 citations
- Fair Wrapping for Black-box PredictionsAlexander Soen, Ibrahim M. Alabdulmohsin, Sanmi Koyejo, Yishay Mansour et al.NeurIPS 2022 · 8 citations
- Generative Trees: Adversarial and CopycatRichard Nock, Mathieu Guillame-BertICML 2022 · 6 citations
- Random Classification Noise does not defeat All Convex Potential Boosters Irrespective of Model ChoiceYishay Mansour, Richard Nock, Robert C. WilliamsonICML 2023 · 4 citations
- Enhancing Robustness of Last Layer Two-Stage Fair Model CorrectionsNathan Stromberg, Rohan Ayyagari, Sanmi Koyejo, Richard Nock et al.NeurIPS 2024 · 3 citations
Builds on4
- Calibrating Deep Neural Networks using Focal LossJishnu Mukhoti, Viveka Kulharia, Amartya Sanyal, Stuart Golodetz et al.NeurIPS 2020 · 674 citations
- Learning Noise Transition Matrix from Only Noisy Labels via Total Variation RegularizationYivan Zhang, Gang Niu, Masashi SugiyamaICML 2021 · 107 citations
- Projection Efficient Subgradient Method and Optimal Nonsmooth Frank-Wolfe MethodKiran Koshy Thekumparampil, Prateek Jain, Praneeth Netrapalli, Sewoong OhNeurIPS 2020 · 31 citations
- Boosting simple learnersNoga Alon, Alon Gonen, Elad Hazan, Shay MoranSTOC 2021 · 2 citations
Related papers
- Lower-Bounded Proper Losses for Weakly Supervised ClassificationShuhei M. Yoshida, Takashi Takenouchi, Masashi SugiyamaICML 2021 · 3 citations
- Robust Minimax Boosting with Performance GuaranteesSantiago Mazuelas, Verónica ÁlvarezNeurIPS 2025
- On the Error Resistance of Hinge-Loss MinimizationKunal TalwarNeurIPS 2020 · 7 citations
- Learning from Noisy Labels with No Change to the Training ProcessMingyuan Zhang, Jane H. Lee, Shivani AgarwalICML 2021 · 38 citations
- Boosting for Predictive SufficiencyAbbavaram Gowtham Reddy, Rajeev Verma, Celia Rubio-Madrigal, Krikamol Muandet et al.ICLR 2026
