Dist Loss: Enhancing Regression in Few-Shot Region through Distribution Distance Constraint
Guangkun Nie, Gongzheng Tang, Shenda Hong
Abstract
Imbalanced data distributions are prevalent in real-world scenarios, presenting significant challenges in both classification and regression tasks. This imbalance often causes deep learning models to overfit in regions with abundant data (manyshot regions) while underperforming in regions with sparse data (few-shot regions). Such characteristics limit the applicability of deep learning models across various domains, notably in healthcare, where rare cases often carry greater clinical significance. While recent studies have highlighted the benefits of incorporating distributional information in imbalanced classification tasks, similar strategies have been largely unexplored in imbalanced regression. To address this gap, we propose Dist Loss, a novel loss function that integrates distributional information into model training by jointly optimizing the distribution distance between model predictions and target labels, alongside sample-wise prediction errors. This dual-objective approach encourages the model to balance its predictions across different label regions, leading to significant improvements in accuracy in fewshot regions. We conduct extensive experiments across three datasets spanning computer vision and healthcare: IMDB-WIKI-DIR, AgeDB-DIR, and ECG-K-DIR. The results demonstrate that Dist Loss effectively mitigates the impact of imbalanced data distributions, achieving state-of-the-art performance in few-shot regions. Furthermore, Dist Loss is easy to integrate and complements existing methods. To facilitate further research, we provide our implementation at https://github.com/Ngk03/DIR-Dist-Loss.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0c166102-ba5f-4dfe-b9ed-80ba3f836ac5Cited by top-tier papers1
Ask how each one uses itBuilds on9
- Distributionally Robust Neural NetworksShiori Sagawa, Pang Wei Koh, Tatsunori B. Hashimoto, Percy LiangICLR 2020 · 1,578 citations
- Delving into Deep Imbalanced RegressionYuzhe Yang, Kaiwen Zha, Ying-Cong Chen, Hao Wang et al.ICML 2021 · 385 citations
- Fast Differentiable Sorting and RankingMathieu Blondel, Olivier Teboul, Quentin Berthet, Josip DjolongaICML 2020 · 285 citations
- Balanced MSE for Imbalanced Visual RegressionJiawei Ren, Mingyuan Zhang, Cunjun Yu, Ziwei LiuCVPR 2022 · 163 citations
- Posterior Re-calibration for Imbalanced DatasetsJunjiao Tian, Yen-Cheng Liu, Nathaniel Glaser, Yen-Chang Hsu et al.NeurIPS 2020 · 86 citations
Related papers
- A step towards understanding why classification helps regressionSilvia L. Pintea, Yancong Lin, Jouke Dijkstra, Jan C. van GemertICCV 2023 · 17 citations
- RankSim: Ranking Similarity Regularization for Deep Imbalanced RegressionYu Gong, Greg Mori, Frederick TungICML 2022 · 68 citations
- Leveraging Group Classification with Descending Soft Labeling for Deep Imbalanced RegressionRuizhi Pu, Gezheng Xu, Ruiyi Fang, Bingkun Bao et al.AAAI 2025 · 7 citations
- Variational Imbalanced Regression: Fair Uncertainty Quantification via Probabilistic SmoothingZiyan Wang, Hao WangNeurIPS 2023 · 7 citations
- Deep Imbalanced Regression via Hierarchical Classification AdjustmentHaipeng Xiong, Angela YaoCVPR 2024 · 8 citations
