Binary Losses for Density Ratio Estimation
Werner Zellinger
Abstract
Estimating the ratio of two probability densities from a finite number of observations is a central machine learning problem. A common approach is to construct estimators using binary classifiers that distinguish observations from the two densities. However, the accuracy of these estimators depends on the choice of the binary loss function, raising the question of which loss function to choose based on desired error properties. For example, traditional loss functions, such as logistic or boosting loss, prioritize accurate estimation of small density ratio values over large ones, even though the latter are more critical in many applications. In this work, we start with prescribed error measures in a class of Bregman divergences and characterize all loss functions that result in density ratio estimators with small error. Our characterization extends results on composite binary losses from Reid & Williamson (2010) and their connection to density ratio estimation as identified by Menon & Ong (2016) . As a result, we obtain a simple recipe for constructing loss functions with certain properties, such as those that prioritize an accurate estimation of large density ratio values. Our novel loss functions outperform related approaches for resolving parameter choice issues of 11 deep domain adaptation algorithms in average performance across 484 real-world tasks including sensor signals, texts, and images. 2 Denote by y = 1 the event "drawn from P " and by y = -1 the event "drawn from Q". Then, the probability measure ρ on X × -1, 1 is defined by ρ(x|y = 1) := P (x), ρ(x|y = -1) := Q(x) and the marginal (w.r.t. y) Bernoulli measure ρY realizing y = 1 with probability 1 2 and y = -1 otherwise. It follows that β(x) = ρ(x|y=1) ρ(x|y=-1) = ρ(x,y=1)ρ Y (y=-1) ρ(x,y=-1)ρ Y (y=1) = ρ(x,y=1) ρ(x,y=-1) = ρ(x,y=1)ρ X (x) ρ(x,y=-1)ρ X (x) = ρ(y=1|x) ρ(y=-1|x) .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on9
- Moment Matching for Multi-Source Domain AdaptationXingchao Peng, Qinxun Bai, Xide Xia, Zijun Huang et al.ICCV 2019 · 2,239 citations
- Telescoping Density-Ratio EstimationBenjamin Rhodes, Kai Xu, Michael U. GutmannNeurIPS 2020 · 148 citations
- Refining Generative Process with Discriminator Guidance in Score-based Diffusion ModelsDongjun Kim, Yeongmin Kim, Se Jung Kwon, Wanmo Kang et al.ICML 2023 · 109 citations
- Non-Negative Bregman Divergence Minimization for Deep Direct Density Ratio EstimationMasahiro Kato, Takeshi TeshimaICML 2021 · 53 citations
- Training Unbiased Diffusion Models From Biased DatasetYeongmin Kim, Byeonghu Na, Minsang Park, JoonHo Jang et al.ICLR 2024 · 37 citations
Related papers
- Overcoming Saturation in Density Ratio Estimation by Iterated RegularizationLukas Gruber, Markus Holzleitner, Johannes Lehner, Sepp Hochreiter et al.ICML 2024 · 6 citations
- Bounds on Lp Errors in Density Ratio Estimation via f-Divergence Loss FunctionsYoshiaki KitazawaICLR 2025
- Meta-Learning for Relative Density-Ratio EstimationAtsutoshi Kumagai, Tomoharu Iwata, Yasuhiro FujiwaraNeurIPS 2021 · 13 citations
- Density Ratio Estimation with Doubly Strong RobustnessRyosuke Nagumo, Hironori FujisawaICML 2024 · 8 citations
- Variational Weighting for Kernel Density RatiosSangwoong Yoon, Frank C. Park, Gunsu S. Yun, Iljung Kim et al.NeurIPS 2023 · 1 citation
