Conditional Density Estimation with Histogram Trees
Lincen Yang, Matthijs van Leeuwen
Abstract
Conditional density estimation (CDE) goes beyond regression by modeling the full conditional distribution, providing a richer understanding of the data than just the conditional mean in regression. This makes CDE particularly useful in critical application domains. However, interpretable CDE methods are understudied. Current methods typically employ kernel-based approaches, using kernel functions directly for kernel density estimation or as basis functions in linear models. In contrast, despite their conceptual simplicity and visualization suitability, tree-based methods -- which are arguably more comprehensible -- have been largely overlooked for CDE tasks. Thus, we propose the Conditional Density Tree (CDTree), a fully non-parametric model consisting of a decision tree in which each leaf is formed by a histogram model. Specifically, we formalize the problem of learning a CDTree using the minimum description length (MDL) principle, which eliminates the need for tuning the hyperparameter for regularization. Next, we propose an iterative algorithm that, although greedily, searches the optimal histogram for every possible node split. Our experiments demonstrate that, in comparison to existing interpretable CDE methods, CDTrees are both more accurate (as measured by the log-loss) and more robust against irrelevant features. Further, our approach leads to smaller tree sizes than existing tree-based models, which benefits interpretability.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dddb39dd-9a2b-4f91-91c1-e8a317644507Cited by top-tier papers2
- Continuous multinomial logistic regression for neural decodingRupasinghe Arachchige Anuththara Rupasinghe, Jonathan W. PillowICLR 2026
- Learning Subgroups with Maximum Treatment Effects Without Causal HeuristicsLincen Yang, Zhong Li, Matthijs van Leeuwen, Saber SalehkaleybarAAAI 2026
Builds on1
Related papers
- Deconvolutional Density Network: Modeling Free-Form Conditional DistributionsBing Chen, Mazharul Islam, Jisuo Gao, Lin WangAAAI 2022 · 8 citations
- Semi-supervised Conditional Density Estimation with Wasserstein Laplacian RegularisationOlivier Graffeuille, Yun Sing Koh, Jörg Wicker, Moritz K. LehmannAAAI 2022 · 4 citations
- Convex Polytope Trees and its Application to VAEMohammadreza Armandpour, Ali Sadeghian, Mingyuan ZhouNeurIPS 2021 · 2 citations
- Neural Prototype Trees for Interpretable Fine-Grained Image RecognitionMeike Nauta, Ron van Bree, Christin SeifertCVPR 2021
- Decision Trees with Short Explainable RulesVictor Feitosa Souza, Ferdinando Cicalese, Eduardo Sany Laber, Marco MolinaroNeurIPS 2022 · 26 citations
