On Statistical Learning Theory for Distributional Inputs
Christian Fiedler, Pierre-François Massiani, Friedrich Solowjow, Sebastian Trimpe
Abstract
In supervised learning with distributional inputs in the twostage sampling setup, relevant to applications like learningbased medical screening or causal learning, the inputs (which are probability distributions) are not accessible in the learning phase, but only samples thereof. This problem is particularly amenable to kernel-based learning methods, where the distributions or samples are first embedded into a Hilbert space, often using kernel mean embeddings (KMEs), and then a standard kernel method like Support Vector Machines (SVMs) is applied, using a kernel defined on the embedding Hilbert space. In this work, we contribute to the theoretical analysis of this latter approach, with a particular focus on classification with distributional inputs using SVMs. We establish a new oracle inequality and derive consistency and learning rate results. Furthermore, for SVMs using the hinge loss and Gaussian kernels, we formulate a novel variant of an established noise assumption from the binary classification literature, under which we can establish learning rates. Finally, some of our technical tools like a new feature space for Gaussian kernels on Hilbert spaces are of independent interest.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4b08f96f-abaf-49b1-bbb1-0d327e88916cCited by top-tier papers2
- Kernel conditional tests from learning-theoretic boundsPierre-François Massiani, Christian Fiedler, Lukas Haverbeck, Friedrich Solowjow et al.NeurIPS 2025 · 1 citation
- Statistical Learning Theory for Distributional ClassificationChristian FiedlerAAAI 2026
Builds on3
- Distribution Regression with Sliced Wasserstein KernelsDimitri Meunier, Massimiliano Pontil, Carlo CilibertoICML 2022 · 24 citations
- Kernelized Cumulants: Beyond Kernel Mean EmbeddingsPatric Bonnier, Harald Oberhauser, Zoltán SzabóNeurIPS 2023 · 9 citations
- Learning to Embed Distributions via Maximum Kernel EntropyOleksii Kachaiev, Stefano RecanatesiNeurIPS 2024 · 3 citations
Related papers
- On the Consistency of Kernel Methods with Dependent ObservationsPierre-François Massiani, Sebastian Trimpe, Friedrich SolowjowICML 2024 · 2 citations
- Optimal Rates for Regularized Conditional Mean Embedding LearningZhu Li, Dimitri Meunier, Mattes Mollenhauer, Arthur GrettonNeurIPS 2022 · 69 citations
- A Measure-Theoretic Approach to Kernel Conditional Mean EmbeddingsJunhyung Park, Krikamol MuandetNeurIPS 2020 · 123 citations
- Towards a Unified Analysis of Kernel-based Methods Under Covariate ShiftXingdong Feng, Xin He, Caixing Wang, Chao Wang et al.NeurIPS 2023 · 17 citations
- On kernel-based statistical learning theory in the mean field limitChristian Fiedler, Michael Herty, Sebastian TrimpeNeurIPS 2023
