Learning to Embed Distributions via Maximum Kernel Entropy
Oleksii Kachaiev, Stefano Recanatesi
Abstract
Empirical data can often be considered as samples from a set of probability distributions. Kernel methods have emerged as a natural approach for learning to classify these distributions. Although numerous kernels between distributions have been proposed, applying kernel methods to distribution regression tasks remains challenging, primarily because selecting a suitable kernel is not straightforward. Surprisingly, the question of learning a data-dependent distribution kernel has received little attention. In this paper, we propose a novel objective for the unsupervised learning of data-dependent distribution kernel, based on the principle of entropy maximization in the space of probability measure embeddings. We examine the theoretical properties of the latent embedding space induced by our objective, demonstrating that its geometric structure is well-suited for solving downstream discriminative tasks. Finally, we demonstrate the performance of the learned kernel across different modalities.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e4326d1f-0c1c-4c61-a7fc-db6887d77160Cited by top-tier papers3
- On Statistical Learning Theory for Distributional InputsChristian Fiedler, Pierre-François Massiani, Friedrich Solowjow, Sebastian TrimpeICML 2024 · 3 citations
- Statistical Learning Theory for Distributional ClassificationChristian FiedlerAAAI 2026
- Semi-Supervised Learning with Noisy Proxy Covariates: Generalization Bounds and Distribution RegressionKwangho Kim, Jisu KimICML 2026
Builds on10
- Understanding Contrastive Representation Learning through Alignment and Uniformity on the HypersphereTongzhou Wang, Phillip IsolaICML 2020 · 2,360 citations
- Kernel Methods Through the Roof: Handling Billions of Points EfficientlyGiacomo Meanti, Luigi Carratino, Lorenzo Rosasco, Alessandro RudiNeurIPS 2020 · 138 citations
- Learning Deep Features in Instrumental Variable RegressionLiyuan Xu, Yutian Chen, Siddarth Srinivasan, Nando de Freitas et al.ICLR 2021 · 85 citations
- Deep Proxy Causal Learning and its Application to Confounded Bandit Policy EvaluationLiyuan Xu, Heishiro Kanagawa, Arthur GrettonNeurIPS 2021 · 52 citations
- Sliced-Wasserstein on Symmetric Positive Definite Matrices for M/EEG SignalsClément Bonet, Benoît Malézieux, Alain Rakotomamonjy, Lucas Drumetz et al.ICML 2023 · 28 citations
Related papers
- Unsupervised Metric Learning with Synthetic ExamplesUjjal Kr Dutta, Mehrtash Harandi, C. Chandra SekharAAAI 2020 · 14 citations
- Distribution Regression with Sliced Wasserstein KernelsDimitri Meunier, Massimiliano Pontil, Carlo CilibertoICML 2022 · 24 citations
- Isolation Distributional Kernel: A New Tool for Kernel based Anomaly DetectionKai Ming Ting, Bi-Cun Xu, Takashi Washio, Zhi-Hua ZhouKDD 2020 · 49 citations
- Learning Probabilistic Ordinal Embeddings for Uncertainty-Aware RegressionWanhua Li, Xiaoke Huang, Jiwen Lu, Jianjiang Feng et al.CVPR 2021
- A Measure-Theoretic Approach to Kernel Conditional Mean EmbeddingsJunhyung Park, Krikamol MuandetNeurIPS 2020 · 123 citations
