Optimal Rates for Regularized Conditional Mean Embedding Learning
Zhu Li, Dimitri Meunier, Mattes Mollenhauer, Arthur Gretton
摘要
We address the consistency of a kernel ridge regression estimate of the conditional mean embedding (CME), which is an embedding of the conditional distribution of given into a target reproducing kernel Hilbert space . The CME allows us to take conditional expectations of target RKHS functions, and has been employed in nonparametric causal and Bayesian inference. We address the misspecified setting, where the target CME is in the space of Hilbert-Schmidt operators acting from an input interpolation space between and , to . This space of operators is shown to be isomorphic to a newly defined vector-valued interpolation space. Using this isomorphism, we derive a novel and adaptive statistical learning rate for the empirical CME estimator under the misspecified setting. Our analysis reveals that our rates match the optimal rates without assuming to be finite dimensional. We further establish a lower bound on the learning rate, which shows that the obtained upper bound is optimal.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper26
- IRPruneDet: Efficient Infrared Small Target Detection via Wavelet Structure-Regularized Soft Channel PruningMingjin Zhang, Handi Yang, Jie Guo, Yunsong Li 等AAAI 2024 · 被引用 159 次
- Sharp Spectral Rates for Koopman Operator LearningVladimir Kostic, Karim Lounici, Pietro Novelli, Massimiliano PontilNeurIPS 2023 · 被引用 57 次
- Estimating Koopman operators with sketching to provably learn large scale dynamical systemsGiacomo Meanti, Antoine Chatalic, Vladimir Kostic, Pietro Novelli 等NeurIPS 2023 · 被引用 22 次
- On the Optimality of Misspecified Kernel Ridge RegressionHaobo Zhang, Yicheng Li, Weihao Lu, Qian LinICML 2023 · 被引用 19 次
- Learning invariant representations of time-homogeneous stochastic dynamical systemsVladimir R. Kostic, Pietro Novelli, Riccardo Grazzi, Karim Lounici 等ICLR 2024 · 被引用 17 次
它引用的顶会 Paper1
相关 Paper
- Deconditional Downscaling with Gaussian ProcessesSiu Lun Chau, Shahine Bouabid, Dino SejdinovicNeurIPS 2021 · 被引用 31 次
- Towards a Unified Analysis of Kernel-based Methods Under Covariate ShiftXingdong Feng, Xin He, Caixing Wang, Chao Wang 等NeurIPS 2023 · 被引用 17 次
- On Statistical Learning Theory for Distributional InputsChristian Fiedler, Pierre-François Massiani, Friedrich Solowjow, Sebastian TrimpeICML 2024 · 被引用 3 次
- On the Consistency of Kernel Methods with Dependent ObservationsPierre-François Massiani, Sebastian Trimpe, Friedrich SolowjowICML 2024 · 被引用 2 次
- Statistical Learning Theory for Distributional ClassificationChristian FiedlerAAAI 2026
