Distribution Compression in Near-Linear Time
Abhishek Shetty, Raaz Dwivedi, Lester Mackey
摘要
In distribution compression, one aims to accurately summarize a probability distribution using a small number of representative points. Near-optimal thinning procedures achieve this goal by sampling points from a Markov chain and identifying points with discrepancy to . Unfortunately, these algorithms suffer from quadratic or super-quadratic runtime in the sample size . To address this deficiency, we introduce Compress++, a simple meta-procedure for speeding up any thinning algorithm while suffering at most a factor of in error. When combined with the quadratic-time kernel halving and kernel thinning algorithms of Dwivedi and Mackey (2021), Compress++ delivers points with integration error and better-than-Monte-Carlo maximum mean discrepancy in time and space. Moreover, Compress++ enjoys the same near-linear runtime given any quadratic-time input and reduces the runtime of super-quadratic algorithms by a square-root factor. In our benchmarks with high-dimensional Monte Carlo samples and Markov chains targeting challenging differential equation posteriors, Compress++ matches or nearly matches the accuracy of its input algorithm in orders of magnitude less time.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- Generalized Kernel ThinningRaaz Dwivedi, Lester MackeyICLR 2022 · 被引用 37 次
- Positively Weighted Kernel Quadrature via SubsamplingSatoshi Hayakawa, Harald Oberhauser, Terry J. LyonsNeurIPS 2022 · 被引用 35 次
- Estimating the Rate-Distortion Function by Wasserstein Gradient DescentYibo Yang, Stephan Eckstein, Marcel Nutz, Stephan MandtNeurIPS 2023 · 被引用 21 次
- Sampling-based Nyström Approximation and Kernel QuadratureSatoshi Hayakawa, Harald Oberhauser, Terry J. LyonsICML 2023 · 被引用 20 次
- Kernel Quadrature with Randomly Pivoted CholeskyEthan Epperly, Elvira MorenoNeurIPS 2023 · 被引用 16 次
它引用的顶会 Paper1
相关 Paper
- Debiased Distribution CompressionLingxiao Li, Raaz Dwivedi, Lester MackeyICML 2024 · 被引用 7 次
- Low-Rank ThinningAnnabelle Michael Carrell, Albert Gong, Abhishek Shetty, Raaz Dwivedi 等ICML 2025
- Supervised Kernel ThinningAlbert Gong, Kyuseong Choi, Raaz DwivediNeurIPS 2024 · 被引用 6 次
- Efficient and Accurate Explanation Estimation with Distribution CompressionHubert Baniecki, Giuseppe Casalicchio, Bernd Bischl, Przemyslaw BiecekICLR 2025
- Conditional Distribution Compression via the Kernel Conditional Mean EmbeddingDominic Broadbent, Nick Whiteley, Robert Allison, Tom LovettNeurIPS 2025 · 被引用 1 次
