Lune

ICML2026Top-tier venue

Approximation Preserving Coresets

Milind Prabhu, Chris Schwiegelshohn, Sudarshan Shyam

2026Year

Abstract

Clustering in a big data setting is an intensively studied problem, with coresets emerging as one of the important paradigms in this line of work. Given a cost function cost(P,S)\text{cost}(P,S) mapping input points PP and a solution SS to an objective value, a coreset is a typically weighted sketch Ω⊆P\Omega\subseteq P such that cost(Ω,S)≈cost(P,S)\text{cost}(\Omega,S)\approx \text{cost}(P,S). In practice, coreset sizes much smaller than those suggested by theoretical guarantees are often found to be sufficient. In this paper, we offer an explanation for this phenomenon. Smaller coreset sizes suffice if we only wish to preserve the costs of good solutions, i.e., solutions with low cost. We define and devise approximation-preserving coresets, which provide a weaker guarantee than strong coresets, which apply to all solutions, while providing stronger guarantees than weak coresets, which apply only to the optimum solution. We complement this result by showing that even a very small distortion in the approximation factor cannot admit coresets of this size.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext cfc6d770-113d-4690-8808-2d8fe1cfce9f

Builds on21

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines