Peeling Close to the Orientability Threshold - Spatial Coupling in Hashing-Based Data Structures
Stefan Walzer
Abstract
In multiple-choice data structures each element x in a set S of m keys is associated with a random set e(x) ⊆ [n] of buckets with capacity ℓ ≥ 1 by hash functions. This setting is captured by the hypergraph H = ([n], e(x) | x ∈ S). Accomodating each key in an associated bucket amounts to finding an ℓ-orientation of H assigning to each hyperedge an incident vertex such that each vertex is assigned at most ℓ hyperedges. If each subhypergraph of H has minimum degree at most ℓ, then an ℓ-orientation can be found greedily and H is called ℓ-peelable. Peelability has a central role in invertible Bloom lookup tables and can speed up the construction of retrieval data structures, perfect hash functions and cuckoo hash tables.
Many hypergraphs exhibit sharp density thresholds with respect to ℓ-orientability and ℓ-peelability, i.e. as the density c = m n grows past a critical value, the probability of these properties drops from almost 1 to almost 0. In fully random k-uniform hypergraphs the thresholds c * k,ℓ for ℓ-orientability significantly exceed the thresholds for ℓ-peelability. In this paper, for every k ≥ 2 and ℓ ≥ 1 with (k, ℓ) = (2, 1) and every z > 0, we construct a new family of random k-uniform hypergraphs with i.i.d. random hyperedges such that both the ℓ-peelability and the ℓ-orientability thresholds approach c * k,ℓ as z → ∞. In particular we achieve 1-peelability at densities arbitrarily close to 1, extending the reach of greedy algorithms.
Our construction is simple: The n vertices are linearly ordered and each hyperedge selects its k elements uniformly at random from a random range of n z+1 consecutive vertices. We thus exploit the phenomenon of threshold saturation via spatial coupling discovered in the context of low-density parity-check codes. Once the connection to data structures is in plain sight, a framework by Kudekar, Richardson and Urbanke [39] does the heavy lifting in our proof.
We demonstrate the usefulness of our construction using our hypergraphs as a drop-in replacement in a retrieval data structure by Botelho et al. [8]. This reduces memory usage from ≈ 1.23m bits to ≈ 1.12m bits (for input size m). Using k > 3 attains, at small sacrifices in running time, further improvements to memory usage.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4ba11cf2-fd77-42d0-a17b-7b3861d9aa8bCited by top-tier papers2
- ChainedFilter: Combining Membership Filters by Chain RuleHaoyu Li, Liuhui Wang, Qizhi Chen, Jianan Ji et al.SIGMOD 2024 · 5 citations
- CodingSketch: A Hierarchical Sketch with Efficient Encoding and Recursive DecodingQizhi Chen, Yisen Hong, Yuhan Wu, Tong Yang et al.ICDE 2024 · 5 citations
Related papers
- A Hash Table Without Hash Functions, and How to Get the Most Out of Your Random BitsWilliam KuszmaulFOCS 2022 · 7 citations
- ℓ2/ℓ2 Sparse Recovery via Weighted Hypergraph PeelingNick Fischer, Vasileios NakosFOCS 2025 · 1 citation
- Perfect Matchings in Random Sparsifications of Dense HypergraphsJie Han, Jingwen ZhaoSODA 2026
- Optimal and Efficient Partite Decompositions of HypergraphsAndrew Krapivin, Benjamin Przybocki, Nicolás Sanhueza-Matamala, Bernardo SubercaseauxSTOC 2026 · 2 citations
- Densest Subgraph: Supermodularity, Iterative Peeling, and FlowChandra Chekuri, Kent Quanrud, Manuel R. TorresSODA 2022 · 34 citations
