USENIX Security2022Top-tier venue
One-off Disclosure Control by Heterogeneous Generalization
Olga Gkountouna, Katerina Doka, Mingqiang Xue, Jianneng Cao, Panagiotis Karras
Abstract
How can we orchestrate an one-off sharing of informative data about individuals, while bounding the risk of disclosing sensitive information to an adversary who has access to the global distribution of such information and to personal identifiers? Despite intensive efforts, current privacy protection techniques fall short of this objective. Differential privacy provides strong guarantees regarding the privacy risk incurred by one's participation in the data at the cost of high information loss and is vulnerable to learning-based attacks exploiting correlations among data. Syntactic anonymization bounds the risk on specific sensitive information incurred by data publication, yet typically resorts to a superfluous clustering of individuals into groups that forfeits data utility.
In this paper, we develop algorithms for disclosure control that abide to sensitive-information-oriented syntactic privacy guarantees and gain up to 77% in utility against current methods. We achieve this feat by recasting data heterogeneously, via bipartite matching, rather than homogeneously via clustering. We show that our methods resist adversaries who know the employed algorithm and its parameters. Our experimental study featuring synthetic and real data, as well as real learning and data analysis tasks, shows that these methods enhance data utility with a runtime overhead that is small and reducible by data partitioning, while the 𝛽-likeness guarantee with heterogeneous generalization staunchly resists machine-learning-based attacks, hence offers practical value.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on2
Related papers
- One-sided Differential PrivacyIos Kotsogiannis, Stelios Doudalis, Samuel Haney, Ashwin Machanavajjhala et al.ICDE 2020 · 36 citations
- A Linear Reconstruction Approach for Attribute Inference Attacks against Synthetic DataMeenatchi Sundaram Muthu Selva Annamalai, Andrea Gadotti, Luc RocherUSENIX Security 2024 · 37 citations
- Epistemic Parity: Reproducibility as an Evaluation Metric for Differential PrivacyLucas Rosenblatt, Bernease Herman, Anastasia Holovenko, Wonkwon Lee et al.VLDB 2023 · 11 citations
- Synthetic Data - Anonymisation Groundhog DayTheresa Stadler, Bristena Oprisanu, Carmela TroncosoUSENIX Security 2022
- PrivSynth: Alternating and Control-Based Optimization for Privacy and Utility in Synthetic DataXinyuan Zhao, Hanlin Gu, Guibao Song, Gongxi Zhu et al.CVPR 2026
