FreqyWM: Frequency Watermarking for the New Data Economy
Devris Isler, Elisa Cabana, Álvaro García-Recuero, Georgia Koutrika, Nikolaos Laoutaris
摘要
We present a novel technique for modulating the appearance frequency of a few tokens within a dataset for encoding an invisible watermark that can be used to protect ownership rights upon data. We develop optimal as well as fast heuristic algorithms for creating and verifying such watermarks. We also demonstrate the robustness of our technique against various attacks and derive analytical bounds for the false positive probability of erroneously “detecting” a watermark on a dataset that does not carry it. Our technique is applicable to both single dimensional and multidimensional datasets, is independent of token type, allows for a fine control of the introduced distortion, and can be used in a variety of use cases that involve buying and selling data in contemporary data marketplaces.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper7
- RIGA: Covert and Robust White-Box Watermarking of Deep Neural NetworksTianhao Wang, Florian KerschbaumWWW 2021 · 被引用 128 次
- Frequency Estimation under Local Differential PrivacyGraham Cormode, Samuel Maddock, Carsten MapleVLDB 2021 · 被引用 70 次
- VeriDB: An SGX-based Verifiable DatabaseWenchao Zhou, Yifan Cai, Yanqing Peng, Sheng Wang 等SIGMOD 2021 · 被引用 52 次
- HEDA: Multi-Attribute Unbounded Aggregation over Homomorphically Encrypted DatabaseXuanle Ren, Le Su, Zhen Gu, Sheng Wang 等VLDB 2023 · 被引用 42 次
- PRISM: Private Verifiable Set Computation over Multi-Owner Outsourced DatabasesYin Li, Dhrubajyoti Ghosh, Peeyush Gupta, Sharad Mehrotra 等SIGMOD 2021 · 被引用 26 次
相关 Paper
- STMark: Watermarking Scalar and Textual Data Robustly Using Frequency StatisticsJiongyang Ji, Yanguo Peng, Zhen Lv, Gengquan Guo 等CCS 2026
- B2Mark: A Blind and Buyer-Traceable Watermarking Scheme for Tabular DatasetsYihao Zheng, Jinfei Liu, Kui Ren, Li XiongSIGMOD 2026 · 被引用 1 次
- PuzzleMark: Implicit Jigsaw Learning for Robust Code Dataset Watermarking in Neural Code Completion ModelsHaocheng Huang, Yuchen Chen, Weisong Sun, Peizhuo Lv 等FSE 2026
- Provable Watermarking for Data Poisoning AttacksYifan Zhu, Lijia Yu, Xiao-Shan GaoNeurIPS 2025 · 被引用 3 次
- TabularMark: Watermarking Tabular Datasets for Machine LearningYihao Zheng, Haocheng Xia, Junyuan Pang, Jinfei Liu 等CCS 2024 · 被引用 5 次
