Watermarking Generative Tabular Data

Hengzhi He; Peiyu Yu; Junpeng Ren; Ying Nian Wu; Guang; Cheng

arXiv:2405.14018·cs.CR·May 28, 2024·1 cites

Watermarking Generative Tabular Data

Hengzhi He, Peiyu Yu, Junpeng Ren, Ying Nian Wu, Guang, Cheng

PDF

Open Access

TL;DR

This paper presents a simple, statistically sound method for watermarking tabular data that ensures data integrity, robustness against noise, and easy detection through a hypothesis-testing framework.

Contribution

It introduces a novel watermarking technique based on data binning with theoretical guarantees and practical robustness for tabular datasets.

Findings

01

Effective watermark detection with statistical guarantees

02

High robustness against additive noise attacks

03

Preserves data fidelity in watermark embedding

Abstract

In this paper, we introduce a simple yet effective tabular data watermarking mechanism with statistical guarantees. We show theoretically that the proposed watermark can be effectively detected, while faithfully preserving the data fidelity, and also demonstrates appealing robustness against additive noise attack. The general idea is to achieve the watermarking through a strategic embedding based on simple data binning. Specifically, it divides the feature's value range into finely segmented intervals and embeds watermarks into selected ``green list" intervals. To detect the watermarks, we develop a principled statistical hypothesis-testing framework with minimal assumptions: it remains valid as long as the underlying data distribution has a continuous density function. The watermarking efficacy is demonstrated through rigorous theoretical analysis and empirical validation, highlighting…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsAdvanced Steganography and Watermarking Techniques · Cellular Automata and Applications · Chaos-based Image/Signal Encryption