Geometry of Sensitivity: Twice Sampling and Hybrid Clipping in   Differential Privacy with Optimal Gaussian Noise and Application to Deep   Learning

Hanshen Xiao; Jun Wan; Srinivas Devadas

arXiv:2309.02672·cs.CR·September 29, 2023·1 cites

Geometry of Sensitivity: Twice Sampling and Hybrid Clipping in Differential Privacy with Optimal Gaussian Noise and Application to Deep Learning

Hanshen Xiao, Jun Wan, Srinivas Devadas

PDF

Open Access 1 Repo

TL;DR

This paper investigates the geometry of high-dimensional sensitivity sets in differential privacy, introducing twice sampling to improve privacy-utility tradeoffs and providing optimal Gaussian noise bounds under various conditions.

Contribution

It characterizes the optimal Gaussian noise for high-dimensional sensitivity sets, introduces twice sampling for better privacy amplification, and analyzes the geometry of sensitivity sets in differential privacy.

Findings

01

Curse of dimensionality is tight for symmetric sensitivity sets.

02

Asymmetric sensitivity sets can have dimension-independent optimal noise bounds.

03

Twice sampling enhances privacy amplification especially at small sampling rates.

Abstract

We study the fundamental problem of the construction of optimal randomization in Differential Privacy. Depending on the clipping strategy or additional properties of the processing function, the corresponding sensitivity set theoretically determines the necessary randomization to produce the required security parameters. Towards the optimal utility-privacy tradeoff, finding the minimal perturbation for properly-selected sensitivity sets stands as a central problem in DP research. In practice, l_2/l_1-norm clippings with Gaussian/Laplace noise mechanisms are among the most common setups. However, they also suffer from the curse of dimensionality. For more generic clipping strategies, the understanding of the optimal noise for a high-dimensional sensitivity set remains limited. In this paper, we revisit the geometry of high-dimensional sensitivity sets and present a series of results to…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

hanshen-xiao/twice_sampling_and_hybrid_clipping
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsPrivacy-Preserving Technologies in Data · Markov Chains and Monte Carlo Methods · Stochastic Gradient Optimization Techniques