PrivQuant: Communication-Efficient Private Inference with Quantized   Network/Protocol Co-Optimization

Tianshi Xu; Shuzhang Zhong; Wenxuan Zeng; Runsheng Wang; Meng Li

arXiv:2410.09531·cs.CR·October 15, 2024

PrivQuant: Communication-Efficient Private Inference with Quantized Network/Protocol Co-Optimization

Tianshi Xu, Shuzhang Zhong, Wenxuan Zeng, Runsheng Wang, Meng Li

PDF

TL;DR

PrivQuant is a framework that combines network quantization and protocol optimization to significantly reduce communication and latency in private DNN inference using secure two-party computation.

Contribution

It introduces a joint optimization approach for quantized network inference and 2PC protocols, achieving communication efficiency and high accuracy.

Findings

01

Reduces communication by up to 11 times.

02

Achieves latency reductions of up to 8.7 times.

03

Outperforms prior frameworks like SiRNN, COINN, and CoPriv.

Abstract

Private deep neural network (DNN) inference based on secure two-party computation (2PC) enables secure privacy protection for both the server and the client. However, existing secure 2PC frameworks suffer from a high inference latency due to enormous communication. As the communication of both linear and non-linear DNN layers reduces with the bit widths of weight and activation, in this paper, we propose PrivQuant, a framework that jointly optimizes the 2PC-based quantized inference protocols and the network quantization algorithm, enabling communication-efficient private inference. PrivQuant proposes DNN architecture-aware optimizations for the 2PC protocols for communication-intensive quantized operators and conducts graph-level operator fusion for communication reduction. Moreover, PrivQuant also develops a communication-aware mixed precision quantization algorithm to improve…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.