Privacy-preserving Automatic Speaker Diarization

Francisco Teixeira; Alberto Abad; Bhiksha Raj; Isabel Trancoso

arXiv:2210.14995·eess.AS·April 19, 2023·1 cites

Privacy-preserving Automatic Speaker Diarization

Francisco Teixeira, Alberto Abad, Bhiksha Raj, Isabel Trancoso

PDF

Open Access

TL;DR

This paper introduces a privacy-preserving automatic speaker diarization system that leverages cryptographic techniques to protect user privacy during voice data processing, achieving real-time performance.

Contribution

It is the first to combine Secure Multiparty Computation and Secure Modular Hashing for privacy-preserving ASD, addressing a previously overlooked area.

Findings

01

Achieves real-time processing with factors of 1.1 and 1.6.

02

Balances privacy and performance effectively.

03

Introduces a novel cryptographic approach to ASD.

Abstract

Automatic Speaker Diarization (ASD) is an enabling technology with numerous applications, which deals with recordings of multiple speakers, raising special concerns in terms of privacy. In fact, in remote settings, where recordings are shared with a server, clients relinquish not only the privacy of their conversation, but also of all the information that can be inferred from their voices. However, to the best of our knowledge, the development of privacy-preserving ASD systems has been overlooked thus far. In this work, we tackle this problem using a combination of two cryptographic techniques, Secure Multiparty Computation (SMC) and Secure Modular Hashing, and apply them to the two main steps of a cascaded ASD system: speaker embedding extraction and agglomerative hierarchical clustering. Our system is able to achieve a reasonable trade-off between performance and efficiency,…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsSpeech Recognition and Synthesis · Music and Audio Processing · Speech and Audio Processing