Privacy-Preserving Adversarial Representation Learning in ASR: Reality   or Illusion?

Brij Mohan Lal Srivastava; Aur\'elien Bellet; Marc Tommasi; Emmanuel; Vincent

arXiv:1911.04913·cs.CL·November 13, 2019

Privacy-Preserving Adversarial Representation Learning in ASR: Reality or Illusion?

Brij Mohan Lal Srivastava, Aur\'elien Bellet, Marc Tommasi, Emmanuel, Vincent

PDF

Open Access

TL;DR

This paper investigates whether adversarial training can effectively anonymize speaker identity in speech representations for ASR, finding that while it reduces closed-set recognition accuracy, it does not significantly improve open-set speaker privacy.

Contribution

The study evaluates the effectiveness of adversarial training for speaker anonymization in ASR and highlights its limitations in real-world privacy protection.

Findings

01

Adversarial training reduces closed-set speaker classification accuracy.

02

Open-set speaker verification error does not significantly increase.

03

Standard representations still carry substantial speaker information.

Abstract

Automatic speech recognition (ASR) is a key technology in many services and applications. This typically requires user devices to send their speech data to the cloud for ASR decoding. As the speech signal carries a lot of information about the speaker, this raises serious privacy concerns. As a solution, an encoder may reside on each user device which performs local computations to anonymize the representation. In this paper, we focus on the protection of speaker identity and study the extent to which users can be recognized based on the encoded representation of their speech as obtained by a deep encoder-decoder architecture trained for ASR. Through speaker identification and verification experiments on the Librispeech corpus with open and closed sets of speakers, we show that the representations obtained from a standard architecture still carry a lot of information about speaker…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsSpeech Recognition and Synthesis · Adversarial Robustness in Machine Learning · Privacy-Preserving Technologies in Data