Audio Codec Augmentation for Robust Collaborative Watermarking of Speech Synthesis
Lauri Juvela, Xin Wang

TL;DR
This paper enhances collaborative audio watermarking techniques to improve robustness against various audio codecs, ensuring reliable detection of synthetic speech with minimal perceptual quality loss.
Contribution
It extends channel augmentation methods to non-differentiable codecs using waveform estimators and demonstrates transferability across different codec types and bitrates.
Findings
Robust watermarking can be achieved with black-box audio codecs.
Channel augmentation transfers well to traditional codecs.
Negligible perceptual degradation at high bitrates.
Abstract
Automatic detection of synthetic speech is becoming increasingly important as current synthesis methods are both near indistinguishable from human speech and widely accessible to the public. Audio watermarking and other active disclosure methods of are attracting research activity, as they can complement traditional deepfake defenses based on passive detection. In both active and passive detection, robustness is of major interest. Traditional audio watermarks are particularly susceptible to removal attacks by audio codec application. Most generated speech and audio content released into the wild passes through an audio codec purely as a distribution method. We recently proposed collaborative watermarking as method for making generated speech more easily detectable over a noisy but differentiable transmission channel. This paper extends the channel augmentation to work with…
Peer Reviews
No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.
Code & Models
Videos
No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.
Taxonomy
TopicsAdvanced Steganography and Watermarking Techniques · Speech and Audio Processing · Music and Audio Processing
MethodsDynamic Algorithm Configuration
