What to Remember: Self-Adaptive Continual Learning for Audio Deepfake   Detection

Xiaohui Zhang; Jiangyan Yi; Chenglong Wang; Chuyuan Zhang; Siding; Zeng; Jianhua Tao

arXiv:2312.09651·cs.SD·December 18, 2023·1 cites

What to Remember: Self-Adaptive Continual Learning for Audio Deepfake Detection

Xiaohui Zhang, Jiangyan Yi, Chenglong Wang, Chuyuan Zhang, Siding, Zeng, Jianhua Tao

PDF

Open Access 1 Repo 1 Video

TL;DR

This paper introduces Radian Weight Modification, a continual learning method that improves audio deepfake detection by effectively distinguishing genuine and fake audio classes, reducing forgetting and enhancing knowledge retention.

Contribution

The paper proposes RWM, a novel continual learning approach that categorizes classes based on feature distribution and applies gradient modifications, advancing audio deepfake detection and broader machine learning tasks.

Findings

01

RWM outperforms existing continual learning methods in deepfake detection.

02

RWM effectively mitigates forgetting and enhances knowledge acquisition.

03

Applicable to other domains like image recognition.

Abstract

The rapid evolution of speech synthesis and voice conversion has raised substantial concerns due to the potential misuse of such technology, prompting a pressing need for effective audio deepfake detection mechanisms. Existing detection models have shown remarkable success in discriminating known deepfake audio, but struggle when encountering new attack types. To address this challenge, one of the emergent effective approaches is continual learning. In this paper, we propose a continual learning approach called Radian Weight Modification (RWM) for audio deepfake detection. The fundamental concept underlying RWM involves categorizing all classes into two groups: those with compact feature distributions across tasks, such as genuine audio, and those with more spread-out distributions, like various types of fake audio. These distinctions are quantified by means of the in-class cosine…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

cecile-hi/regularized-adaptive-weight-modification
pytorch

Videos

What to Remember: Self-Adaptive Continual Learning for Audio Deepfake Detection· underline

Taxonomy

TopicsSpeech and Audio Processing · Digital Media Forensic Detection · Music and Audio Processing