The INTERSPEECH 2020 Deep Noise Suppression Challenge: Datasets,   Subjective Speech Quality and Testing Framework

Chandan K. A. Reddy; Ebrahim Beyrami; Harishchandra Dubey; Vishak; Gopal; Roger Cheng; Ross Cutler; Sergiy Matusevych; Robert Aichner; Ashkan; Aazami; Sebastian Braun; Puneet Rana; Sriram Srinivasan; Johannes Gehrke

arXiv:2001.08662·cs.SD·April 21, 2020·67 cites

The INTERSPEECH 2020 Deep Noise Suppression Challenge: Datasets, Subjective Speech Quality and Testing Framework

Chandan K. A. Reddy, Ebrahim Beyrami, Harishchandra Dubey, Vishak, Gopal, Roger Cheng, Ross Cutler, Sergiy Matusevych, Robert Aichner, Ashkan, Aazami, Sebastian Braun, Puneet Rana, Sriram Srinivasan, Johannes Gehrke

PDF

Open Access 1 Repo

TL;DR

The INTERSPEECH 2020 Deep Noise Suppression Challenge promotes research in real-time speech enhancement by providing datasets, a subjective testing framework, and a focus on improving perceptual speech quality in real-world scenarios.

Contribution

It introduces open-source datasets and an online subjective testing framework to better evaluate noise suppression methods on real recordings.

Findings

01

Open-source large speech and noise datasets provided.

02

Subjective evaluation framework based on ITU-T P.808 released.

03

Challenge results emphasize real-world performance over synthetic metrics.

Abstract

The INTERSPEECH 2020 Deep Noise Suppression Challenge is intended to promote collaborative research in real-time single-channel Speech Enhancement aimed to maximize the subjective (perceptual) quality of the enhanced speech. A typical approach to evaluate the noise suppression methods is to use objective metrics on the test set obtained by splitting the original dataset. Many publications report reasonable performance on the synthetic test set drawn from the same distribution as that of the training set. However, often the model performance degrades significantly on real recordings. Also, most of the conventional objective metrics do not correlate well with subjective tests and lab subjective tests are not scalable for a large test set. In this challenge, we open-source a large clean speech and noise corpus for training the noise suppression models and a representative test set to…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

microsoft/DNS-Challenge
noneOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsSpeech and Audio Processing · Acoustic Wave Phenomena Research · Hearing Loss and Rehabilitation

MethodsTest