Supervised Classifiers for Audio Impairments with Noisy Labels

Chandan K A Reddy; Ross Cutler; Johannes Gehrke

arXiv:1907.01742·cs.SD·July 4, 2019

Supervised Classifiers for Audio Impairments with Noisy Labels

Chandan K A Reddy, Ross Cutler, Johannes Gehrke

PDF

TL;DR

This paper investigates how supervised audio impairment classifiers trained on noisy user feedback labels perform, demonstrating CNNs' robustness to label noise and the need for larger datasets with increased noise levels.

Contribution

It provides an analysis of CNN performance on noisy labels in audio impairment classification and highlights the importance of larger datasets for noisy label training.

Findings

01

CNNs outperform dense networks on noisy labels

02

Training with noisy labels requires larger datasets

03

CNNs generalize better with noisy label data

Abstract

Voice-over-Internet-Protocol (VoIP) calls are prone to various speech impairments due to environmental and network conditions resulting in bad user experience. A reliable audio impairment classifier helps to identify the cause for bad audio quality. The user feedback after the call can act as the ground truth labels for training a supervised classifier on a large audio dataset. However, the labels are noisy as most of the users lack the expertise to precisely articulate the impairment in the perceived speech. In this paper, we analyze the effects of massive noise in labels in training dense networks and Convolutional Neural Networks (CNN) using engineered features, spectrograms and raw audio samples as inputs. We demonstrate that CNN can generalize better on the training data with a large number of noisy labels and gives remarkably higher test performance. The classifiers were trained…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.