Transfer Learning-Based Deep Residual Learning for Speech Recognition in   Clean and Noisy Environments

Noussaiba Djeffal; Djamel Addou; Hamza Kheddar; Sid Ahmed Selouani

arXiv:2505.01632·eess.AS·May 6, 2025

Transfer Learning-Based Deep Residual Learning for Speech Recognition in Clean and Noisy Environments

Noussaiba Djeffal, Djamel Addou, Hamza Kheddar, Sid Ahmed Selouani

PDF

TL;DR

This paper presents a transfer learning approach using residual neural networks to improve speech recognition accuracy in both clean and noisy environments, outperforming CNN and LSTM models.

Contribution

The study introduces a novel ResNet-based transfer learning framework for robust speech recognition in diverse acoustic conditions.

Findings

01

Achieved 98.94% accuracy in clean environments

02

Achieved 91.21% accuracy in noisy environments

03

Outperformed CNN and LSTM models in recognition accuracy

Abstract

Addressing the detrimental impact of non-stationary environmental noise on automatic speech recognition (ASR) has been a persistent and significant research focus. Despite advancements, this challenge continues to be a major concern. Recently, data-driven supervised approaches, such as deep neural networks, have emerged as promising alternatives to traditional unsupervised methods. With extensive training, these approaches have the potential to overcome the challenges posed by diverse real-life acoustic environments. In this light, this paper introduces a novel neural framework that incorporates a robust frontend into ASR systems in both clean and noisy environments. Utilizing the Aurora-2 speech database, the authors evaluate the effectiveness of an acoustic feature set for Mel-frequency, employing the approach of transfer learning based on Residual neural network (ResNet). The…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

MethodsSparse Evolutionary Training