CorrectSpeech: A Fully Automated System for Speech Correction and Accent Reduction
Daxin Tan, Liqun Deng, Nianzu Zheng, Yu Ting Yeung, Xin Jiang, Xiao, Chen, Tan Lee

TL;DR
CorrectSpeech is an automated system designed to identify and correct speech errors and reduce accents by combining speech recognition, alignment, and editing, demonstrated on multiple speech corpora with promising results.
Contribution
This paper introduces CorrectSpeech, a novel fully automated pipeline for speech correction and accent reduction that integrates recognition, alignment, and editing modules.
Findings
Effective correction of mispronunciations demonstrated
Accent reduction achieved on multiple corpora
System performance depends on recognition and alignment quality
Abstract
This study propose a fully automated system for speech correction and accent reduction. Consider the application scenario that a recorded speech audio contains certain errors, e.g., inappropriate words, mispronunciations, that need to be corrected. The proposed system, named CorrectSpeech, performs the correction in three steps: recognizing the recorded speech and converting it into time-stamped symbol sequence, aligning recognized symbol sequence with target text to determine locations and types of required edit operations, and generating the corrected speech. Experiments show that the quality and naturalness of corrected speech depend on the performance of speech recognition and alignment modules, as well as the granularity level of editing operations. The proposed system is evaluated on two corpora: a manually perturbed version of VCTK and L2-ARCTIC. The results demonstrate that our…
Peer Reviews
No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.
Videos
No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.
Taxonomy
TopicsSpeech and Audio Processing · Speech Recognition and Synthesis · Phonetics and Phonology Research
