Title :
Performance estimation of noisy speech recognition using spectral distortion and SNR of noise-reduced speech
Author :
Guo Ling ; Yamada, Tomoaki ; Makino, Shigeru ; Kitawaki, Nobuhiko
Author_Institution :
Grad. Sch. of Syst. & Inf. Eng., Univ. of Tsukuba, Tsukuba, Japan
Abstract :
To ensure a satisfactory QoE (Quality of Experience) and facilitate system design in speech recognition services, it is essential to establish a method that can be used to efficiently investigate recognition performance in different noise environments. Previously, we proposed a performance estimation method using the PESQ (Perceptual Evaluation of Speech Quality) as a spectral distortion measure. However, there is the problem that the relationship between the recognition performance and the distortion value differs depending on the noise reduction algorithm used. To solve this problem, we propose a novel performance estimation method that uses an estimator defined as a function of the distortion value and the SNR (Signal to Noise Ratio) of noise-reduced speech. The estimator is applicable to different noise reduction algorithms without any modification. We confirmed the effectiveness of the proposed method by experiments using the AURORA-2J connected digit recognition task and four different noise reduction algorithms.
Keywords :
distortion; quality of experience; signal denoising; spectral analysis; speech recognition; AURORA-2J; PESQ; QoE; SNR; digit recognition task; distortion value; noise environments; noise reduction algorithm; noise-reduced speech; noisy speech recognition; perceptual evaluation of speech quality; performance estimation method; quality of experience; recognition performance; signal to noise ratio; spectral distortion measure; speech recognition services; system design; Accuracy; Estimation; Noise reduction; Signal to noise ratio; Speech; Speech recognition; SNR; noise reduction; noisy speech recognition; performance estimation; spectral distortion;
Conference_Titel :
TENCON 2013 - 2013 IEEE Region 10 Conference (31194)
Conference_Location :
Xi´an
Print_ISBN :
978-1-4799-2825-5
DOI :
10.1109/TENCON.2013.6718993