Title :
Choice of Mel filter bank in computing MFCC of a resampled speech
Author :
Kopparapu, Sunil Kumar ; Laxminarayana, M.
Author_Institution :
TCS Innovation Labs.-Mumbai, Tata Consultancy Services, Thane West, India
Abstract :
Mel Frequency Cepstral Coefficients (MFCCs) are the most popularly used speech features in many speech and speaker recognition applications. In this paper, we study the effect of resampling a speech signal on these speech features. We first derive a relationship between the MFCC parameters of the resampled speech and the MFCC parameters of the original speech. We propose six methods of calculating the MFCC parameters of downsampled speech by transforming the Mel filter bank used to compute MFCC of the original speech. We then experimentally compute the MFCC parameters of the down sampled speech using the proposed methods and compute the Pearson coefficient between the MFCC parameters of the downsampled speech and that of the original speech to identify the most effective choice of Mel-filter band that enables the computed MFCC of the resampled speech to be as close as possible to the MFCC of the original speech.
Keywords :
filtering theory; speech recognition; MFCC; Mel filter bank; Mel frequency cepstral coefficients; resampled speech; speaker recognition applications; speech recognition applications; Discrete cosine transforms; Filter bank; HTML; Mel frequency cepstral coefficient; Speech; MFCC; Time scale modification; time compression; time expansion;
Conference_Titel :
Information Sciences Signal Processing and their Applications (ISSPA), 2010 10th International Conference on
Conference_Location :
Kuala Lumpur
Print_ISBN :
978-1-4244-7165-2
DOI :
10.1109/ISSPA.2010.5605491