Title of article :
Comparison and combination of features in a hybrid HMM/MLP and a HMM/GMM speech recognition system
Author/Authors :
P.، Pujol, نويسنده , , S.، Pol, نويسنده , , C.، Nadeu, نويسنده , , A.، Hagen, نويسنده , , H.، Bourlard, نويسنده ,
Issue Information :
روزنامه با شماره پیاپی سال 2004
Pages :
-13
From page :
14
To page :
0
Abstract :
Recently, the advantages of the spectral parameters obtained by frequency filtering (FF) of the logarithmic filter-bank energies (logFBEs) have been reported. These parameters, which are frequency derivatives of the logFBEs, lie in the frequency domain, and have shown good recognition performance with respect to the conventional mel-frequency cepstral coefficients (MFCCs) for hidden Markov models (HMM) based systems. In this paper, the FF features are first compared with the MFCCs and the relative spectral perceptual linear prediction (Rasta-PLP) features using both a hybrid HMM/MLP and a usual HMM/Gaussian mixture models (HMM/GMM) based recognition system, for both clean and noisy speech. Taking advantage of the ability of the hybrid system to deal with correlated features, the inclusion of both the frequency secondderivatives and the raw logFBEs as additional features is proposed and tested. Moreover, the robustness of these features in noisy conditions is enhanced by combining the FF technique with the Rasta temporal filtering approach. Finally, a study of the FF features in the framework of multistream processing is presented. The best recognition results for both clean and noisy speech are obtained from the multistream combination of the J-Rasta-PLP features and the FF features.
Keywords :
waist circumference , Abdominal obesity , Food patterns , Prospective study
Journal title :
IEEE TRANSACTIONS ON SPEECH AND AUDIO PROCESSING
Serial Year :
2004
Journal title :
IEEE TRANSACTIONS ON SPEECH AND AUDIO PROCESSING
Record number :
86842
Link To Document :
بازگشت