Title :
Asymmetric acoustic modeling of mixed language speech
Author :
Li, Ying ; Fung, Pascale ; Xu, Ping ; Liu, Yi
Author_Institution :
Dept. of Electron. & Comput. Eng., Hong Kong Univ. of Sci. & Technol., Hong Kong, China
Abstract :
We propose to improve speech recognition performance on speaker-independent, mixed language speech by asymmetric acoustic modeling. Mixed language is either inter-sentential code switching from the source matrix language to a foreign language or intra-sentential code mixing between the matrix language and embedded foreign words or phrases. In either case, the foreign phrases are pronounced by the matrix language speaker with varying degrees of accent. Our pro posed system using selective decision tree merging between a bilingual model and an accented embedded speech model outperforms previous approaches of either using a bilingual model with model retraining by 21.51%, or using adaptation by 15.88%. It outperforms all models on both code mixing and code switching cases. We successfully improved recognition on embedded foreign speech without degrading the performance on the matrix language speech.
Keywords :
decision trees; speech coding; speech recognition; accented embedded speech model; asymmetric acoustic modeling; code mixing; code switching; foreign language; intersentential code switching; intrasentential code mixing; language speaker; matrix language speech; mixed language speech; selective decision tree; source matrix language; speaker-independent speech; speech recognition; speech recognition performance; Acoustics; Adaptation models; Data models; Hidden Markov models; Speech; Speech recognition; Switches; code mix; code switch; mixed language acoustic modeling;
Conference_Titel :
Acoustics, Speech and Signal Processing (ICASSP), 2011 IEEE International Conference on
Conference_Location :
Prague
Print_ISBN :
978-1-4577-0538-0
Electronic_ISBN :
1520-6149
DOI :
10.1109/ICASSP.2011.5947480