Title :
Emotion Conversion Based on Prosodic Unit Selection
Author :
Erro, Daniel ; Navas, Eva ; Hernáez, Inma ; Saratxaga, Ibon
Author_Institution :
Electron. & Telecommun. Dept., Univ. of the Basque Country (UPVEHU), Bilbao, Spain
fDate :
7/1/2010 12:00:00 AM
Abstract :
Voice conversion has been traditionally focused on spectrum. Current systems lack a solid prosody conversion method suitable for different speaking styles. Recently, the unit selection technique has been applied to transform emotional intonation contours. This paper goes one step beyond: it explores strategies for training and configuring the selection cost function in an emotion conversion application. The proposed system, which uses accent groups as basic intonation units and performs conversion also on phoneme durations and intensity, is evaluated by means of a carefully designed subjective test involving the big six emotions. Although the expressiveness of the converted sentences is still far from that of natural emotional speech, satisfactory results are obtained when different configurations are used for different emotions.
Keywords :
natural language processing; speech synthesis; accent groups; emotion conversion; emotional intonation contour transform; natural emotional speech; phoneme durations; phoneme intensity; prosodic unit selection; selection cost function; speech synthesis; voice conversion; Emotional speech synthesis; intonation; prosody; unit selection; voice conversion;
Journal_Title :
Audio, Speech, and Language Processing, IEEE Transactions on
DOI :
10.1109/TASL.2009.2038658