DocumentCode :
3703341
Title :
Learning speech emotion features by joint disentangling-discrimination
Author :
Wentao Xue;Zhengwei Huang;Xin Luo;Qirong Mao
Author_Institution :
Department of Computer Science and Communication Engineering, Jiangsu University, Zhenjiang, China
fYear :
2015
Firstpage :
374
Lastpage :
379
Abstract :
Speech plays an important part in human-computer interaction. As a major branch of speech processing, speech emotion recognition (SER) has drawn much attention of researchers. Excellent discriminant features are of great importance in SER. However, emotion-specific features are commonly mixed with some other features. In this paper, we introduce an approach to pull apart these two parts of features as much as possible. First we employ an unsupervised feature learning framework to achieve some rough features. Then these rough features are further fed into a semi-supervised feature learning framework. In this phase, efforts are made to disentangle the emotion-specific features and some other features by using a novel loss function, which combines reconstruction penalty, orthogonal penalty, discriminative penalty and verification penalty. Orthogonal penalty is utilized to disentangle emotion-specific features and other features. The discriminative penalty enlarges inter-emotion variations, while the verification penalty reduces the intra-emotion variations. Evaluations on the FAU Aibo emotion database show that our approach can improve the speech emotion classification performance.
Keywords :
"Feature extraction","Speech","Kernel","Support vector machines","Speech recognition","Computer aided engineering","Yttrium"
Publisher :
ieee
Conference_Titel :
Affective Computing and Intelligent Interaction (ACII), 2015 International Conference on
Electronic_ISBN :
2156-8111
Type :
conf
DOI :
10.1109/ACII.2015.7344598
Filename :
7344598
Link To Document :
بازگشت