• Title of article

    Speech Emotion Recognition based on Improved SOAR Model

  • Author/Authors

    Ramzani Shahrestani ، Matin Department of Computer Engineering - Islamic Azad University, Rasht Branch , Motamed ، Sara Department of Computer Engineering - Islamic Azad University, Fouman and Shaft Branch , Yamaghani ، Mohammadreza Department of Computer Engineering - Islamic Azad University, Lahijan Branch

  • From page
    39
  • To page
    48
  • Abstract
    In recent years, emotion recognition as a new method for human-computer interaction has attracted the attention of researchers. Automatic speech emotion recognition has become one of the practical methods to increase engagement in most industries. It is expected that emotion recognition based on audio information can result in better accuracy. The purpose of this article is to present an efficient method for recognizing emotional states from speech signals, based on a new cognitive model. Due to the importance of the topic, this article presents an efficient method for recognizing emotional states from speech signals based on a mixed deep learning and cognitive model called SOAR. To implement each part of this model, two main steps have been introduced. The first step is reading the video and converting it to images and preprocessing it. The next step is to use the combination of convolutional neural network (CNN) and learning automata (LA) to classify and detect the rate of facial emotional recognition. The reason for choosing CNN in our model is that no dimension is removed from the speech signal and considering the temporal information in dynamic speech leads to more efficient and better classification. Also, the training of the CNN network in calculating the backpropagation error is adjusted by LA so that the efficiency of the proposed model is increased and the working memory part of the SOAR model can be implemented. In the proposed model, audio databases available in the field of multimodal emotion recognition eNTERFACE’ 05 and SAVEE have been used for various experiments. The recognition accuracy of the presented model in the best case from eNTERFACE’ 05 and SAVEE databases is equal to 85.3% and 84.5%, respectively.
  • Keywords
    Speech emotion recognition , Convolutional Neural Network (CNN) , Learning Automata , Improved SOAR model
  • Journal title
    Journal of Computing and Security
  • Journal title
    Journal of Computing and Security
  • Record number

    2772952