• DocumentCode
    638692
  • Title

    Extraction of glottal features for speaker recognition

  • Author

    Ostrogonac, Stevan ; Secujski, Milan ; Knezevic, Dragan ; Suzic, Sinisa

  • Author_Institution
    Fac. of Tech. Sci., Univ. of Novi Sad, Novi Sad, Serbia
  • fYear
    2013
  • fDate
    8-10 July 2013
  • Firstpage
    369
  • Lastpage
    373
  • Abstract
    This paper presents an extension to the SEDREAMS algorithm for extracting the information about glottal opening and glottal closure instants (GCI and GOI) directly from the speech signal. Accurate detection of GCIs and GOIs is crucial for estimating the glottal features which are to be used in speaker recognition systems. Many different approaches resulted in a variety of algorithms dealing with this problem. The algorithm that showed the best results so far consists of two steps. First, a mean-based signal is computed to determine the intervals in which the GC! and GOI moments should be searched for. Then, discontinuities are sought in the LP residual of the speech signal and they represent the estimation of glottal events. This algorithm (in literature found under the name SEDREAMS) is widely used in glottal excitation estimation systems. However, the mean-based signal calculated in the first step of the process sometimes contains unwanted spectral components which significantly degrade the performances. This paper describes one way to address this problem. By applying an adaptive filter to the mean-based signal significant improvement has been achieved in glottal features estimation. This was confirmed by a speaker recognition experiment which showed very encouraging results.
  • Keywords
    adaptive filters; feature extraction; speaker recognition; GCI; GOI; LP residual; SEDREAMS algorithm; adaptive filter; glottal closure instant; glottal excitation estimation systems; glottal feature extraction; glottal opening instant; mean-based signal; speaker recognition systems; speech signal; Adaptive filters; Estimation; Feature extraction; Filtering; Speaker recognition; Speech; Speech recognition;
  • fLanguage
    English
  • Publisher
    ieee
  • Conference_Titel
    Computational Cybernetics (ICCC), 2013 IEEE 9th International Conference on
  • Conference_Location
    Tihany
  • Print_ISBN
    978-1-4799-0060-2
  • Type

    conf

  • DOI
    10.1109/ICCCyb.2013.6617621
  • Filename
    6617621