• DocumentCode
    2199230
  • Title

    Combining spectral features of standard and Throat Microphones for speaker identification

  • Author

    Mubeen, Nafeesa ; Shahina, A. ; Khan, A. Nayeemulla ; Vinoth, G.

  • Author_Institution
    Dept. of Inf. Technol., SSN Coll. of Eng., Chennai, India
  • fYear
    2012
  • fDate
    19-21 April 2012
  • Firstpage
    119
  • Lastpage
    122
  • Abstract
    The objective of this paper is to improve the performance of the speaker recognition system by combining speaker specific evidences present in the spectral characteristics of the standard microphone speech and the throat microphone speech. Certain vocal tract spectral features extracted from these two speech signals are distinct and could be complimentary to one another. These features could also be speech specific as well as speaker specific. These distinguishing and complimentary nature of the spectral features are due to the difference in the placement of the two microphones. Auto associative neural networks are used to model the speaker characteristics based on the system features represented by weighted linear prediction cepstral coefficients. The speaker recognition system based on Throat Microphone (TM) spectral features is comparable (though slightly less accurate) to that based on standard (or Normal) Microphone (NM) features. By combining the evidence from both the NM and TM based systems using late integration, an improvement in performance is observed from about 91% (obtained using NM features alone) to 94% (NM and TM combined). This shows the potential of combining various other speaker specific characteristics of the NM and two speech signals for further improvement in performance.
  • Keywords
    cepstral analysis; feature extraction; microphones; neural nets; speaker recognition; NM based systems; TM based systems; auto associative neural networks; normal microphone features; speaker characteristics; speaker identification; speaker recognition system; spectral features; standard microphone speech; throat microphone spectral features; throat microphone speech; vocal tract spectral features extraction; weighted linear prediction cepstral coefficients; Microphones; Neural networks; Speaker recognition; Speech; Speech recognition; Standards; Vectors;
  • fLanguage
    English
  • Publisher
    ieee
  • Conference_Titel
    Recent Trends In Information Technology (ICRTIT), 2012 International Conference on
  • Conference_Location
    Chennai, Tamil Nadu
  • Print_ISBN
    978-1-4673-1599-9
  • Type

    conf

  • DOI
    10.1109/ICRTIT.2012.6206769
  • Filename
    6206769