• DocumentCode
    1168743
  • Title

    Reviewing automatic language identification

  • Author

    Muthusamy, Yeshwant K. ; Barnard, Etienne ; Cole, Ronald A.

  • Author_Institution
    Syst. & Inf. Sci. Lab., Texas Instrum. Inc., Dallas, TX, USA
  • Volume
    11
  • Issue
    4
  • fYear
    1994
  • Firstpage
    33
  • Lastpage
    41
  • Abstract
    The Oregon Graduate Institute Multi-language Telephone Speech Corpus (OGI-TS) was designed specifically for language identification research. It currently consists of spontaneous and fixed-vocabulary utterances in 11 languages: English, Farsi, French, German, Hindi, Japanese, Korean, Mandarin, Spanish, Tamil, and Vietnamese. These utterances were produced by 90 native speakers in each language over real telephone lines. Language identification is related to speaker-independent speech recognition and speaker identification in several interesting ways. It is therefore not surprising that many of the recent developments in language identification can be related to developments in those two fields. We review some of the more important recent approaches to language identification against the background of successes in speaker and speech recognition. In particular, we demonstrate how approaches to language identification based on acoustic modeling and language modeling, respectively, are similar to algorithms used in speaker-independent continuous speech recognition. Thereafter, prosodic and duration-based information sources are studied. We then review an approach to language identification that draws heavily on speaker identification. Finally, the performance of some representative algorithms is reported.<>
  • Keywords
    acoustic signal processing; natural languages; speech recognition; stochastic processes; English; Farsi; French; German; Hindi; Japanese; Korean; Mandarin; Multi-language Telephone Speech Corpus; OGI-TS; Oregon Graduate Institute; Spanish; Tamil; Vietnamese; acoustic modeling; automatic language identification; language identification research; language modeling; speaker identification; speaker recognition; speaker-independent speech recognition; speech recognition; Auditory system; Delay; Displays; Frequency; Humans; Loudspeakers; Natural languages; Speech recognition; Stress; Telephony;
  • fLanguage
    English
  • Journal_Title
    Signal Processing Magazine, IEEE
  • Publisher
    ieee
  • ISSN
    1053-5888
  • Type

    jour

  • DOI
    10.1109/79.317925
  • Filename
    317925