• DocumentCode
    2705887
  • Title

    Speech Recognition System Combination for Machine Translation

  • Author

    Gales, Mark J.F. ; Liu, Xindong ; Sinha, Roopak ; Woodland, Philip C. ; Yu, Kaiyuan ; Matsoukas, Spyros ; Ng, Timothy ; Nguyen, Khanh ; Nguyen, L. ; Gauvain, J. -L. ; Lamel, Lori ; Messaoudi, A.

  • Author_Institution
    Cambridge Univ., UK
  • Volume
    4
  • fYear
    2007
  • fDate
    15-20 April 2007
  • Abstract
    The majority of state-of-the-art speech recognition systems make use of system combination. The combination approaches adopted have traditionally been tuned to minimising word error rates (WERs). In recent years there has been a growing interest in taking the output from speech recognition systems in one language and translating it into another. This paper investigates the use of cross-site combination approaches in terms of both WER and impact on translation performance. In addition, the stages involved in modifying the output from a speech-to-text (STT) system to be suitable for translation are described. Two source languages, Mandarin and Arabic, are recognised and then translated using a phrase-based statistical machine translation system into English. Performance of individual systems and cross-site combination using cross-adaptation and ROVER are given. Results show that the best STT combination scheme in terms of WER is not necessarily the most appropriate when translating speech.
  • Keywords
    error statistics; language translation; speech recognition; Arabic language; English language; Mandarin language; ROVER; cross-site combination; machine translation; phrase-based statistical machine translation system; speech-to-text system; state-of-the-art speech recognition system; system combination; word error rates; Appropriate technology; Computer numerical control; Costs; Diversity reception; Error analysis; Natural languages; Speech recognition; Surface-mount technology; Voting; Machine Translation; Speech Recognition;
  • fLanguage
    English
  • Publisher
    ieee
  • Conference_Titel
    Acoustics, Speech and Signal Processing, 2007. ICASSP 2007. IEEE International Conference on
  • Conference_Location
    Honolulu, HI
  • ISSN
    1520-6149
  • Print_ISBN
    1-4244-0727-3
  • Type

    conf

  • DOI
    10.1109/ICASSP.2007.367310
  • Filename
    4218341