• DocumentCode
    868676
  • Title

    A High-Quality Speech and Audio Codec With Less Than 10-ms Delay

  • Author

    Valin, Jean-Marc ; Terriberry, Timothy B. ; Montgomery, Christopher ; Maxwell, Gregory

  • Author_Institution
    Octasic, Inc., Montreal, QC, Canada
  • Volume
    18
  • Issue
    1
  • fYear
    2010
  • Firstpage
    58
  • Lastpage
    67
  • Abstract
    With increasing quality requirements for multimedia communications, audio codecs must maintain both high quality and low delay. Typically, audio codecs offer either low delay or high quality, but rarely both. We propose a codec that simultaneously addresses both these requirements, with a delay of only 8.7 ms at 44.1 kHz. It uses gain-shape algebraic vector quantization in the frequency domain with time-domain pitch prediction. We demonstrate that the proposed codec operating at 48 kb/s and 64 kb/s out-performs both G.722.1C and MP3 and has quality comparable to AAC-LD, despite having less than one fourth of the algorithmic delay of these codecs.
  • Keywords
    algebraic codes; audio coding; speech codecs; vector quantisation; audio codec; frequency 44.1 kHz; gain-shape algebraic vector quantization; speech codec; time-domain pitch prediction; Audio coding; low-delay; speech coding; super-wideband; transform coding;
  • fLanguage
    English
  • Journal_Title
    Audio, Speech, and Language Processing, IEEE Transactions on
  • Publisher
    ieee
  • ISSN
    1558-7916
  • Type

    jour

  • DOI
    10.1109/TASL.2009.2023186
  • Filename
    4926218