DocumentCode
2022461
Title
Estimation of the parameters of a long-term model for accurate representation of voiced speech
Author
Stettine, Yoram ; Malah, David ; Chazan, Dan
Author_Institution
Dept. of Electr. Eng., Technion, Haifa, Israel
Volume
2
fYear
1993
fDate
27-30 April 1993
Firstpage
534
Abstract
The model is able to describe the slow time-variation (nonstationary) of speech and hence enables the analysis of a whole phoneme in a single frame. This of great importance in the separation of close pitch harmonics common in speech separation problems. It also has potential in speech coding and synthesis applications. The model considered is an extension of the model proposed by L.B. Almeida and J.M. Tribolet (1983). Contrary to their model, which uses a Taylor series approximation and a fixed pitch in the analysis interval, the authors present an efficient iterative algorithm for explicit estimation of the model parameters, including the time-warping function which describes the pitch variation in the analysis frame. Preliminary simulations with voiced speech show that the model has potential in accurately describing whole voiced phonemes, even those several hundred milliseconds in duration, subject to appropriate segmentation.<>
Keywords
iterative methods; parameter estimation; speech coding; speech synthesis; close pitch harmonics; iterative algorithm; long-term model; parameter estimation; segmentation; simulations; speech coding; speech synthesis; time-warping function; voiced speech; whole voiced phonemes;
fLanguage
English
Publisher
ieee
Conference_Titel
Acoustics, Speech, and Signal Processing, 1993. ICASSP-93., 1993 IEEE International Conference on
Conference_Location
Minneapolis, MN, USA
ISSN
1520-6149
Print_ISBN
0-7803-7402-9
Type
conf
DOI
10.1109/ICASSP.1993.319362
Filename
319362
Link To Document