Voiced speech model including source-tract interaction

Ashok Kumar Krishnamurthy, Ju Li · The Journal of the Acoustical Society of America · 1988

This paper considers the problem of estimating the parameters of a voiced speech model that includes the effects of source-tract interaction. Based on work by Ananthapadmanabha and Fant [Speech Commun. 1, 167–184 (1982)], the voiced speech signal is modeled as the output of a time-varying vocal tract filter excited by a parametrized glottal source waveform. The glottal source waveform [Fant et al., STL/QPSR (1984)] models the main pulse shape of the glottal volume velocity and is described by four parameters. The effect of source-tract interaction is modeled by varying the frequency and bandwidth of the first formant in synchrony with the glottal source waveform. An analysis-by-synthesis approach is used to estimate the parameters of the model. Algorithms for estimating the parameters of the glottal source waveform and the vocal tract filter are described. Comparisons of the spectrum of the original and synthesized speech waveforms are presented. [Work supported by the OSU Seed grant program.]

Read the paper · More papers on PaperTik