Voice source parameters estimation by fitting the glottal formant and the inverse filtering open phase

Thomas Drugman, Thomas M. DuBuisson, Nicolas d’Alessandro, Alexis Moinet, Thierry Dutoit · 2008

This paper presents two approaches to the problem of ex-tracting the parameters of the LF source model directly from the speech waveform. The first approach relies on the glot-tal formant estimated from the anticausal contribution of speech. Indeed the ZZT technique has recently shown its ability to deconvolve speech into its causal and anticausal components. The second method is based on the glottal open phase obtained by inverse filtering. The notion of unanalyz-able frames and the way to detect and correct them are also presented. Once source parameters are extracted, the coef-ficients of the ARX speech production model are estimated by spectral division. Decomposition on both synthetic and natural speech, as well as an analysis-synthesis test confirm the accuracy of methods exposed. 1.

Read the paper · More papers on PaperTik