A full-band adaptive harmonic representation of speech
Gilles Degottex, Yannis Stylianou · 2012
In this paper we present a full-band Adaptive Harmonic Model (aHM) that is able to accurately reconstruct stationary and non stationary parts of speech. The model does not require any voiced/unvoiced decision, neither an accurate estimation of the pitch contour. Its robustness is based on the previously sug-gested adaptive Quasi-Harmonic model (aQHM), which pro-vides a mechanism for frequency correction and adaptivity of its basis functions to the characteristics of the input signal. The suggested method overcomes limitations of the initial method based on aQHM in detecting frequency tracks over time, espe-cially at mid and high frequencies, by employing a bandlimited iterative procedure for the re-estimation of the fundamental fre-quency. Listening tests show that reconstructed speech using aHM is mainly indistinguishable from the original signal, out-performing standard sinusoidal models (SM) and the aQHM-based method, while it uses less parameters for the reconstruc-tion than SM.