Closed-form estimation of the amplitude commands in the automatic extraction of the Fujisaki's model
S.D.S. Silva, Sergio Lima Netto · 2004
Generation of F0 contours is required for natural-sounding text-to-speech systems. This task can be accomplished using the Fujisaki model, proven to be very good to describe F0 contours based on simple linguistically motivated parameters. However, the extraction of the Fujisaki model parameters is a very intricate problem. Several methods were proposed to solve this problem using iterative optimization techniques. This paper presents a new method capable of extracting the amplitude parameters of the Fujisaki model analytically. The time-marking commands are still obtained via iterative optimization. The result is a more accurate and less computationally intensive amplitude determination due to the proposed closed-form solution. Examples are included illustrating the application of the proposed method.