Shape-invariant pitch-synchronous text-to-speech conversion

Eduardo Rodríguez Banga, Carmén García Mateo · 2002

Text-to-speech (T-T-S) systems based on the concatenation of speech units need a prosodic modification algorithm to adjust the prosodic features of the stored speech units to the desired output values. We discuss the application of a sinusoidal shape-invariant model to a T-T-S system for Spanish, paying special attention to the concatenation issues and phase treatment. The resulting speech waveform resembles the waveform of its contributory units, without sounding reverberant as in other sinusoidal implementations.

Read the paper · More papers on PaperTik