Design and Development of a Prosody Generator for Arabic TTS Systems

Zied Mnasri, Fatouma Boukadida, Noureddine Ellouze · International Journal of Computer Applications · 2010

Prosody modeling has become the backbone of TTS synthesis systems.Amongst all the prosodic modeling approaches, phonetic methods aiming to predict duration and F0 contour are being very praised, thanks to the development of regression tools, such as neural networks (NN).Besides, parametric representations like Fujisaki model for F 0 contour generation help to reduce the problem into the approximation of parameters only.But, prior to the prediction process, text analysis should be carried out first, to select and encode the necessary input features.In our purpose to promote Arabic TTS synthesis, an Integrated Model of Arabic Prosody for Speech Synthesis (IMAPSS) tool has been designed to integrate our developed models for text analysis, NN-based phonemic duration prediction and Fujisaki-inspired F 0 contour.Hence, the yielding parameters provide a command file to be read by speech synthesis systems, like MBROLA.

Read the paper · More papers on PaperTik