An end-to-end synthesis method for Korean text-to-speech systems

Yeunju Choi, Youngmoon Jung, Younggwan Kim, Youngjoo Suh, Hoirin Kim · Phonetics and Speech Sciences · 2018

A typical statistical parametric speech synthesis (text-to-speech, TTS) system consists of separate modules, such as a text analysis module, an acoustic modeling module, and a speech synthesis module. This causes two problems: 1) expert knowledge of each module is required, and 2) errors generated in each module accumulate passing through each module. An end-to-end TTS system could avoid such problems by synthesizing voice signals directly from an input string. In this study, we implemented an end-to-end Korean TTS system using Google

Read the paper · More papers on PaperTik