An end-to-end synthesis method for Korean text-to-speech systems
Yeunju Choi, Youngmoon Jung, Younggwan Kim, Youngjoo Suh, Hoirin Kim · Phonetics and Speech Sciences · 2018
A typical statistical parametric speech synthesis (text-to-speech, TTS) system consists of separate modules, such as a text analysis module, an acoustic modeling module, and a speech synthesis module. This causes two problems: 1) expert knowledge of each module is required, and 2) errors generated in each module accumulate passing through each module. An end-to-end TTS system could avoid such problems by synthesizing voice signals directly from an input string. In this study, we implemented an end-to-end Korean TTS system using Google