Polyglot Synthesis Using a Mixture of Monolingual Corpora
Javier Latorre, Koji Iwano, Sadaoki Furui · 2006
The paper proposes a new approach to multilingual synthesis based on an HMM synthesis technique. The idea consists of combining data from different monolingual speakers in different languages to create a single polyglot average voice. This average voice is then transformed into any real speaker's voice of one of these languages. The speech synthesized in this way has the same intelligibility and retains the same individuality for all the languages mixed to create the average voice, regardless of the target speaker's own language.