Synthesizing a human-like voice is the easy way
Sébastien Le Maguer, Benjamin R. Cowan · 2021
Deep-learning based technologies produce speech that is almost indistinguishable from humans. However, focusing on producing human-like voices poses ethical, security and societal issues. Considering the flexibility and the regression power of new technologies based on deep-learning, it is now time to consider a new type of synthesis: natural non-human-like speech synthesis. This paper aims to convince you that such research opens new research directions, that it brings another perspective to address human-like speech challenges, and that enough material is available to start to investigate non-human-like speech.