Text-Free Image-to-Speech Synthesis Using Learned Segmental Units

Wei-Ning Hsu, David F. Harwath, Tyler Miller, Christopher Song, James Glass · 2021

Wei-Ning Hsu, David Harwath, Tyler Miller, Christopher Song, James Glass. Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers). 2021.

Read the paper · More papers on PaperTik