The Kyoto Speech-to-Speech Translation System for IWSLT 2023
Zhengdong Yang, Shuichiro Shimizu, Wangjin Zhou, Sheng Li, Chenhui Chu · 2023
This paper describes the Kyoto speech-tospeech translation system for IWSLT 2023.Our system is a combination of speech-to-text translation and text-to-speech synthesis.For the speech-to-text translation model, we used the dual-decoder Transformer model.For the text-to-speech synthesis model, we took a cascade approach of an acoustic model and a vocoder.