Mongolian speech corpus extension for Text-to-Speech development
Altangerel Chagnaa, Kerey Esenbek, Purev Jaimai · 2013
This paper presents an extension of Mongolian speech corpus that designed for data-driven speech synthesis. The aim of the speech corpus is to develop a high-quality Mongolian TTS for blinds to use with screen reader. The new speech corpus contains nearly 10 hours of Mongolian male speech that is designed to cover all Mongolian phones. It well provides Cyrillic text transcription and its phonetic transcription with stress marking. It also provides context information including phone context, stressing levels, syntactic position in word, phrase and utterance for modeling speech acoustics and characteristics for speech synthesis.