Corpus and Statistical Analysis of F0 Variation for Vietnamese Dialect Identification
Pham Ngoc Hung, Trịnh Văn Loan, Nguyễn Hồng Quang · Advanced science and technology letters · 2015
The performance of speech recognition systems will be improved if the corpus is organized in specialized domain and is applied in a consistent way for speech recognition in specific situations. Vietnamese dialects are various. Building of corpus for Vietnamese dialect is the first step to implement the system of dialect identification used for increasing the performance of Vietnamese recognition in general. This paper presents a method of building corpus for Vietnamese dialect identification. Vietnamese corpus VDSPEC is built with topic-based recording and tonal balance. The duration of corpus is 33.79 hours with 6 topics in total. The basic characteristics and preliminary evaluations of the corpus are also described. The statistical analysis of F0 variation showed that there are distinctions of pronunciation modality for Vietnamese tones toward Hue voice and Hanoi voice. These distinctions can be used as the important features for identifying these dialects