LJSing: Large-Scale Singing Voice Corpus of Single Japanese Singer

Takuto Fujimura, Takashi Nose, Akinori Ito · 2020

This paper describes the construction of the LJSing, a large-scale Japanese singing voice corpus by a single female singer for singing voice synthesis based on statistical methods. Singing voice synthesis systems based on machine learning have been widely studied. However, most Japanese singing voice corpora are not enough for the training of recent deep-learningbased synthesis. Furthermore, those corpora were designed without phonetic and prosodic balance. Therefore, we recorded and labeled a five-hour phonetically and prosodically balanced singing corpus sung by a Japanese singer. The corpus consists of two data sets, SongSet and PhraseSet, which are constructed based on songs and phrases, respectively.

Read the paper · More papers on PaperTik