A Japanese language speech database

S. Itahashi · 2005

A database of spoken Japanese sound has been collected for use in designing and evaluating algorithms for automatic speech recognition. The database is composed of 323 words. It is a special feature of this database that all samples are uttered four times by each speaker (i. e. four tokens per word). Seventy-five male and 75 female data are collected at 15 recording places. Speaker data include sex, age, height, etc. Fifteen research institutions and private enterprises engaged in speech research and development have taken part in the data collection. This is a result of four years' effort by a committee supported by JEIDA (Japan Electronic Industry Development Association). Part of the database has been distributed among the members of the committee.

Read the paper · More papers on PaperTik