Chinese Speech Recognition System based on Deep Learning

Pengyuan Shao · Journal of Physics Conference Series · 2020

Abstract This paper builds a complete Chinese speech recognition system, including acoustic model and linguistic model, which can recognize the input audio signal into Chinese characters. The system realizes the modeling of acoustic model and linguistic model in speech recognition based on deep framework, of which the acoustic model is CNN-CTC and linguistic model is transformer. The data set uses THCHS-30, which refers to 30-hour Chinese speech database of Tsinghua University. The experimental results show that the Chinese speech recognition system based on deep learning achieves 90% accuracy on the test set and has an excellent effect on Mandarin speech recognition in quiet environment.

Read the paper · More papers on PaperTik