Exploiting the potential of auditory preprocessing for robust speech recognition by locally recurrent neural networks

Klaus Kasper, Herbert Reininger, D. Wolf · 2002

We present a robust speaker independent speech recognition system consisting of a feature extraction based on a model of the auditory periphery, and a locally recurrent neural network for scoring of the derived feature vectors. A number of recognition experiments were carried out to investigate the robustness of this combination against different types of noise in the test data. The proposed method is compared with cepstral, RASTA, and JAH-RASTA processing for feature extraction and hidden Markov models for scoring. The presented results show that the information in features from the auditory model can be best exploited by locally recurrent neural networks. The robustness achieved by this combination is comparable to that of JAH-RASTA in combination with HMM but without any requirement for an explicit adaptation to the noise in speech pauses.

Read the paper · More papers on PaperTik