Acoustic Feature Comparison of MFCC and CZT-Based Cepstrum for Speech Recognition

Zhengfeng Jiang, Hanming Huang, Shanxi Yang, Shijun Lu, Zhiqiang Hao · 2009

The speech cepstral features are important parameter in Automatic Speech Recognition (ASR), which symbolizes the property of human auditory system (HAS). The Mel-Frequency Cepstral Coefficients (MFCC) are the most widely used features in speech recognition field. This paper discusses about the algorithm of Chirp Z-Transform (CZT), and the CZT-based cepstral coefficients are proposed along with the corresponding method of feature extraction. We used MATLAB to perform the experiments. Simulation results show the correctness and effectiveness of the MFCC and the CZT-based cepstrum in speech recognition for Mandarin digits recognition. The recognition rate of MFCC algorithm is compared with Chirp Z-Transform for speech recognition system. The inclusion of cepstrum CZT-based features in parameters space may improve the correct rate of speech recognition.

Read the paper · More papers on PaperTik