Performance Analysis of Implemented MFCC and HMM-based Speech Recognition System
Marlyn Maseri, Mazlina Binti Mamat · 2020
This paper describes the performance analysis of designed speech recognition system whereby the front end method uses MFCCs feature extraction algorithm and defined HMM recognition as the back end. The dataset includes 30 phonemes and 2200 utterances by different speakers. Each speech signal is sampled to 16kHz, 16-bit PCM, and in a mono channel format. The extracted feature of each signal consists of 39 feature vectors which are 12 Mel Cepstrum Coefficients, Log Energy, Delta (first-order derivative) coefficients, and Acceleration coefficients (second-order derivative). The Baum-Welch algorithm is applied for HMM training and the Viterbi algorithms for decoding. The overall system performance accuracy of this experiment is 95.00%.