Text independent speaker identification with finite multivariate generalised Gaussian mixture model and k-means algorithm
V. Sailaja, Kommanaboyana Srinivasa Rao, K. V. V. S. Reddy · International Journal of Signal and Imaging Systems Engineering · 2013
In this paper, we propose text independent speaker identification with a finite multivariate generalised Gaussian Mixture Model (GMM) with a k–means algorithm. Each speaker's speech spectra are characterised with a mixture of generalised Gaussian distribution that includes Gaussian and Laplacian distribution as a particular case. Speech analysis is done with the Mel Frequency Cepstral Coefficients (MFCC) extracted from the front end process. Using the EM algorithm and k–means algorithm the model parameters the numbers of acoustic classes associated with each speech spectra are determined. The performance of the proposed algorithm is studied through experimental evaluation and observed that this algorithm outperforms the existing speaker identification algorithm with GMM. It is also observed that this algorithm performs efficiently even with a heterogeneous population with small (less than 2 seconds) utterances.