A Comparative Study of Text-Independent Speaker Identification using Statistical Features
Sherman Ong, Cheng-Hong Yang · 1998
This paper concerns a comparative study on long term text-independent speaker identification using statistical features. Performances of six statistical methods are compared. Four of them are the distance measures (the City block, the Euclidean, the Weighted Euclidean, and the Mahalanobis distance measures). The other two are the Gaussian probability density estimation and the probability estimation after the Karhunen-Loeve orthogonal transformation. Comparisons are based on two statistical tests (the Friedman test and the multiple comparison approach). Experimental results show that (1) probability calculation is generally better than most distance measures, (2) the orthogonally transformed feature vectors of dimension 15 (originally 20) performs better than all the other methods, (3) the Weighted Euclidean distance measure performs