Speaker Recognition Using Spectral Dimension Features

Wen‐Shiung Chen, Jr-Feng Huang · 2009

Biometric recognition is more and more important due to security applications all over the world. Mobile phone becomes popular in recent years. Therefore voice recognition for recognizing a speakerpsilas identity also plays a potential role. This paper presents a speaker recognition that combines a non-linear feature, named spectral dimension (SD), with Mel frequency cepstral coefficients (MFCC). In order to improve the performance of the proposed scheme, the Mel-scale method is adopted for allocating sub-bands and the pattern matching is trained by Gaussian mixture model. Some problems related to spectral dimension are discussed and the comparison with other simple spectral features is made. We observe that our proposed methods can improve the performance in different components. For instance, speaker verification combining MFCC with our proposed SD features gives a good performance of EER=2.3140% by 32_Multi-GMM. The relative improvement is about 22% better than the method that is based only on MFCC with EER=2.9631%.

Read the paper · More papers on PaperTik