Speaker verification using the spectral and time parameters of voice signal
V. N. Sorokin, A. I. Tsyplikhin · Journal of Communications Technology and Electronics · 2010
The speaker verification is based on variations in formant frequencies at stationary fragments and transient processes of vowels, the spectral features of fricative sounds, and the duration of speech segments. The best features are chosen for each word from the fixed list of Russian numerals ranging from zero to nine. The password phrase is randomly generated by the system at each verification. The compensation for dynamic noise and the counteraction with respect to interference using the reproduction of the intercepted and recorded speech are provided by the repeated reproduction of several words. The total error probabilities for male and female voices are 0.006 and 0.025%, respectively, for 30 million tests, 429 speakers, and a maximum length of the password phrase of 10 words. Note that the probabilities of false identification and false rejection are almost equal.