Feature divergence of pathological speech
Steven Sandoval, Rene L. Utianski, Visar Berisha, Julie Liss, Andreas Spanias · The Journal of the Acoustical Society of America · 2013
Many state of the art speaker verification systems are implemented by modeling the probability distribution of a feature set using Gaussian mixture models. In these systems, a decision is made by comparing a likelihood of an observation using both a Gaussian mixture model corresponding to an individual, and a Gaussian mixture model universal back ground model. In this study we propose to use a similar framework to instead characterize the divergence of the feature set distribution between healthy and pathological speech. We accomplish this by determining the difference between a universal background model trained on healthy speech and model of an individual's pathological speech. There are several known methods to evaluate the difference between two probability distributions, one example being the Kullback-Leibler divergence. By building a universal background model using healthy speech, we hope to capture the expected distribution of our feature space. Then by computing a difference between a dysathric individual's feature distribution, and the universal background model, we can determine the features that are most likely to capture the effects of a specific motor speech disorder.