Automatic detection of voice impairments due to vocal misuse by means of Gaussian mixture models
Juan Ignacio Godino-Llorente, Santiago Aguilera-Navarro, Pedro Gómez‐Vilda · 2005
There is an increasing risk of vocal and voice diseases due to the modern way of life. It is well known that most of the vocal and voice diseases cause changes in the acoustic voice signal. These diseases have to be diagnosed and treated at an early stage. Acoustic analysis is a non-invasive technique based on digital processing of speech signal. Acoustic analysis could be a useful tool to diagnose this kind of diseases, furthermore it presents several advantages: it is a non-invasive tool, provides an objective diagnostic, moreover, it can be used for the evaluation of surgical and pharmacological treatments and rehabilitation processes. ENT clinicians use acoustic voice analysis to characterise pathological voices. In this paper, we study a well known classification approach-in speaker recognition and identification-applied to the automatic detection of voice disorders. Former and actual works demonstrate that impaired voice detection can be carried out by means of supervised neural nets: multilayer perceptron. We have focused our task in detection of impaired voices by means of Gaussian mixture models and parameters such as mel frequency cepstral coefficients extracted from the windowed voice signal.