Phoneme modelling using continuous mixture densities

Hermann Ney, Andreas Noll · 2003

Deals with the use of continuous mixture densities for phenome modelling in large vocabulary continuous speech recognition. The concept of continuous mixture densities is applied to the emission probability density functions of hidden Markov models for phonemes in order to take into account phonetic-context dependencies. It is shown that the advantage of continuous mixture densities is the ability to lead to parameter estimates that are accurate and at the same time robust with respect to the limited amount of training data. Training and recognition algorithms for mixture densities in the framework of phoneme modelling are described. Recognition results for a 917-word task, requiring only 7 min of speech for training and an overlap of 43 words between training vocabulary and test vocabulary, are presented.>

Read the paper · More papers on PaperTik