The SPHINX speech recognition system
K.-F. Lee, Hsiao-Wuen Hon, Minjoo Hwang, Sanjoy Mahajan, R. Reddy · International Conference on Acoustics, Speech, and Signal Processing · 2003
A description is given of SPHINX an accurate large-vocabulary speaker-independent continuous speech recognition system. The authors have made several recent enhancements, including generalized triphone models, word duration modeling, function-phrase modeling, between-word coarticulation modeling, and corrective training. On the 997-word resource management task, SPHINX attained a word accuracy of 96% with a grammar (perplexity 60), and 82% without grammar (perplexity 997).>