An integrated approach to robust speaker identification and speech recognition
Chiman Kwan, Jihao Yin, Bulent Ayhan, S. Chu, X. Liu, K. Puckett, Yunxin Zhao, K. C. Ho, Martin Krüger, I. Sityar · 2008
Conventional speaker identification and speech recognition algorithms cannot deal with noisy and multiple speaker environments. For example, IBM via Voice has low recognition rates if dictation is done in a noisy environment. In order to achieve high performance in speaker identification and speech recognition, we propose an integrated approach that takes every facet of the process into account. Here we summarize some preliminary results from the application of this integrated approach to robust speaker identification and speech recognition. A real-time stand-alone software prototype has been developed to evaluate the effectiveness of the approach.