The RWTH speech recognition system and spoken document retrieval

Hermann Ney, Lutz Welling, Stefan Ortmanns, Klaus Beulen, Frank Wessel · 2002

We present an overview of the RWTH Aachen large vocabulary continuous speech recognizer. The recognizer is based on continuous density hidden Markov models and a time-synchronous left-to-right beam search strategy. Experimental results on the ARPA Wall Street Journal (WSJ) corpus verify the effects of several system components, namely linear discriminant analysis, vocal tract normalization, pronunciation lexicon and cross-word triphones, on the recognition performance. Finally, the extension of the recognition system towards spoken document retrieval is discussed.

Read the paper · More papers on PaperTik