Acoustic modeling and language modeling for cantonese LVCSR

Y. W. Wong, Ka-Fai Chow, Wai H. Lau, Wan-Yi Lo, Tan Lee, P.C. Ching · 1999

This paper describes our recent work on the development of a large-vocabulary, speaker-independent continuous speech recognition system for Cantonese (a major Chinese dialect). Both acoustic modeling and language modeling are being addressed. For acoustic modeling, we focus on right-context-dependent sub-syllable units. Tying of HMM at model as well as state level is applied based on phonetic knowledge and the decision-tree approach. Statistical language model is built from large amount of newspaper text. The overall recognition accuracy for syllable and Chinese character are 81.83% and 68.94% respectively. Keywords: LVCSR, Cantonese speech recognition, acoustic modeling, language modeling 1 INTRODUCTION Cantonese is one of the major Chinese dialects spoken by tens of millions of people in Hong Kong, Southern China as well as many overseas Chinese communities. With the great advancement of computer and information technology, there is an ever-increasing demand of largevocabulary con...

Read the paper · More papers on PaperTik