Solving the Polysemy Problems in A Chinese to Taiwanese TTS System by Using Rough Set Theory
Yih-Jeng Lin, Ming‐Shing Yu, Jhen-Yuan Liao · 2011
This paper proposes a rough set with confidence measurement approach to the polysemy problems in a Chinese to Taiwanese text-to-speech (TTS) system. Polysemy means there are words with more than one meaning or pronunciation. For example, there are two kinds of pronunciation for the word ”他” (he) in Taiwanese. They are /yi7/ and /yin7/. The first pronunciation, /yi7/, can mean 'he' or 'him'; and the other one, /yin7/, means 'his'. The correct pronunciation of a word affects the comprehensibility (or clarity) and fluency of Taiwanese speech. We had shown that using language models to solve the polysemy problems can have very good results. There are six words with polysemy problems to be solved in this paper. They are ”你” (you), ”我” (I), ”他” (he), ”不” (no), ”上” (up), and ”下” (down). The numbers of kinds of pronunciation of these six words are 2, 2, 2, 6, 3, and 4, respectively. Results show that the proposed approach can achieve higher accuracy for the words ”上” (up) and ”不” (no) than the existing methods.