Word accentuation prediction using a neural net classifier.

Taniya Mishra, Emily Tucker Prud’hommeaux, Jan P. H. van Santen · 2007

Automaticpredictionof pitch accentassignmentis an important but challenging task in text-to-speech synthesis (TTS). Early work in accent prediction relied on simple word-class distinctions, but recentlymore sophisticatedinductive learningmodels using multiple features have been applied to the problem. For our neural network accent classifier, we developed a corpus that was labeled according to judgments of accent assignment appropriatenessin synthesized speech rather than the usual ToBI annotation guidelines. Because the resulting training set was imbalanced, the baseline neural network we developed for this task had a very high accuracy rate (84%) but performed only slightly better than chance accordingto our ROC analysis. Balancing our training data using downsizing, oversampling, and cost-based post-processing yielded significant improvement in this informative measure. We anticipate that balance adjustments and the inclusion of more complex features will lead to further improvement. 1.

Read the paper · More papers on PaperTik