Pronunciation assessment based upon the phonological distortions observed in language learners' utterances
Nobuaki Minematsu · 2004
Speech representation provided by acoustic phonetics, spectro-gram, is very noisy representation in that it shows every acoustic aspect of speech. Age, gender, size, shape, microphone, room and line are completely irrelevant to speech recognition, pro-nunciation assessment, and so on. But the spectrogram is af-fected easily by these factors. This is the very essential reason why speech systems are sometimes unreliable and the author supposes that the education should not endure this inevitable characteristics. The author proposed a novel method of acous-tic representation of speech where no dimensions of the above factors exist. The method was derived by implementing struc-tural phonology on physics. This paper examines whether the new representation of speech can provide a good tool of pro-nunciation assessment. Results of the experiments with good and intentionally-bad pronunciations of a single speaker showed that all the students are acoustically located between the two pronunciations, indicating that all the students are judged to be acoustically closer to the speaker than the speaker himself is. This result shows that the proposed method can delete the irrel-evant factors and is extremely reliable and effective in CALL. 1.