Computational analysis of Mandarin sounds with reference to the English language
Ching Y. Suen · 1982
In the analysis of Mandarin, the author used a corpus composed of over 750,000 samples transcribed automatically from Chinese characters by the computer through the sequential application of a set of phonetic rules developed by the author. The result is a classification and rank distribution of all speech sounds, the phonetic properties, frequency distribution of symbols, phonemes, syllables, tones, and their combinations. These statistical properties are compared with those of the English language.