On using formants to improve SCHMM speaker adaptation
Tae-Young Yang, Won-Ho Shin, Weon-Goo Kim, Dae Hee Youn, Il-Whan Cha · IEEE Transactions on Speech and Audio Processing · 1999
A speaker adaptation algorithm using formant frequencies is proposed. The formants extracted from the cepstral means in the reference codebook are iteratively shifted toward the formants of a test speaker. The number of cepstral means selected at each iteration decreases as the iteration increases. The decision of the number of selected cepstral means and a formant based distance measure are formulated. The proposed algorithm was implemented in two schemes and evaluated by speaker-independent, male speaker dependent, and female speaker-dependent recognition experiments. A combined scheme with the Bayesian adaptation obtained 9.7% enhancement for the average recognition accuracy in speaker-independent experiments and 52.6% in speaker-dependent recognition experiments.