Context-adaptive smoothing for concatenative speech synthesis

Ki-Seung Lee, Sang-Ryoung Kim · IEEE Signal Processing Letters · 2002

In text-to-speech synthesis, spectral smoothing is often employed to reduce artifacts at unit-joining points. A context-adaptive smoothing method is proposed in this letter, where the amount of smoothing is determined according to context information. Discontinuities at unit boundaries are predicted by a regression tree, and smoothing factors are computed by using predicted discontinuities and real discontinuities at unit boundaries. Experimental results are presented to demonstrate the effectiveness of the proposed method.

Read the paper · More papers on PaperTik