Mixed-phase AR models for voiced speech and perceptual cost functions

William R. Gardner, Bhaskar D. Rao · 2002

Mixed-phase AR models are introduced for encoding the magnitudes and phases of the harmonics of voiced speech. Motivation for the use of the mixed-phase AR models is given and several cost functions are introduced, forming the basis for algorithms which estimate the model parameters. An efficient algorithm based on a quasi-linear least squares approach is presented, and a more sophisticated algorithm based on the perceptual masking properties of the ear is described. When the algorithms are used to model voiced speech signals using a 14th order mixed-phase model, high quality speech can be produced.>

Read the paper · More papers on PaperTik