Linear transformations in sub-band groups for speech recognition

B. J. S. Doherty, Saeed V. Vaseghi, Paul McCourt · 1999

Linear transforms have been demonstrated to successfully achieve on-line speaker and environmental adaptation for robust recognition. This paper explores the gains in computational speed, speaker adaptation convergence rate and recognition performance obtained through the use of multi-resolution sub-band linear transforms in speech recognition. A useful feature of multiresolution processing is that significant savings can be attained as regards transform calculation. In this paper we appraise the relative merits of multiband processing over that of full-band and present evaluation results on the WSJCAM0 continuous speech database.

Read the paper · More papers on PaperTik