Feature extraction for automatic speech recognition (ASR)
Brandon T. Swartz, Neeraj Magotra · 2002
This paper presents a new speech feature extraction technique for use in automatic speech recognition (ASR). The technique is based on a new two-dimensional series expansion that is applied to the spectrogram of a sampled speech signal. The series expansion allows for global analysis in frequency and local multiresolution analysis in time. Multiresolution analysis in time is useful because the duration of vowels is almost an order of magnitude greater than that of consonants.