Perceptual Properties of Current Speech Recognition Technology

Hynek Heřmanský, Jordan R. Cohen, Richard M. Stern · Proceedings of the IEEE · 2013

In recent years, a number of feature extraction procedures for automatic speech recognition (ASR) systems have been based on models of human auditory processing, and one often hears arguments in favor of implementing knowledge of human auditory perception and cognition into machines for ASR. This paper takes a reverse route, and argues that the engineering techniques for automatic recognition of speech that are already in widespread use are often consistent with some well-known properties of the human auditory system.

Read the paper · More papers on PaperTik