Noise estimation without explicit speech, non-speech detection: a comparison of mean, modal and median based approaches
Nicholas Evans, John S. D. Mason · 2001
Automatic speech recognition performance tends to be degraded in noisy conditions. Spectral subtraction is a simple, popular approach of noise compensation. In conventional spectral subtraction [1, 2], noise statistics are updated during speech gaps and subtracted from a corrupt signal during speech intervals. Some means of explicit speech, non-speech detection is therefore essential. Recent proposals have avoided the problem of speech, non-speech detection [3, 4, 5, 6, 7] by continually updating noise estimates whether speech is present or not.