Theoretical analysis of biased MMSE short-time spectral amplitude estimator and its extension to musical-noise-free speech enhancement
Shunsuke Nakai, Hiroshi Saruwatari, Ryoichi Miyazaki, Satoshi Nakamura, Kazunobu Kondo · 2014
In this paper, we provide a theoretical analysis of the minimum mean-square error short-time spectral amplitude (MMSE-STSA) estimator with the biased a priori SNR estimation and its extension to musical-noise-free speech enhancement. Recently, musical-noise-free speech enhancement has been proposed, where no musical noise is generated in iterative spectral subtraction. However, no existence of the musical-noise-free condition in the MMSE-STSA estimator has been reported. Therefore, in this paper, we show that the musical-noise-free condition exists in the biased MMSE-STSA estimator via the theoretical analysis. In addition, we perform comparative experiments and clarify the efficacy of the proposed musical-noise-free speech enhancement.