Developing a Corpus for Polish Speech Enhancement by Reducing Noise, Reverberation, and Disruptions

Mariusz Kleć, Krzysztof Szklanny, Alicja Wieczorkowska · Proceedings of the International Conference on Information Systems Development · 2024

This paper presents a solution for generating corpora of simulated Polish speech recordings in complex acoustic environments. The proposed method introduces an additional layer of unpredictable sound events, in addition to the acoustic scene noise and reverberation, making the solution unique. We generated a corpus comprising over 277 hours of training examples and over 5.5 hours for testing purposes using publicly available data sources. Next, we trained the Conv-TasNet network on the generated data to enhance single speech and separate two speakers from complex noise. The results of the experiments indicated the potential of the generated corpora for solving these tasks. Researchers can use publicly available codes to create their corpora tailored to the Polish language and solve various speech-related tasks.

Read the paper · More papers on PaperTik