Automatic Language Identification from Spectorgam Images

Kaiyr Aizada, Shirali Kadyrov, Andrey Bogdanchikov · 2021 IEEE International Conference on Smart Information Systems and Technologies (SIST) · 2021

The main idea of the work is to apply CNN and LSTM algorithms on the preprocessed audio data converted to spectrograms. In the experiment 7 languages are used English, Kazakh, French, German, Italian, Russian and Spanish to test that algorithm which show over 99% training accuracy and the maximum of 94.28% accuracy on the test set. In the end there is a discussion of the 100% of classification of the Kazakh language and the possible influences of dataset on the result.

Read the paper · More papers on PaperTik