Continuous Sinhala Speech Recognizer

Thilini Nadungodage, Ruvan Weerasinghe · 2011

Automatic Speech Recognition has been successfully developed for many Western languages including English. Continuous, speaker independent speech recognition however has still not achieved high levels of accuracy owing to the variations in pronunciation between members of even a single community. This paper describes an effort to implement a speaker dependent continuous speech recognizer for a less resourced non-Latin language, namely, Sinhala. Using readily available open source tools, it shows that fairly accurate speaker dependent ASR systems for continuous speech can be built for newly digitized languages. The paper is expected to serve as a starting point for those interested in initiating projects in speech recognition for such ‘new’ languages from non-Latin linguistic traditions.

Read the paper · More papers on PaperTik