Language modelling for Turkish as an agglutinative language

Tolga Çiloğlu, M. Comez, S. Sahin · 2004

Two types of language models have been considered for Turkish continuous speech recognition. In one case, words are separated into their stems and the rest, and language models are calculated based on this new set of units. In the other case, words are considered as a whole, but language models are calculated with respect to the stems of the words. Studies are carried out for bigram and trigram formalisms.

Read the paper · More papers on PaperTik