Statistical Language Models for Croatian Weather-domain Corpus

Lucia Načinović Prskalo, Sanda Martinčić-Ipšić, Ivo Ipšić · Repozitorij Filozofskog fakulteta u Zagrebu' at University of Zagreb (University of Zagreb) · 2009

Statistical language modelling estimates the regularities in natural languages.Language models are used in speech recognition, machine translation and other applications for speech and language technologies.In this paper we will present a procedure for language models building for the Croatian weatherdomain corpus.Different types of n-gram statistic language models and smoothing methods for language modelling are presented.Those models are compared in terms of their estimated perplexity.

Read the paper · More papers on PaperTik