ccGigafida ARPA language model 1.0

Jože Kadivec, Marko Robnik‐Šikonja, Špela Vintar · Americanae (AECID Library) · 2017

The ccGigafida ARPA language model was created from the ccGigafida written corpus of Slovenian (https://www.clarin.si/repository/xmlui/handle/11356/1035) using the KenLM algorithm in the Moses machine translation framework. It is a general language model of contemporary standard Slovenian language that can be used as a language model in statistical machine translation systems. The language model was created as a part of the master thesis: Kadivec, Jože. 2016. Prilagoditev statističnega strojnega prevajalnika za specifično domeno v slovenskem jeziku (Domain specific adaptation of a statistical machine translation engine in Slovene language). Master's thesis, Faculty of computer and information science, University of Ljubljana. https://repozitorij.uni-lj.si/IzpisGradiva.php?id=84815

Read the paper · More papers on PaperTik