Wikipedia as an SMT Training Corpus
Dan Ioan Tufis, Radu Ion, Ștefan Daniel Dumitrescu, Dan C. Stefanescu · Recent Advances in Natural Language Processing · 2013
This article reports on mass experiments sup� porting the idea that data extracted from strongly comparable corpora may successfully be used to build statistical machine translation systems of reasonable translation quality for