Manual and automatic evaluation of machine translation between European languages
Philipp Koehn, Christof Monz · 2006
We evaluated machine translation performance for six European language pairs that participated in a shared task: translating French, German, Spanish texts to English and back. Evaluation was done automatically using the Bleu score and manually on fluency and adequacy.