Corpus-based comprehensive and diagnostic MT evaluation: initial Arabic, Chinese, French, and Spanish results

Kishore Papineni, Salim Roukos, Todd R. Ward, John C. Henderson, Florence M. Reeder · 2002

We describe two metrics for automatic evaluation of machine trans-lation quality. These metrics, BLEU and NEE, are compared to hu-man judgment of quality of translation of Arabic, Chinese, French, and Spanish documents into English. 1.

Read the paper · More papers on PaperTik