MT evaluation

Enrique Amigó, Jesus Vicente Gimenez, Julio A. Gonzalo, Lluı́s Màrquez · 2006

We present a comparative study on Machine Translation Evaluation according to two different criteria: Human Likeness and Human Acceptability. We provide empirical evidence that there is a relationship between these two kinds of evaluation: Human Likeness implies Human Acceptability but the reverse is not true. From the point of view of automatic evaluation this implies that metrics based on Human Likeness are more reliable for system tuning.

Read the paper · More papers on PaperTik