MT evaluation
Enrique Amigó, Jesus Vicente Gimenez, Julio A. Gonzalo, Lluı́s Màrquez · 2006
We present a comparative study on Machine Translation Evaluation according to two different criteria: Human Likeness and Human Acceptability. We provide empirical evidence that there is a relationship between these two kinds of evaluation: Human Likeness implies Human Acceptability but the reverse is not true. From the point of view of automatic evaluation this implies that metrics based on Human Likeness are more reliable for system tuning.