Machine Translation Evaluation Inside QARLA
Jesus Vicente Gimenez, Enrique Amigó, Chiori Hori, Educación Distancia · 2005
In this work we present the fundamentals of the IQMT frame-work for MT evaluation. IQMT offers a common work-bench on which existing evaluation metrics can be utilized. We suggest the IQ measure and test it on the Chinese-to-English data from the IWSLT 2004 Evaluation Campaign. We show how the correlation with human assessments at the system level improves substantially for most individual met-rics. Moreover, IQMT allows to robustly combine several met-rics avoiding scaling problems and metric weightings. Sev-eral metric combinations were tried, but correlations did not further improve significantly. 1.