Evaluation of Translation Technology
Walter M. P. Daelemans, Véronique Hoste · Linguistica Antverpiensia New Series – Themes in Translation Studies · 2021
Lacking widely accepted and reliable evaluation measures, the evaluation of Machine Translation (MT) and translation tools remains an open issue.MT developers focus on automatic evaluation measures such as BLEU (Papineni et al., 2002) and NIST (Doddington, 2002) which primarily count n-gram overlap with reference translations and which are only indirectly linked to translation usability and quality.Commercial translation tools such as translation memories and translation workbenches are widely used and their developers claim usefulness in terms of productivity, consistency or quality.However, these claims are rarely proven using objective comparative studies.This collection dissects the state of the art in translation technology and translation tool development and provides quantitative and qualitative answers to the question how useful translation technology is.Evaluation of translation technology requires a multifaceted approach.It involves the evaluation of the textual output quality in terms of intelligibility, accuracy, fidelity to its source text, and appropriateness of style and register.But it also takes into account the usability of supportive tools for creating and updating dictionaries, for post-editing texts, for controlling the source language, for customization of documents, for extendibility to new languages and for domain adaptability, etc.Finally, evaluation involves contrasting the costs and benefits of translation technology with those of human translation performance.This collection comprises 10 original contributions from researchers and developers in the field.The volume is divided into two parts.The first addresses evaluation of Machine Translation, the second evaluation of Translation Tools.