Removing the distinction between a translation memory, a bilingual dictionary and a parallel corpus
Vincent Vandeghinste · 2007
This paper presents a prototype MT system which does not make the dis-tinction between a dictionary, a sub-sentential aligned parallel corpus, and post-edited information (translators output) like a translation memory. The system is based on the METIS-approach (Vandeghinste et al, 2006), and uses an XML-based dictionary format in which not only simple word-to-word translations can be included, but which also contains complex dictionary en-tries, including discontinuous entries, like idioms and proverbs. The pre-sented prototype is a system that automatically adapts its dictionary and tar-get language corpus depending on the post-edited output as made by the users of the system, and will therefore have a learning curve in its performance. 1 1