Corpus-driven bilingual lexicon extraction
Michael Rosner · OAR@UM (University of Malta) · 2004
Abstract. This paper introduces some key aspects of machine translation in order to situate the role of the bilingual lexicon in transfer-based systems. It then discusses the data-driven approach to extracting bilingual knowledge automatically from bilingual texts, tracing the processes of alignment at different levels of granularity. The paper concludes with some suggestions for future work. 1 Machine Translation transfer parsing generation source language parse tree target language parse tree source language words target language words Fig. 1. The Machine Translation (MT) problem is almost as old as Computer Science, the first efforts at