FCC: Modeling Probabilities with GIZA++ for Task 2 and 3 of SemEval-2

Darnes Vilariño Ayala, Carlos Balderas Posada, David Pinto, Miguel Rodríguez Hernández, Saúl León Silverio · Meeting of the Association for Computational Linguistics · 2010

In this paper we present a naive approach to tackle the problem of cross-lingual WSD and cross-lingual lexical substitution which correspond to the Task #2 and #3 of the SemEval-2 competition. We used a bilingual statistical dictionary, which is calculated with Giza++ by using the EU-ROPARL parallel corpus, in order to calculate the probability of a source word to be translated to a target word (which is assumed to be the correct sense of the source word but in a different language). Two versions of the probabilistic model are tested: unweighted and weighted. The obtained values show that the unweighted version performs better thant the weighted one.

Read the paper · More papers on PaperTik