Using Parallel Texts and Lexicons for Verbal Word Sense Disambiguation
Ondřej Dušek, Eva Fučíková, Jan Hajič, Martin Popel, Jana Šindlerová, Zdeňka Urešová · 2015
We present a system for verbal Word Sense Disambiguation (WSD) that is able to exploit additional information from parallel texts and lexicons. It is an extension of our previous WSD method (Dusek et al., 2014), which gave promising results but used only monolingual features. In the follow-up work described here, we have explored two additional ideas: using English-Czech bilingual resources (as features only – the task itself remains a monolingual WSD task), and using a “hybrid” approach, adding features extracted both from a parallel corpus and from manually aligned bilingual valency lexicon entries, which contain subcategorization information. Albeit not all types of features proved useful, both ideas and additions have led to significant improvements for both languages explored.