Bilingual Dictionary Extraction for Special Domain Based on Web Data
Yongchen Zhang · Zhongwen xinxi xuebao · 2006
Bilingual dictionary is the base of many NLP applications such as multi-lingual information retrieval and machine translation.This paper proposes a method of extracting bilingual dictionary for the special domain from the non-parallel corpora: first,discusses the fundamental postulate and reviews the related research,second,presents an algorithm of extracting the bilingual dictionary for the special domain based on the non-parallel corpora with the word relation matrix,and finally,analyzes the influence of the seed word on the extraction of the bilingual dictionary with abundant of experimentation.The experiments demonstrate that the quantity and average frequency of the seed word pairs contribute to the results effectively.