Word Sense Disambiguation in Bengali Using Sense Induction

Anindya Sau, Tarik Aziz Amin, Nabagata Barman, Alok Ranjan Pal · 2019

In this paper an algorithm is proposed for Word Sense Disambiguation in Bengali language using Sense Induction technique. The overall work is carried out in two phases. In the first phase, different sense clusters are created using Sense Induction technique and in the second phase, Word Sense Disambiguation is developed using Semantic Similarity Measure. The data sets are prepared from the corpus, developed under the TDIL (Technology Development for Indian Languages) project of the Government of India. The developed model is tested on 10 commonly used Bengali ambiguous words, each of which is having approximately 200 sentences. The overall accuracy is achieved as 63.71% in Word Sense Disambiguation task. The challenges and the pitfalls of this work are explained in detail at the end of this paper.

Read the paper · More papers on PaperTik