A Comparative Survey on Parts of Speech Taggers for the Marathi Language

Ishita Tambat, Minakshee Narayankar · 2022 4th International Conference on Inventive Research in Computing Applications (ICIRCA) · 2022

Natural Language Processing relies heavily on the POS tagger. The POS tagger is a useful tool for tagging each word in a phrase with parts of speech tags. NLP Applications performing various tasks use POS tagging as a crucial initial step. In terms of data tagging, the POS tagger for English is widely available, however, there is no similar tagger for Marathi. Marathi is a morphologically complex language with regional speech differences. Because of the ambiguity in the language, as well as its highly inflectional structure and free word order, establishing a successful POS tagger in Marathi is tough. This article provides a comprehensive overview of POS tagging for the Marathi language and its variations. Various POS Tagging models and approaches are investigated in this research. Tokenization in computer languages is similar to tagging in natural language processing. Choosing the appropriate tag for the scenario might be challenging for POS taggers. Research has been conducted to discover a solution to this conundrum.

Read the paper · More papers on PaperTik