Part-of-speech tagging and parsing of Kannada text using Conditional Random Fields (CRFs)

N M Suraksha, K Reshma, Keshav Kumar · 2017

Parts of Speech tagging is consider as the second step in Natural Language Processing. In this paper we present Parts of Speech tagging and Chunking using Conditional Random Fields. We used Kannada corpus of 3000 sentences collected from newspaper. We train with 2500 sentences and tested with 500 sentences. The comparison between Machine output and Human tagging yield an accuracy of 96.86% in tagging and chunking. We propose the parsing model for Kannada sentences using Natural Language Tool Kit (NLTK).

Read the paper · More papers on PaperTik