Issues and Challenges in Marathi Named Entity Recognition
Nita Sanjay Patil, Ajay S. Patil, Pawar B.V · International Journal on Natural Language Computing · 2016
Information Extraction (IE) is a sub discipline of Artificial Intelligence. IE identifies information inunstructured information source that adheres to predefined semantics i.e. people, location etc. Recognition of named entities (NEs) from computer readable natural language text is significant task of IE and natural language processing (NLP).Named entity (NE) extraction is important step for processing unstructured content.Unstructured data is computationally opaque.Computers require computationally transparent data for processing.IE adds meaning to raw data so that it can be easily processed by computers.There are various different approaches that are applied for extraction of entities from text.This paper elaborates need of NE recognition for Marathi and discusses issues and challenges involved in NE recognition tasks for Marathi language.It also explores various methods and techniques that are useful for creation of learning resources and lexicons that are important for extraction of NEs from natural language unstructured text.