Degraded Script Identification of Urdu and Devanagari Document-A Survey
Sobia Habib, Manoj Kumar Shukla, Rajiv Kapoor · 2019
Script identification especially for non-Latin script have gained the attention of researchers from both the academics and industry. There are lots of challenges associated with this since most of the existing research focuses on Latin scripts. Most of the researches in this field are working only with the latest or modern documents and font types. Historical and Degraded documents have not been given much importance in the OCR research. This paper provides the different stages required for the identification of the scripts. A brief overview of the different techniques for identifying and classifying the characters in Devanagari and Urdu Script. This has been performed especially for degraded and historical texts. The paper has been concluded with a strong future scope.