Sanskrit-Gujarati Constituency Mapper for Machine Translation System

Jaideepsinh K. Raulji, Jatinderkumar R. Saini · 2019

Looking at vastness, depth and precise nature of Sanskrit grammar and geographically wide proliferation of Gujarati language and its native speaker, it becomes necessary to spotlight on constituency characteristics and features of Sanskrit and Gujarati. Both the languages fall under Indo-Iranian language sub-tree, but there are grammatical divergences which are discussed here so as to reflect in implementation of Machine Translation System (MTS). The content revolves around divergence pattern for a rule base MT system, due to scarce or unavailability of parallel aligned corpora to incorporate statistical or Example based methodology. The Sanskrit grammatical constituents like indeclinables, pronouns, verbs and nouns are analyzed. The Sanskrit inflectional affixes are mapped to its Gujarati inflectional affixes for each equivalent grammar constituent.

Read the paper · More papers on PaperTik