Improving transliteration accuracy using word-origin detection and lexicon lookup
Mitesh M. Khapra, Pushpak Bhattacharyya · 2009
We propose a framework for transliteration which uses (i) a word-origin detection engine (pre-processing) (ii) a CRF based transliteration engine and (iii) a re-ranking model based on lexicon-lookup (post-processing). The results obtained for English-Hindi and English-Kannada transliteration show that the preprocessing and post-processing modules improve the top-1 accuracy by 7.1%.