A complete character recognition and transliteration technique for Devanagari script - Inria - Institut national de recherche en sciences et technologies du numérique Accéder directement au contenu
Pré-Publication, Document De Travail Année : 2020

A complete character recognition and transliteration technique for Devanagari script

Jasmine Kaur
  • Fonction : Auteur
Vinay Kumar

Résumé

Transliteration involves transformation of one script to another based on phonetic similarities between the characters of two distinctive scripts. In this paper, we present a novel technique for automatic transliteration of Devanagari script using character recognition. One of the first tasks performed to isolate the constituent characters is segmentation. Line segmentation methodology in this manuscript discusses the case of overlapping lines. Character segmentation algorithm is designed to segment conjuncts and separate shadow characters. Presented shadow character segmentation scheme employs connected component method to isolate the character, keeping the constituent characters intact. Statistical features namely different order moments like area, variance, skewness and kurtosis along with structural features of characters are employed in two phase recognition process. After recognition, constituent Devanagari characters are mapped to corresponding roman alphabets in way that resulting roman alphabets have similar pronunciation to source characters.

Dates et versions

hal-03020043 , version 1 (23-11-2020)

Identifiants

Citer

Jasmine Kaur, Vinay Kumar. A complete character recognition and transliteration technique for Devanagari script. 2020. ⟨hal-03020043⟩
36 Consultations
0 Téléchargements

Altmetric

Partager

Gmail Facebook X LinkedIn More