Arabic Handwritten Documents Segmentation into Text-lines and Words using Deep Learning - Inria - Institut national de recherche en sciences et technologies du numérique Accéder directement au contenu
Communication Dans Un Congrès Année : 2019

Arabic Handwritten Documents Segmentation into Text-lines and Words using Deep Learning

Résumé

One of the most important steps in a handwriting recognition system is text-line and word segmentation. But, this step is made difficult by the differences in handwriting styles, problems of skewness, overlapping and touching of text and the fluctuations of text-lines. It is even more difficult for ancient and calligraphic writings, as in Arabic manuscripts, due to the cursive connection in Arabic text, the erroneous position of diacritic marks, the presence of ascending and descending letters, etc. In this work, we propose an effective segmentation of Arabic handwritten text into text-lines and words, using deep learning. For text-line segmentation, we used an RU-net which allows a pixel-wise classification to separate text-lines pixels from the background ones. For word segmentation, we resorted to the text-line transcription, as we have not got a ground truth at word level. A BLSTM-CTC (Bidirectional Long Short Term Memory followed by a Connectionist Temporal Classification) is then used to perform the mapping between the transcription and text-line image, avoiding the need of the input segmentation. A CNN (Convolutional Neural Network) precedes the BLST-CTC to extract the features and to feed the BLSTM with the essential of the text-line image. Tested on the standard KHATT Arabic database, the experimental results confirm a segmentation success rate of no less than 96.7% for text-lines and 80.1% for words.
Fichier principal
Vignette du fichier
LineSeg-ASAR2019.pdf (1.23 Mo) Télécharger le fichier
Origine : Fichiers produits par l'(les) auteur(s)
Loading...

Dates et versions

hal-02460880 , version 1 (30-01-2020)

Identifiants

  • HAL Id : hal-02460880 , version 1

Citer

Chemseddine Neche, Abdel Belaïd, Afef Kacem-Echi. Arabic Handwritten Documents Segmentation into Text-lines and Words using Deep Learning. ASAR, Sep 2019, Sydney, Australia. ⟨hal-02460880⟩
168 Consultations
847 Téléchargements

Partager

Gmail Facebook X LinkedIn More