Arabic Handwritten Documents Segmentation into Text-lines and Words using Deep Learning

Chemseddine Neche; Abdel Belaïd; Afef Kacem-Echi

Communication Dans Un Congrès Année : 2019

Arabic Handwritten Documents Segmentation into Text-lines and Words using Deep Learning

(1) , (1) , (2)

1
2

Chemseddine Neche

Fonction : Auteur
PersonId : 1064002

Recognition of writing and analysis of documents

Abdel Belaïd

Fonction : Auteur
PersonId : 830137

Recognition of writing and analysis of documents

Afef Kacem-Echi

Fonction : Auteur

Laboratoire de Recherche en Technologies de l’Information et de la Communication & Génie Electrique [Tunis]

Résumé

One of the most important steps in a handwriting recognition system is text-line and word segmentation. But, this step is made difficult by the differences in handwriting styles, problems of skewness, overlapping and touching of text and the fluctuations of text-lines. It is even more difficult for ancient and calligraphic writings, as in Arabic manuscripts, due to the cursive connection in Arabic text, the erroneous position of diacritic marks, the presence of ascending and descending letters, etc. In this work, we propose an effective segmentation of Arabic handwritten text into text-lines and words, using deep learning. For text-line segmentation, we used an RU-net which allows a pixel-wise classification to separate text-lines pixels from the background ones. For word segmentation, we resorted to the text-line transcription, as we have not got a ground truth at word level. A BLSTM-CTC (Bidirectional Long Short Term Memory followed by a Connectionist Temporal Classification) is then used to perform the mapping between the transcription and text-line image, avoiding the need of the input segmentation. A CNN (Convolutional Neural Network) precedes the BLST-CTC to extract the features and to feed the BLSTM with the essential of the text-line image. Tested on the standard KHATT Arabic database, the experimental results confirm a segmentation success rate of no less than 96.7% for text-lines and 80.1% for words.

Domaines

Informatique [cs] Intelligence artificielle [cs.AI]

Fichier principal

LineSeg-ASAR2019.pdf (1.23 Mo)

Origine : Fichiers produits par l'(les) auteur(s)

Abdel Belaid : Connectez-vous pour contacter le contributeur

https://inria.hal.science/hal-02460880

Soumis le : jeudi 30 janvier 2020-12:32:44

Dernière modification le : lundi 11 septembre 2023-17:41:19

Dates et versions

hal-02460880 , version 1 (30-01-2020)

Identifiants

HAL Id : hal-02460880 , version 1

Citer

Chemseddine Neche, Abdel Belaïd, Afef Kacem-Echi. Arabic Handwritten Documents Segmentation into Text-lines and Words using Deep Learning. ASAR, Sep 2019, Sydney, Australia. ⟨hal-02460880⟩

Exporter

BibTeX XML-TEI Dublin Core DC Terms EndNote DataCite

Collections

CNRS INRIA UNIV-LORRAINE LORIA LORIA-NLPKD

168 Consultations

847 Téléchargements

Arabic Handwritten Documents Segmentation into Text-lines and Words using Deep Learning

Résumé

Domaines

Dates et versions

Identifiants

Citer

Exporter

Collections

Partager