Construction of language models for an handwritten mail reading system - Archive ouverte HAL Accéder directement au contenu
Communication Dans Un Congrès Année : 2012

Construction of language models for an handwritten mail reading system

Résumé

This paper presents a system for the recognition of unconstrained handwritten mails. The main part of this system is an HMM recognizer which uses trigraphs to model contextual information. This recognition system does not require any segmentation into words or characters and directly works at line level. To take into account linguistic information and enhance performance, a language model is introduced. This language model is based on bigrams and built from training document transcriptions only. Different experiments with various vocabulary sizes and language models have been conducted. Word Error Rate and Perplexity values are compared to show the interest of specific language models, fit to handwritten mail recognition task.
Fichier principal
Vignette du fichier
8297_27.PDF (450.32 Ko) Télécharger le fichier
Origine : Fichiers produits par l'(les) auteur(s)
Loading...

Dates et versions

hal-00737407 , version 1 (01-10-2012)

Identifiants

Citer

Olivier Morillot, Laurence Likforman-Sulem, Emmanuèle Grosicki. Construction of language models for an handwritten mail reading system. IS&T/SPIE 24th Annual Symposium on Electronic Imaging - Document Recognition and Retrieval XIX, Jan 2012, San Francisco, United States. ⟨10.1117/12.911965⟩. ⟨hal-00737407⟩
78 Consultations
232 Téléchargements

Altmetric

Partager

Gmail Facebook X LinkedIn More