Word confidence estimation for speech translation - Archive ouverte HAL Accéder directement au contenu
Communication Dans Un Congrès Année : 2014

Word confidence estimation for speech translation

Résumé

Word Confidence Estimation (WCE) for machine transla-tion (MT) or automatic speech recognition (ASR) consists in judging each word in the (MT or ASR) hypothesis as correct or incorrect by tagging it with an appropriate label. In the past, this task has been treated separately in ASR or MT con-texts and we propose here a joint estimation of word confi-dence for a spoken language translation (SLT) task involving both ASR and MT. This research work is possible because we built a specific corpus which is first presented. This cor-pus contains 2643 speech utterances for which a quintuplet containing: ASR output (src-asr), verbatim transcript (src-ref), text translation output (tgt-mt), speech translation out-put (tgt-slt) and post-edition of translation (tgt-pe), is made available. The rest of the paper illustrates how such a corpus (made available to the research community) can be used for evaluating word confidence estimators in ASR, MT or SLT scenarios. WCE for SLT could help rescoring SLT output graphs, improving translators productivity (for translation of lectures or movie subtitling) or it could be useful in interac-tive speech-to-speech translation scenarios. Word confidence estimation (WCE), Spoken Language Translation (SLT), Corpus, Joint features.
Fichier principal
Vignette du fichier
iwslt2014-FINAL.pdf (194.35 Ko) Télécharger le fichier
Origine : Fichiers produits par l'(les) auteur(s)
Loading...

Dates et versions

hal-01110393 , version 1 (28-01-2015)

Identifiants

  • HAL Id : hal-01110393 , version 1

Citer

Laurent Besacier, Benjamin Lecouteux, Ngoc-Quang Luong, K Hour, Marwa Hadj Salah. Word confidence estimation for speech translation. International Workshop on Spoken Language Translation, Dec 2014, Lake Tahoe, United States. ⟨hal-01110393⟩
229 Consultations
570 Téléchargements

Partager

Gmail Facebook X LinkedIn More