Skip to Main content Skip to Navigation
Conference papers

Word confidence estimation for speech translation

Abstract : Word Confidence Estimation (WCE) for machine transla-tion (MT) or automatic speech recognition (ASR) consists in judging each word in the (MT or ASR) hypothesis as correct or incorrect by tagging it with an appropriate label. In the past, this task has been treated separately in ASR or MT con-texts and we propose here a joint estimation of word confi-dence for a spoken language translation (SLT) task involving both ASR and MT. This research work is possible because we built a specific corpus which is first presented. This cor-pus contains 2643 speech utterances for which a quintuplet containing: ASR output (src-asr), verbatim transcript (src-ref), text translation output (tgt-mt), speech translation out-put (tgt-slt) and post-edition of translation (tgt-pe), is made available. The rest of the paper illustrates how such a corpus (made available to the research community) can be used for evaluating word confidence estimators in ASR, MT or SLT scenarios. WCE for SLT could help rescoring SLT output graphs, improving translators productivity (for translation of lectures or movie subtitling) or it could be useful in interac-tive speech-to-speech translation scenarios. Word confidence estimation (WCE), Spoken Language Translation (SLT), Corpus, Joint features.
Document type :
Conference papers
Complete list of metadata

Cited literature [23 references]  Display  Hide  Download
Contributor : Laurent Besacier Connect in order to contact the contributor
Submitted on : Wednesday, January 28, 2015 - 2:30:02 PM
Last modification on : Thursday, October 21, 2021 - 3:53:27 AM
Long-term archiving on: : Wednesday, April 29, 2015 - 10:30:48 AM


Files produced by the author(s)


  • HAL Id : hal-01110393, version 1


Laurent Besacier, Benjamin Lecouteux, Ngoc-Quang Luong, K Hour, Marwa Hadj Salah. Word confidence estimation for speech translation. International Workshop on Spoken Language Translation, Dec 2014, Lake Tahoe, United States. ⟨hal-01110393⟩



Record views


Files downloads