Skip to Main content Skip to Navigation
Conference papers

Corpus generation for voice command in smart home and the effect of speech synthesis on End-to-End SLU

Abstract : Massive amounts of annotated data greatly contributed to the advance of the machine learning field. However such large data sets are often unavailable for novel tasks performed in realistic environments such as smart homes. In this domain, semantically annotated large voice command corpora for Spoken Language Understanding (SLU) are scarce, especially for non-English languages. We present the automatic generation process of a synthetic semantically-annotated corpus of French commands for smart-home to train pipeline and End-to-End (E2E) SLU models. SLU is typically performed through Automatic Speech Recognition (ASR) and Natural Language Understanding (NLU) in a pipeline. Since errors at the ASR stage reduce the NLU performance, an alternative approach is End-to-End (E2E) SLU to jointly perform ASR and NLU. To that end, the artificial corpus was fed to a text-to-speech (TTS) system to generate synthetic speech data. All models were evaluated on voice commands acquired in a real smart home. We show that artificial data can be combined with real data within the same training set or used as a stand-alone training corpus. The synthetic speech quality was assessed by comparing it to real data using dynamic time warping (DTW).
Complete list of metadatas

Cited literature [52 references]  Display  Hide  Download

https://hal.archives-ouvertes.fr/hal-02861770
Contributor : Michel Vacher <>
Submitted on : Tuesday, June 9, 2020 - 10:59:04 AM
Last modification on : Saturday, July 11, 2020 - 3:00:35 AM

File

2020_LREC_Desot_4_proceedings....
Files produced by the author(s)

Identifiers

  • HAL Id : hal-02861770, version 1

Collections

Citation

Thierry Desot, François Portet, Michel Vacher. Corpus generation for voice command in smart home and the effect of speech synthesis on End-to-End SLU. 12th Conference on Language Resources and Evaluation (LREC 2020), ELRA, May 2020, Marseille, France. pp.6395-6404. ⟨hal-02861770⟩

Share

Metrics

Record views

11

Files downloads

8