Towards Speaker and Environmental Robustness in ASR: The HIWIRE Project

Abstract : In this paper, we present algorithms for dealing with variability and mismatch in speech recognition due to environmental conditions and non-native speaker populations. The proposed algorithms cover a broad spectrum of ideas including robust feature extraction, feature compensation and speech enhancement. Specifically the following algorithms are presented and evaluated: beamforming for multi-microphone speech recognition, robust modulation and fractal features, Teager energy cepstrum coefficients, parametric feature equalization, speech enhancement, and acoustic modeling for non-native speech recognition. Also the problem of feature fusion and voice activity detection are discussed. Evaluation results on the AURORA databases under the auspices of the HIWIRE project show that significant gains can be achieved under adverse or mismatched conditions using these algorithms. Relative error rate reduction of up to 50% was shown for multi-microphone speech recognition, robust feature combination and speech enhancement. 30-40% reduction was shown for parametric feature equalization and non-native acoustic models.
Type de document :
Communication dans un congrès
SRIV'06 ITRW on Speech Recognition and Intrinsic Variation, May 2006, Toulouse, France. 2006
Liste complète des métadonnées

Littérature citée [28 références]  Voir  Masquer  Télécharger

https://hal.archives-ouvertes.fr/hal-00110502
Contributeur : Dominique Fohr <>
Soumis le : lundi 30 octobre 2006 - 12:34:20
Dernière modification le : samedi 1 décembre 2018 - 19:56:17
Document(s) archivé(s) le : mardi 6 avril 2010 - 21:16:16

Identifiants

  • HAL Id : hal-00110502, version 1

Collections

Citation

A. Potamianos, Ghazi Bouselmi, D. Dimitriadis, Dominique Fohr, R. Gemello, et al.. Towards Speaker and Environmental Robustness in ASR: The HIWIRE Project. SRIV'06 ITRW on Speech Recognition and Intrinsic Variation, May 2006, Toulouse, France. 2006. 〈hal-00110502〉

Partager

Métriques

Consultations de la notice

1192

Téléchargements de fichiers

194