Analysis and Classification of Speech Signals by Generalized Fractal Dimension Features - Archive ouverte HAL Accéder directement au contenu
Article Dans Une Revue Speech Communication Année : 2009

Analysis and Classification of Speech Signals by Generalized Fractal Dimension Features

Vassilis Pitsikalis
  • Fonction : Auteur correspondant
  • PersonId : 895250

Connectez-vous pour contacter l'auteur
Petros Maragos
  • Fonction : Auteur

Résumé

We explore nonlinear signal processing methods inspired by dynamical systems and fractal theory in order to analyze and characterize speech sounds.A speech signal is at first embedded in a multidimensional phase-space and further employed for the estimation of measurements related to the fractal dimensions.Our goals are to compute these raw measurementsin the practical cases of speech signals, to further utilize them for the extraction of simple descriptive features and to address issues on the efficacy of the proposed features to characterize speech sounds.We observe that distinct feature vector elements obtain values or show statistical trends that on average depend on general characteristics such as the voicing, the manner and the place of articulation of broad phoneme classes. Moreover the way that the statistical parameters of the features are altered as an effect of the variation of phonetic characteristics seem to follow some roughly formed patterns.We also discuss some qualitative aspects concerning the linear phoneme-wise correlation between the fractal features and the commonly employed mel-frequency cepstral coefficients (MFCC) demonstrating phonetic cases of maximal and minimal correlation. In the same context we also investigate the fractal features' spectral content, in terms of the most and least correlated components with the MFCC.Further the proposed methods are examined under the light of indicative phoneme classification experiments.These quantify the efficacy of the features to characterize broad classes of speech sounds.The results are shown to be comparable for some classification scenarios with the corresponding ones of the MFCC features.
Fichier principal
Vignette du fichier
PEER_stage2_10.1016%2Fj.specom.2009.06.005.pdf (1009.12 Ko) Télécharger le fichier
Origine : Fichiers produits par l'(les) auteur(s)
Loading...

Dates et versions

hal-00575231 , version 1 (10-03-2011)

Identifiants

Citer

Vassilis Pitsikalis, Petros Maragos. Analysis and Classification of Speech Signals by Generalized Fractal Dimension Features. Speech Communication, 2009, 51 (12), pp.1206. ⟨10.1016/j.specom.2009.06.005⟩. ⟨hal-00575231⟩

Collections

PEER
132 Consultations
247 Téléchargements

Altmetric

Partager

Gmail Facebook X LinkedIn More