Robust and complex approach of pathological speech signal analysis

Jiri Mekyska,Eva Janousova,Pedro Gomez-Vilda,Zdenek Smekal,Irena Rektorova,Ilona Eliasova,Milena Kostalova,Martina Mrackova,Jesus B Alonso-Hernandez,Marcos Faundez-Zanuy,Karmele López-De-Ipiña

doi:10.1016/j.neucom.2015.02.085

Abstract

This paper presents a study of the approaches in the state-of-the-art in the field of pathological speech signal analysis with a special focus on parametrization techniques. It provides a description of 92 speech features where some of them are already widely used in this field of science and some of them have not been tried yet (they come from different areas of speech signal processing like speech recognition or coding). As an original contribution, this work introduces 36 completely new pathological voice measures based on modulation spectra, inferior colliculus coefficients, bicepstrum, sample and approximate entropy and empirical mode decomposition. The significance of these features was tested on 3 (English, Spanish and Czech) pathological voice databases with respect to classification accuracy, sensitivity and specificity. To our best knowledge the introduced approach based on complex feature extraction and robust testing outperformed all works that have been published already in this field. The results (accuracy, sensitivity and specificity equal to 100.0±0.0%) are discussable in the case of Massachusetts Eye and Ear Infirmary (MEEI) database because of its limitation related to a length of sustained vowels, however in the case of Príncipe de Asturias (PdA) Hospital in Alcalá de Henares of Madrid database we made improvements in classification accuracy (82.1±3.3%) and specificity (83.8±5.1%) when considering a single-classifier approach. Hopefully, large improvements may be achieved in the case of Czech Parkinsonian Speech Database (PARCZ), which are discussed in this work as well. All the features introduced in this work were identified by Mann–Whitney U test as significant (p<0.05) when processing at least one of the mentioned databases. The largest discriminative power from these proposed features has a cepstral peak prominence extracted from the first intrinsic mode function (p=6.9443×10−32) which means, that among all newly designed features those that quantify especially hoarseness or breathiness are good candidates for pathological speech identification. The paper also mentions some ideas for the future work in the field of pathological speech signal analysis that can be valuable especially under the clinical point of view.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Robust and complex approach of pathological speech signal analysis

Abstract

Talk to us

Similar Papers

More From: Neurocomputing

Lead the way for us

Journal: Neurocomputing	Publication Date: May 21, 2015
Citations: 98

Similar Papers

Pathological speech signal analysis and classification using empirical mode decomposition
Muhammad Kaleem ... Behnaz Ghoraani
Medical & Biological Engineering & Computing | VOL. 51
Muhammad Kaleem, et. al.Muhammad Kaleem ... Behnaz Ghoraani
05 Mar 2013
Medical & Biological Engineering & Computing | VOL. 51

Feature analysis for automatic detection of pathological speech
A.A Dibazar ... S Narayanan
-
A.A Dibazar, et. al.A.A Dibazar ... S Narayanan
23 Oct 2002
23 Oct 2002

Pathological voice detection based on gammatone short time spectral self-similarity
Denghuang Zhao ... Zhi Tao
Sheng wu yi xue gong cheng xue za zhi = Journal of biomedical engineering = Shengwu yixue gongchengxue zazhi | VOL. 39
Denghuang Zhao, et. al.Denghuang Zhao ... Zhi Tao
25 Aug 2022
Sheng wu yi xue gong cheng xue za zhi = Journal of biomedical engineering = Shengwu yixue gongchengxue zazhi | VOL. 39

Pathological Voice Detection Based on Phase Reconstitution and Convolutional Neural Network
Deli Fu ... Weiping Hu
Journal of Voice | VOL. -
Deli Fu, et. al.Deli Fu ... Weiping Hu
01 Oct 2022
Journal of Voice | VOL. -

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Robust and complex approach of pathological speech signal analysis

Abstract

Talk to us

Similar Papers

More From: Neurocomputing