Detection of Neurogenic Voice Disorders Using the Fisher Vector Representation of Cepstral Features

Madhu Keerthana Yagnavajjula,Paavo Alku,Krothapalli Sreenivasa Rao,Pabitra Mitra

doi:10.1016/j.jvoice.2022.10.016

Abstract

Neurogenic voice disorders (NVDs) are caused by damage or malfunction of the central or peripheral nervous system that controls vocal fold movement. In this paper, we investigate the potential of the Fisher vector (FV) encoding in automatic detection of people with NVDs. FVs are used to convert features from frame level (local descriptors) to utterance level (global descriptors). At the frame level, we extract two popular cepstral representations, namely, Mel-frequency cepstral coefficients (MFCCs) and perceptual linear prediction cepstral coefficients (PLPCCs), from acoustic voice signals. In addition, the MFCC features are also extracted from every frame of the glottal source signal computed using a glottal inverse filtering (GIF) technique. The global descriptors derived from the local descriptors are used to train a support vector machine (SVM) classifier. Experiments are conducted using voice signals from 80 healthy speakers and 80 patients with NVDs (40 with spasmodic dysphonia (SD) and 40 with recurrent laryngeal nerve palsy (RLNP)) taken from the Saarbruecken voice disorder (SVD) database. The overall results indicate that the use of the FV encoding leads to better identification of people with NVDs, compared to the defacto temporal encoding. Furthermore, the SVM trained using the combination of FVs derived from the cepstral and glottal features provides the overall best detection performance.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: Journal of Voice	Publication Date: Nov 1, 2022
Citations: 2	License type: cc-by

R Discovery Prime

R Discovery Prime

Detection of Neurogenic Voice Disorders Using the Fisher Vector Representation of Cepstral Features

Abstract

Talk to us

Similar Papers

More From: Journal of Voice

Lead the way for us

Similar Papers

The automatic detection of heart failure using speech signals
M Kiran Reddy ... Paavo Alku
Computer Speech & Language | VOL. 69
M Kiran Reddy, et. al.M Kiran Reddy ... Paavo Alku
27 Feb 2021
Computer Speech & Language | VOL. 69

Classification of functional dysphonia using the tunable Q wavelet transform
Kiran Reddy Mittapalle ... Paavo Alku
Speech Communication | VOL. 155
Kiran Reddy Mittapalle, et. al.Kiran Reddy Mittapalle ... Paavo Alku
06 Oct 2023
Speech Communication | VOL. 155

Detection of Specific Language Impairment in Children Using Glottal Source Features
Mittapalle Kiran Reddy ... Krothapalli Sreenivasa Rao
IEEE Access | VOL. 8
Mittapalle Kiran Reddy, et. al.Mittapalle Kiran Reddy ... Krothapalli Sreenivasa Rao
01 Jan 2020
IEEE Access | VOL. 8

Whispered Speech Conversion Based on the Inversion of Mel Frequency Cepstral Coefficient Features
Qiang Zhu ... Yunfeng Dou
Algorithms | VOL. 15
Qiang Zhu, et. al.Qiang Zhu ... Yunfeng Dou
20 Feb 2022
Algorithms | VOL. 15

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Detection of Neurogenic Voice Disorders Using the Fisher Vector Representation of Cepstral Features

Abstract

Talk to us

Similar Papers

More From: Journal of Voice