Simultaneous Estimation of Glottal Source Waveforms and Vocal Tract Shapes from Speech Signals Based on ARX-LF Model

Yongwei Li,Masato Akagi,Ken-Ichi Sakakibara

doi:10.1007/s11265-019-01510-4

Abstract

Estimating glottal source waveforms and vocal tract shapes is typically done by processing the speech signal using an inverse filter and then fitting the residual signal using the glottal source model. However, due to source-tract interactions, the estimation accuracy is reduced. In this paper, we propose a method to estimate glottal source waveforms and vocal tract shapes simultaneously based on an analysis-by-synthesis approach with a source-filter model constructed of an Auto-Regressive eXogenous (ARX) model and the Liljencrants-Fant (LF) model. Since the optimization of multiple parameters makes simultaneous estimation difficult, we first initialize the glottal source parameters using the inverse filter method, and then simultaneously estimate the accurate parameters of the glottal sources and the vocal tract shapes using an analysis-by-synthesis approach. Experimental results with synthetic and real speech signals showed that the proposed method has higher estimation accuracy than using the inverse filter.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Simultaneous Estimation of Glottal Source Waveforms and Vocal Tract Shapes from Speech Signals Based on ARX-LF Model

Abstract

Talk to us

Similar Papers

More From: Journal of Signal Processing Systems

Lead the way for us

Journal: Journal of Signal Processing Systems	Publication Date: Dec 23, 2019
Citations: 4

Similar Papers

Estimation of glottal source waveforms and vocal tract shapes from speech signals based on ARX-LF model
Yongwei Li ... Ken-Ichi Sakakibara
-
Yongwei Li, et. al.Yongwei Li ... Ken-Ichi Sakakibara
01 Nov 2018
01 Nov 2018

$F_0$-Noise-Robust Glottal Source and Vocal Tract Analysis Based on ARX-LF Model
Yongwei Li ... Donna Erickson
IEEE/ACM Transactions on Audio, Speech, and Language Processing | VOL. 29
Yongwei Li, et. al.Yongwei Li ... Donna Erickson
01 Jan 2020
IEEE/ACM Transactions on Audio, Speech, and Language Processing | VOL. 29

Investigation of self-supervised pre-trained models for classification of voice quality from speech and neck surface accelerometer signals
Sudarsana Reddy Kadiri ... Paavo Alku
Computer Speech & Language | VOL. 83
Sudarsana Reddy Kadiri, et. al.Sudarsana Reddy Kadiri ... Paavo Alku
28 Jul 2023
Computer Speech & Language | VOL. 83

Estimation of glottal source waveforms and vocal tract shape for singing voices with wide frequency range
Kyoko Takahashi ... Masato Akagi
-
Kyoko Takahashi, et. al.Kyoko Takahashi ... Masato Akagi
01 Nov 2018
01 Nov 2018

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Simultaneous Estimation of Glottal Source Waveforms and Vocal Tract Shapes from Speech Signals Based on ARX-LF Model

Abstract

Talk to us

Similar Papers

More From: Journal of Signal Processing Systems