Finding phoneme trajectories in a feature space of sound and midsagittal ultrasound tongue images

Yuichi Yaguchi,Naoya Horiguchi,Ian Wilson

doi:10.1109/icawst.2012.6469606

Abstract

Supporting the development of a pronunciation learning system, this paper reports an inspection of the trajectory of speech sentences in a feature space that is constructed from midsagittal tongue images and frame-wise speech sounds. One objective of this research is to estimate tongue shape and position from speech sounds, so we focus on determining how best to construct and interpret a feature space we call MUTIS (midsagittal ultrasound tongue image space). Experimental results indicate that higher dimensions of MUTIS are most effective for separating people, and that primarily the lower dimensions of VSS (vocal sound space) data are most effective for separating phonemes. Also, the trajectories within only the VSS data indicate clear differences between first language and second language speakers, but they do not do so within only the MUTIS data. These results indicate that the ultrasound tongue image expresses individual oral cavity over a wide area, and specific tongue shape has a lower contribution in ultrasound tongue images.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Finding phoneme trajectories in a feature space of sound and midsagittal ultrasound tongue images

Abstract

Talk to us

Similar Papers

Lead the way for us

Similar Papers

Comparing L1 and L2 phoneme trajectories in a feature space of sound and midsagittal ultrasound tongue images
Keita Sano ... Yuichi Yaguchi
The Journal of the Acoustical Society of America | VOL. 132
Keita Sano, et. al.Keita Sano ... Yuichi Yaguchi
01 Sep 2012
The Journal of the Acoustical Society of America | VOL. 132

Ultra2Speech - A Deep Learning Framework for Formant Frequency Estimation and Tracking from Ultrasound Tongue Images
Pramit Saha ... Bryan Gick
-
Pramit Saha, et. al.Pramit Saha ... Bryan Gick
01 Jan 2020
01 Jan 2020

DNN-based Acoustic-to-Articulatory Inversion using Ultrasound Tongue Imaging
Dagoberto Porras ... Alexander Sepulveda-Sepulveda
-
Dagoberto Porras, et. al.Dagoberto Porras ... Alexander Sepulveda-Sepulveda
01 Jul 2019
01 Jul 2019

Articulatory feature extraction from ultrasound images using pretrained convolutional neural networks
Kele Xu ... Jian Zhu
The Journal of the Acoustical Society of America | VOL. 144
Kele Xu, et. al.Kele Xu ... Jian Zhu
01 Sep 2018
The Journal of the Acoustical Society of America | VOL. 144

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Finding phoneme trajectories in a feature space of sound and midsagittal ultrasound tongue images

Abstract

Talk to us

Similar Papers