Merge-Weighted Dynamic Time Warping for Speech Recognition

Xiang-Lilan Zhang,Zhi-Gang Luo,Ming Li

doi:10.1007/s11390-014-1491-0

Abstract

Obtaining training material for rarely used English words and common given names from countries where English is not spoken is difficult due to excessive time, storage and cost factors. By considering personal privacy, languageindependent (LI) with lightweight speaker-dependent (SD) automatic speech recognition (ASR) is a convenient option to solve the problem. The dynamic time warping (DTW) algorithm is the state-of-the-art algorithm for small-footprint SD ASR for real-time applications with limited storage and small vocabularies. These applications include voice dialing on mobile devices, menu-driven recognition, and voice control on vehicles and robotics. However, traditional DTW has several limitations, such as high computational complexity, constraint induced coarse approximation, and inaccuracy problems. In this paper, we introduce the merge-weighted dynamic time warping (MWDTW) algorithm. This method defines a template confidence index for measuring the similarity between merged training data and testing data, while following the core DTW process. MWDTW is simple, efficient, and easy to implement. With extensive experiments on three representative SD speech recognition datasets, we demonstrate that our method outperforms DTW, DTW on merged speech data, the hidden Markov model (HMM) significantly, and is also six times faster than DTW overall.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Merge-Weighted Dynamic Time Warping for Speech Recognition

Abstract

Talk to us

Similar Papers

More From: Journal of Computer Science and Technology

Lead the way for us

Journal: Journal of Computer Science and Technology	Publication Date: Nov 1, 2014
Citations: 15

Similar Papers

One-against-All Weighted Dynamic Time Warping for Language-Independent and Speaker-Dependent Speech Recognition in Adverse Conditions
Xianglilan Zhang ... Zhigang Luo
PLoS ONE | VOL. 9
Xianglilan Zhang, et. al.Xianglilan Zhang ... Zhigang Luo
10 Feb 2014
PLoS ONE | VOL. 9

A Novel Weighted Dynamic Time Warping for Light Weight Speaker-Dependent Speech Recognition in Noisy and Bad Recording Conditions
Xiang Lilan Zhang ... Ji Ping Sun
Applied Mechanics and Materials | VOL. 490-491
Xiang Lilan Zhang, et. al.Xiang Lilan Zhang ... Ji Ping Sun
01 Jan 2014
Applied Mechanics and Materials | VOL. 490-491

Classification of Flying Insects with high performance using improved DTW algorithm based on hidden Markov model
S Arif Abdul Rahuman ... J Veerappan
Brazilian Archives of Biology and Technology | VOL. 59
S Arif Abdul Rahuman, et. al.S Arif Abdul Rahuman ... J Veerappan
01 Jan 2015
Brazilian Archives of Biology and Technology | VOL. 59

Downsampling of time-series data for approximated dynamic time warping on nonvolatile memories
Xingni Li ... Po-Chun Huang
-
Xingni Li, et. al.Xingni Li ... Po-Chun Huang
01 Aug 2017
01 Aug 2017

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Merge-Weighted Dynamic Time Warping for Speech Recognition

Abstract

Talk to us

Similar Papers

More From: Journal of Computer Science and Technology