A wavelet-based voice activity detection algorithm in noisy environments

Shi-Huang Chen Shi-Huang Chen,Jhing-Fa Wang Jhing-Fa Wang

doi:10.1109/icecs.2002.1046417

Shi-Huang Chen Shi-Huang Chen, Jhing-Fa Wang Jhing-Fa Wang

https://doi.org/10.1109/icecs.2002.1046417

Copy DOI

Export

Save

Cite

Publication Date: Dec 10, 2002

Citations: 17

Affiliation: National Cheng Kung University

Abstract
Full-Text
Similar Papers

Abstract

Listen

This paper presents a new voice activity detection (VAD) algorithm based on the perceptual wavelet packet transform (PWPT) and the Teager energy operator (TEO). The basic procedure of the proposed VAD algorithm is to make use of the PWPT to decompose the input speech into critical subband signals. Then a parameter called voice activity shape (VAS) can be derived from the TEO of these critical subband signals. It is shown in this paper that the VAS can be used as a robust feature for VAD. The advantage of this new algorithm is that the preset threshold values or a priori knowledge of the SNR usually needed in conventional VAD methods can be completely avoided. Various experimental results show that the proposed VAD algorithm is capable of outperforming to the ITU-T G.729B VAD and can operate reliably in real noisy environments.

Full Text