International Symposium on Chinese Spoken Language Processing (ISCSLP 2002)

Taipei, Taiwan
August 23-24, 2002

A Voice Activity Detection Algorithm Based on Perceptual Wavelet Packet Transform and Teager Energy Operator

Jhing-Fa Wang, Shi-Huang Chen

National Cheng Kung University, Tainan, Taiwan

This paper presents a new voice activity detection (VAD) algorithm based on the perceptual wavelet packet transform (PWPT) and the Teager energy operator (TEO). The basic procedure of the proposed VAD algorithm is to make use of the PWPT to decompose the input speech into critical subband signals. Then a parameter called voice activity shape (VAS) can be derived from the TEO of these critical subband signals. It is shown in this paper that the VAS can be used as a robust feature for VAD. The advantage of this new algorithm is that the preset threshold values or a priori knowledge of the SNR usually needed in conventional VAD methods can be completely avoided. Various experimental results show that the proposed VAD algorithm is capable of outperforming to the ITU-T G.729B VAD and can operate reliably in real noisy environments.


Full Paper

Bibliographic reference.  Wang, Jhing-Fa / Chen, Shi-Huang (2002): "A voice activity detection algorithm based on perceptual wavelet packet transform and teager energy operator", In ISCSLP 2002, paper 125.