Speaker identification using discrete wavelet packet transform technique with irregular decomposition

作者：

Highlights：

•

摘要

This paper presents the study of speaker identification for security systems based on the energy of speaker utterances. The proposed system consisted of a combination of signal pre-process, feature extraction using wavelet packet transform (WPT) and speaker identification using artificial neural network. In the signal pre-process, the amplitude of utterances, for a same sentence, were normalized for preventing an error estimation caused by speakers’ change in volume. In the feature extraction, three conventional methods were considered in the experiments and compared with the irregular decomposition method in the proposed system. In order to verify the effect of the proposed system for identification, a general regressive neural network (GRNN) was used and compared in the experimental investigation. The experimental results demonstrated the effectiveness of the proposed speaker identification system and were compared with the discrete wavelet transform (DWT), conventional WPT and WPT in Mel scale.

论文关键词：Speaker identification,Discrete wavelet transform,Wavelet packet transform,General regressive neural network

论文评审过程：Available online 9 February 2008.

论文官网地址：https://doi.org/10.1016/j.eswa.2008.01.038