Identity Vector Extraction by Perceptual Wavelet Packet Entropy and Convolutional Neural Network for Voice Authentication-Reference-Cited by-同舟云学术

Identity Vector Extraction by Perceptual Wavelet Packet Entropy and Convolutional Neural Network for Voice Authentication

Published:2018-08-13 Issue:8 Volume:20 Page:600
ISSN:1099-4300
Container-title:Entropy
language:en
Short-container-title:Entropy

Author:

Lei Lei^ORCID,She Kun

Abstract

Recently, the accuracy of voice authentication system has increased significantly due to the successful application of the identity vector (i-vector) model. This paper proposes a new method for i-vector extraction. In the method, a perceptual wavelet packet transform (PWPT) is designed to convert speech utterances into wavelet entropy feature vectors, and a Convolutional Neural Network (CNN) is designed to estimate the frame posteriors of the wavelet entropy feature vectors. In the end, i-vector is extracted based on those frame posteriors. TIMIT and VoxCeleb speech corpus are used for experiments and the experimental results show that the proposed method can extract appropriate i-vector which reduces the equal error rate (EER) and improve the accuracy of voice authentication system in clean and noisy environment.

Publisher

MDPI AG

Subject

General Physics and Astronomy

Link

http://www.mdpi.com/1099-4300/20/8/600/pdf

Reference30 articles.

1. A Study of Interspeaker Variability in Speaker Verification

2. Joint Speaker Verification and Antispoofing in the $i$ -Vector Space

3. Average framing linear prediction coding with wavelet transform for text-independent speaker identification system

4. Wavelet packet based Mel frequency cepstral coefficient features for text independent speaker identification;Srivastava;Intell. Inf.,2013

Cited by 7 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Voiceprint Recognition under Cross-Scenario Conditions Using Perceptual Wavelet Packet Entropy-Guided Efficient-Channel-Attention–Res2Net–Time-Delay-Neural-Network Model;Mathematics;2023-10-09

2. Accurate and early prediction of the wound healing outcome of burn injuries using the wavelet Shannon entropy of terahertz time-domain waveforms;Journal of Biomedical Optics;2022-11-08

3. Research on Anti-Frequency Sweeping Jamming Method for Frequency Modulation Continuous Wave Radio Fuze Based on Wavelet Packet Transform Features;Applied Sciences;2022-08-30

4. Machine Learning Techniques for THz Imaging and Time-Domain Spectroscopy;Sensors;2021-02-08

5. Pattern analysis based acoustic signal processing: a survey of the state-of-art;International Journal of Speech Technology;2020-02-03