Title
Improving Performance of Speaker Identification System Using Complementary Information Fusion
Abstract
Feature extraction plays an important role as a front-end processing block in speaker identification (SI) process. Most of the SI systems utilize like Mel-Frequency Cepstral Coefficients (MFCC), Perceptual Linear Prediction (PLP), Linear Predictive Cepstral Coefficients (LPCC), as a feature for representing speech signal. Their derivations are based on short term processing of speech signal and they try to capture the vocal tract information ignoring the contribution from the vocal cord. Vocal cord cues are equally important in SI context, as the information like pitch frequency, phase in the residual signal, etc could convey important speaker specific attributes and are complementary to the information contained in spectral feature sets. In this paper we propose a novel feature set extracted from the residual signal of LP modeling. Higher-order statistical moments are used here to find the nonlinear relationship in residual signal. To get the advantages of complementarity vocal cord based decision score is fused with the vocal tract based score. The experimental results on two public databases show that fused mode system outperforms single spectral features.
Year
Venue
Keywords
2011
Clinical Orthopaedics and Related Research
mel frequency cepstral coefficient,feature extraction,vocal tract,front end
Field
DocType
Volume
Complementarity (molecular biology),Mel-frequency cepstrum,Residual,Speaker identification,Nonlinear system,Pattern recognition,Computer science,Feature extraction,Speech recognition,Artificial intelligence,Vocal tract,Method of moments (statistics)
Journal
abs/1105.2
ISSN
Citations 
PageRank 
Proceedings of 17th International Conference on Advanced Computing and Communications (ADCOM 2009) pp. 182-187 (2009)
1
0.36
References 
Authors
0
3
Name
Order
Citations
PageRank
Md. Sahidullah132624.99
Sandipan Chakroborty2313.34
Goutam Saha3112.21