International Journal of Computer Applications
Foundation of Computer Science (FCS), NY, USA
|
Volume 134 - Issue 15 |
Published: January 2016 |
Authors: S. B. Dhonde, S. M. Jagade |
![]() |
S. B. Dhonde, S. M. Jagade . Comparison of Vector Quantization and Gaussian Mixture Model using Effective MFCC Features for Text-independent Speaker Identification. International Journal of Computer Applications. 134, 15 (January 2016), 11-13. DOI=10.5120/ijca2016908019
@article{ 10.5120/ijca2016908019, author = { S. B. Dhonde,S. M. Jagade }, title = { Comparison of Vector Quantization and Gaussian Mixture Model using Effective MFCC Features for Text-independent Speaker Identification }, journal = { International Journal of Computer Applications }, year = { 2016 }, volume = { 134 }, number = { 15 }, pages = { 11-13 }, doi = { 10.5120/ijca2016908019 }, publisher = { Foundation of Computer Science (FCS), NY, USA } }
%0 Journal Article %D 2016 %A S. B. Dhonde %A S. M. Jagade %T Comparison of Vector Quantization and Gaussian Mixture Model using Effective MFCC Features for Text-independent Speaker Identification%T %J International Journal of Computer Applications %V 134 %N 15 %P 11-13 %R 10.5120/ijca2016908019 %I Foundation of Computer Science (FCS), NY, USA
In this paper, the performance of speaker modeling schemes such as vector quantization (VQ) and Gaussian mixture model (GMM) is compared for speaker identification. Along with the effective size of feature set, model based approaches are typically used as a solution for robustness issues of speaker recognition systems. Gaussian Mixture Model (GMM) is versatile parameter estimation approach whereas; Vector Quantization (VQ) is based on template modeling. Here, first, MFCC features are used to extract speaker specific speech features for text-independent speaker identification. MFCC features are then modeled using Vector Quantization (VQ) and Gaussian mixture model (GMM) and their performance is compared in the context of speaker identification. The average recognition rate achieved for MFCC with GMM is 99.2% and for MFCC with VQ is 98.4% on TIMIT database consisting of 64 speakers.