2009Unpublished venueOpen access

Speaker Verification Based on Different Vector Quantization Techniques with Gaussian Mixture Models

Sheeraz Memon, Margaret Lech, Namunu C. Maddage

Open full text 20 citations

Abstract

The introduction of Gaussian mixture models (GMMs) in the field of speaker verification has led to very good results. This paper illustrates an evolution in state-of-the-art speaker verification by highlighting the contribution of recently established information theoretic based vector quantization technique. We explore the novel application of three different vector quantization algorithms, namely K-means, Linde-Buzo-Gray (LBG) and information theoretic vector quantization (ITVQ) for efficient speaker verification. The expectation maximization (EM) algorithm used by GMM requires a prohibitive amount of iterations to converge. In this paper, comparable alternatives to EM including K-means, LBG and ITVQ algorithm were tested. The GMM-ITVQ algorithm was found to be the most efficient alternative for the GMM-EM. It gives correct classification rates at a similar level to that of GMM-EM. Finally, representative performance benchmarks and system behaviour experiments on NIST SRE corpora are presented.

About this research paper

What this paper is about

The introduction of Gaussian mixture models (GMMs) in the field of speaker verification has led to very good results. This paper illustrates an evolution in state-of-the-art speaker verification by highlighting the contribution of recently established information theoretic based vector quantization technique. We explore the novel application of three different vector quantization algorithms, namely K-means, Linde-Buzo-Gray (LBG) and information theoretic vector quantization (ITVQ) for efficient speaker verification. The expectation maximization (EM) algorithm used by GMM requires a prohibitive amount of iterations to converge. In this paper, comparable alternatives to EM including K-means, LBG and ITVQ algorithm were tested. The GMM-ITVQ algorithm was found to be the most efficient alternative for the GMM-EM. It gives correct classification rates at a similar level to that of GMM-EM. Finally, representative performance benchmarks and system behaviour experiments on NIST SRE corpora are presented.

Why it matters

OpenAlex reports 20 citations for this work. Citation counts describe recorded attention and do not establish research quality.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

The introduction of Gaussian mixture models (GMMs) in the field of speaker verification has led to very good results. This paper illustrates an evolution in state-of-the-art speaker verification by highlighting the contribution of recently established information theoretic based vector quantization technique. We explore the novel application of three different vector quantization algorithms, namely K-means, Linde-Buzo-Gray (LBG) and information theoretic vector quantization (ITVQ) for efficient speaker verification. The expectation maximization (EM) algorithm used by GMM requires a prohibitive amount of iterations to converge. In this paper, comparable alternatives to EM including K-means, LBG and ITVQ algorithm were tested. The GMM-ITVQ algorithm was found to be the most efficient alternative for the GMM-EM. It gives correct classification rates at a similar level to that of GMM-EM. Finally, representative performance benchmarks and system behaviour experiments on NIST SRE corpora are presented.

Key concepts: Vector quantization, Mixture model, Speaker verification, Computer science, Quantization (signal processing), NIST, Gaussian, Expectation–maximization algorithm

Related papers

Back to paper searchBrowse research topicsOriginal source
Speaker Verification Based on Different Vector Quantization Techniques with Gaussian Mixture Models — Research Paper | ScholarLens