Please use this identifier to cite or link to this item: http://hdl.handle.net/10397/114613
PIRA download icon_1.1View/Download Full Text
Title: High-level speaker verification via articulatory-feature based sequence kernels and SVM
Authors: Zhang, SX 
Mak, MW 
Issue Date: 2008
Source: Proceedings of the Annual Conference of the International Speech Communication Association, INTERSPEECH, 2008, p. 1393-1396
Abstract: Articulatory-feature based pronunciation models (AFCPMs) are capable of capturing the pronunciation variations among different speakers and are good for high-level speaker recognition. However, the likelihood-ratio scoring method of AFPCMs is based on a decision boundary created by training the target speaker model and universal background model (UBM) separately. Therefore, the method does not fully utilize the discriminative information available in the training data. To fully harness the discriminative information, this paper proposes training a support vector machine (SVM) for computing the verification scores. More precisely, the models of target speakers, individual background speakers, and claimants are converted to AF-supervectors, which form the inputs to an AF-based kernel of the SVM for computing verification scores. Results show that the proposed AF-kernel scoring is complementary to likelihood-ratio scoring, leading to better performance when the two scoring methods are combined. Further performance enhancement was also observed when the AF scores were combined with acoustic scores derived from a GMM-UBM system.
Publisher: International Speech Communication Association
DOI: 10.21437/interspeech.2008-404
Description: Interspeech 2008, Brisbane, Australia, 22-26 September 2008
Rights: Copyright © 2008 ISCA
The following publication Zhang, S.-X., Mak, M.-W. (2008) High-level speaker verification via articulatory-feature based sequence kernels and SVM. Proc. Interspeech 2008, 1393-1396 is available at https://doi.org/10.21437/Interspeech.2008-404.
Appears in Collections:Conference Paper

Files in This Item:
File Description SizeFormat 
zhang08d_interspeech.pdf518.27 kBAdobe PDFView/Open
Open Access Information
Status open access
File Version Version of Record
Access
View full-text via PolyU eLinks SFX Query
Show full item record

Google ScholarTM

Check

Altmetric


Items in DSpace are protected by copyright, with all rights reserved, unless otherwise indicated.