AMRITATCS-IITGUWAHATI combined system for the Speakers in the Wild (SITW) speaker recognition challenge

Kuruvachan K. George, Rohan Kumar Das, Sarfaraz Jelil, K. Arun Das, C. Santhosh Kumar, S. R. Mahadeva Prasanna, Ashish Kumar Panda · 2016

In this work, the details of AMRITA-TCS and IITGUWAHATI speaker recognition systems submitted to the Speakers in the Wild (SITW) speaker recognition challenge are presented. The AMRITA-TCS system is a fusion of i-vector with a backend probabilistic linear discriminant analysis (i-PLDA) system and a cosine distance features (CDF) with backend support vector machine classifier (CDF-SVM) system, developed using the short term cepstral features, mel frequency cepstral coefficients (MFCC) and power normalized cepstral coefficients (PNCC), respectively. The IITGUWAHATI system is an i-PLDA system using MFCC with a vowel like region (VLR) based feature selection (i-PLDA-VLR). The experimental results reported in this work are based on the core-core condition of the challenge. Finally, a fusion of AMRITA-TCS and IITGUWAHATI speaker recognition systems is carried out that enhances the performance than each of the subsystems.

Read the paper · More papers on PaperTik