The BLZ Systems for the 2011 NIST Language Recognition Evaluation
Mikel Peñagarikano, Amparo Varona, Alberto Abad, Jesús Villalba, Alfonso Ortega, Eduardo Lleida · 2011
This paper briefly describes the language recognition systems developed for the 2011 NIST Language Recognition Evaluation (LRE) by the BLZ (Bilbao-Lisboa-Zaragoza) team, a threesite joint including GTTS from the University of the Basque Country (Spain), L 2 F (Spoken Language Systems Lab) from INESC-ID Lisboa (Portugal) and I3A from the University of Zaragoza (Spain). The primary system fuses 8 (3 acoustic + 5 phonotactic) subsystems: a Linearized Eigenchannel GMM (LE-GMM) subsystem, a JFA subsystem, an iVector subsystem, three Phone-SVM subsystems using the Brno University of Technology phone decoders for Czech, Hungarian and Russian, and two Phone-SVM subsystems using the L 2 F phone decoders for European Portuguese and Brazilian Portuguese. Gaussian backends and multiclass fusion have been applied to get the final scores. Three contrastive systems have been also submitted, featuring: (1) the fusion of the whole set of 13 (6 acoustic + 7 phonotactic) subsystems; (2) the fusion of 3 subsystems, for the combination of one subsystem per site yielding the best performance on development data; and (3) the fusion of the same 8 subsystems used in the primary system under a different configuration.