AHUMADA: A large speech corpus in Spanish for speaker characterization and identification q

Javier Ortega-García, Joaquín González-Rodríguez, Victoria Marrero Aguiar · 2000

Speaker recognition is an emerging task in both commercial and forensic applications. Nevertheless, while in certain applications we can estimate, adapt or hypothesize about our working conditions, most of the commercial applications and almost the whole of the forensic approaches to speaker recognition are still open problems, due to several reasons. Some of these reasons can be stated: environmental conditions are (usually) rapidly changing or highly degraded, acquisition processes are not always under control, incriminated people exhibit low degree of cooperativeness, etc., inducing a wide range of variability sources on speech utterances. In this sense, real approaches to speaker identification necessarily imply taking into account all these variability factors. In order to isolate, analyze and measure the eAect of some of the main variability sources that can be found in real commercial and forensic applications, and their influence in automatic recognition systems, a specific large speech database in Castilian Spanish called AHUMADA (/aum ada/) has been designed and acquired under controlled conditions. In this paper, together with a detailed description of the database, some experimental results including diAerent speech variability factors are also presented. ” 2000 Elsevier Science B.V. All rights reserved.

Read the paper · More papers on PaperTik