Audio features for DSL-S 2023 shared task
Çağrı Çöltekin, Mourhaf Kazzaz, Tommi Jauhiainen, Nikola Ljubešić · Zenodo (CERN European Organization for Nuclear Research) · 2023
These files are features (i-vectors, x-vectors, MFCC features) extracted from the subset of Mozilla Common Voice corpus version 12.0 used in the VarDial 2023 shared task on Discriminating Between Similar Languages - Speech (DSL-S 2023). This data set contains only the features for the training and development section (and the Common Voice meta data) for the nine languages included in the shared task (see shared task website for further information).