Audio features for DSL-S 2023 shared task
<p>These files are features (i-vectors, x-vectors, MFCC features) extracted from the subset of <a href="https://commonvoice.mozilla.org/">Mozilla Common Voice</a> corpus version 12.0 used in the <a href="https://sites.google.com/view/vardial-2023">VarDial 2023</a> shared task on <a href="https://dsl-s.github.io/">Discriminating Between Similar Languages - Speech</a> (DSL-S 2023).</p> <p>This data set contains only the features for the training and development section (and the Common Voice meta data) for the nine languages included in the shared task (see shared task website for further information).</p> <p> </p>
ShareScore
44/100
Overall dataset sharing score
Score breakdown
These five areas show where the dataset supports — or may limit — practical reuse.
- Stewardship
- 8
- Harmonization
- 4
- Access
- 20
- Reuse readiness
- 8
- Engagement
- 4