FAQ: Why can’t I download old versions of Common Voice datasets?

We limit access to old versions of Common Voice to respect those speakers who have withdrawn their consent to be included in the dataset.

Share

In order to better serve our community and to keep up with current changes in best practices for data stewardship, we are changing how previous releases of the Common Voice datasets are accessed. They will now be accessible to interested researchers and others via an email-based request process. You will send us an email using a valid return address and agree to the terms of use, and we will send you a link to download the dataset. This means that we will be able to more comprehensively honour deletion requests. 

We are committed to supporting researchers and teachers in carrying out reproducible research, but we are adding a step so that we can confirm purpose and track usage. 

Read the full post for details and how to request an older version: We’re changing access to older versions of Common Voice datasets

Read more

Mozilla Data Collective datasets now discoverable through CLARIN’s Virtual Language Observatory

New collaboration expands visibility for community-governed language datasets and improves exploration of linguistic resources, services and tools. Mozilla Data Collective datasets are now discoverable through CLARIN’s Virtual Language Observatory, making it easier for researchers, developers and language technology practitioners in Europe to find multilingual and community-centered datasets