FAQ: What is the long-term sustainability model for Mozilla Data Collective?

We ask for a 5% dataset access fee from downloaders of paid datasets and may offer additional other subscription capabilities in the future.

Share

We are a mission-driven, community-centred tech organisation that was incubated at Mozilla Foundation. As of April 2026, Mozilla Data Collective is an independent for-profit entity based in the United Kingdom. We anticipate that we will have both social enterprise and non-profit components eventually, as we think that in today’s volatile, politicised grant funding environment, it’s never been more important to be independent. 

Our business model is simple, ethical and transparent. In the future, if you choose to ask for financial contributions to make use of your dataset, we will charge the downloader a 5% platform fee. That’s it! This will help cover the platform development, storage and hosting costs. One day, we may ask those using datasets at scale - such as major corporations - to pay a bit more for an enterprise-grade API.  

Read more

Mozilla Data Collective datasets now discoverable through CLARIN’s Virtual Language Observatory

New collaboration expands visibility for community-governed language datasets and improves exploration of linguistic resources, services and tools. Mozilla Data Collective datasets are now discoverable through CLARIN’s Virtual Language Observatory, making it easier for researchers, developers and language technology practitioners in Europe to find multilingual and community-centered datasets