Skip to content
RAG Repo

Mozilla Common Voice

Common Voice is Mozilla's effort to build an open, public-domain speech corpus that anyone can use to train voice technology. Speakers record themselves reading short sentences, and a dual-review process (two independent approvals) validates each clip before it enters a release. Because everything is released under CC0, you can use the audio and transcripts for any purpose, including commercial products, with no attribution required.

Each release is versioned and covers a growing set of languages, with hours per language ranging from a handful to thousands. The recordings are paired with transcripts, so the data feeds directly into automatic speech recognition (ASR, turning spoken audio into text) and text-to-speech training pipelines.

Note that from October 2025, Common Voice datasets are only available through the Mozilla Data Collective, not HuggingFace. If you previously pulled releases from HuggingFace, you will need to switch to the Data Collective to get the latest data.

speechmultilingualcrowdsourcedcc0asrnonprofit

Related sources