CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01fixie-ai /common_voice_17_0audio10M<n<100M18 likes197k downloads2y agoHugging Face02anke01 /uyghur-common-voice-tts Uyghur Common Voice TTS Dataset A cleaned and processed Text-to-Speech (TTS) dataset for the Uyghur language, derived from Mozilla Common Voice. Dataset Summary Property Value Language Uyghur (ug) Total Samples 43,054 Train Samples 40,901 Validation Samples 2,153 Audio Format WAV Source Mozilla Common Voice License CC0-1.0 Dataset Structure / ├── train.jsonl # Training data (40,901 samples) ├── val.jsonl #… See the full description on the dataset page: https://huggingface.co/datasets/anke01/uyghur-common-voice-tts.audiotext-to-speech10K<n<100K0 likes5.2k downloads7mo agoHugging Face03SpeechTest /common_voice_16_0audio100K<n<1M0 likes3.2k downloads8mo agoHugging Face04hezarai /common-voice-13-faThe Persian portion of the original CommonVoice 13 dataset at https://huggingface.co/datasets/mozilla-foundation/common_voice_13_0 Load # Using HF Datasets from datasets import load_dataset dataset = load_dataset("hezarai/common-voice-13-fa", split="train") # Using Hezar from hezar.data import Dataset dataset = Dataset.load("hezarai/common-voice-13-fa", split="train") audioautomatic-speech-recognition10K<n<100K1 likes3k downloads2y agoHugging Face05mteb /common_voice_21_0_miniaudio10K<n<100K0 likes2k downloads9mo agoHugging Face06sarulab-speech /commonvoice22_sidongated CV22-Sidon Overview This dataset hosts a release of Mozilla Common Voice 22 restored with the Sidon speech restoration model. Source: Mozilla Common Voice 22.0 Processing: Sidon denoising (sarulab-speech/sidon-v0.1) with 21 s chunks and 48 kHz reconstruction Format: WebDataset shards (.tar.gz) Manifest: paths.yaml enumerates every shard path for Hugging Face–style loading License: Original Common Voice license (CC0 1.0) Languages 137 language folders are… See the full description on the dataset page: https://huggingface.co/datasets/sarulab-speech/commonvoice22_sidon.audiotext-to-speech10M<n<100M30 likes1.8k downloads1y agoHugging Face07TTS-AGI /commonvoice22-sidon-dacvae CommonVoice 22 (Sidon-enhanced) converted to DAC VAE latents Source sarulab-speech/commonvoice22_sidon Format Each tar shard (~2GB) contains samples with three files per sample: {sample_key}.audio.flac # Original audio (FLAC, original sample rate) {sample_key}.dacvae.npy # DAC VAE latent [T_latent, 128] numpy float32 {sample_key}.metadata.json # All metadata + duration_seconds + chars_per_second DAC VAE Latent Format Model:… See the full description on the dataset page: https://huggingface.co/datasets/TTS-AGI/commonvoice22-sidon-dacvae.audioautomatic-speech-recognition1M<n<10M1 likes988 downloads6mo agoHugging Face08tuanmanh28 /VIVOS_CommonVoice_FOSD_Control_processed_dataset Dataset Card for "VIVOS_CommonVoice_FOSD_Control_processed_dataset" More Information needed audio10K<n<100K2 likes974 downloads3y agoHugging Face09Scicom-intl /CommonVoice22-Sidon-HQ CommonVoice22-Sidon-HQ The top-DNSMOS slice of sarulab-speech/commonvoice22_sidon — Mozilla Common Voice 22.0 restored to 48 kHz by sarulab-speech/sidon-v0.1, then scored clip-by-clip with DNSMOS P.835 and cut down to only the cleanest utterances. All 14,946,932 source clips (20,746 h, 2.61 TB of FLAC) were scored; 613,305 (4.1%, 1,021 h) passed and are published here. Filter DNSMOS P.835 (speechmos, ONNX) on a single centred 10 s window at 16 kHz… See the full description on the dataset page: https://huggingface.co/datasets/Scicom-intl/CommonVoice22-Sidon-HQ.audio100K<n<1M0 likes951 downloads2mo agoHugging Face10fixie-ai /common_voice_17_0_timestampsaudio1M<n<10M2 likes706 downloads2y agoHugging Face11BrunoHays /mixed_multilingual_commonvoice_all_languages_100kBuild from mozilla commonvoice 13 using the script commited in this repo. Used to teach a model to ignore languages that are not french audio100K<n<1M0 likes664 downloads2y agoHugging Face12saeedzou /common-voice-17-en-age-gender-accentaudio100K<n<1M0 likes655 downloads2mo agoHugging Face13dmnph /common_voice_16_1_hi_pseudo_labelledaudio100K<n<1M0 likes635 downloads2y agoHugging Face14JacobLinCool /common_voice_19_0_zh-TW Common Voice Corpus 19.0 Chinese (Taiwan) The test set is the same as the original test set, while validated_without_test includes all validated examples except those with sentence IDs that appear in the test set. validated_without_test has about 50,000 examples in total, equivalent to approximately 44 hours, and is intended for use as the training set. test has about 5,000 examples, which is approximately 5 hours. audioautomatic-speech-recognition10K<n<100K3 likes606 downloads2y agoHugging Face15malaysia-ai /common_voice_22_0 Common Voice Corpus 22.0 Originally from https://huggingface.co/datasets/fsicoli/common_voice_22_0, we mirror using multiple zip files also trimmed the silents. How to prepare the dataset huggingface-cli download --repo-type dataset \ --include '*.zip' \ --local-dir './' \ --max-workers 20 \ malaysia-ai/common_voice_22_0 wget https://gist.githubusercontent.com/huseinzol05/2e26de4f3b29d99e993b349864ab6c10/raw/9b2251f3ff958770215d70c8d82d311f82791b78/unzip.py python3… See the full description on the dataset page: https://huggingface.co/datasets/malaysia-ai/common_voice_22_0.audio10M<n<100M1 likes573 downloads1y agoHugging Face16kawsarahmd /common_voice_13_0_bn_multi_splitaudio1M<n<10M0 likes566 downloads2y agoHugging Face17lmms-lab-audio /common_voice_15audio10K<n<100K1 likes525 downloads2y agoHugging Face18laion /common-voice-subset-for-clapaudion<1K1 likes524 downloads9mo agoHugging Face19dmnph /common_voice_17_0_en_pseudo_labelledaudio100K<n<1M0 likes511 downloads2y agoHugging Face20OpenSpeechHub /Common-Voice-17-Jaaudio100K<n<1M2 likes499 downloads1y agoHugging Face21JackyHoCL /common_voice_22_yue2025-08-03 Update: Use MP3 instead of WAV All Right reserved by mozilla-foundation audio100K<n<1M1 likes491 downloads1y agoHugging Face22malaysia-ai /common_voice_17_0 Common Voice Corpus 17.0 Mirror for mozilla-foundation/common_voice_17_0, easy to download and extract instead audio in parquet files. How to prepare the dataset huggingface-cli download --repo-type dataset \ --include '*.zip' \ --local-dir './' \ --max-workers 20 \ malaysia-ai/common_voice_17_0 wget https://gist.githubusercontent.com/huseinzol05/2e26de4f3b29d99e993b349864ab6c10/raw/9b2251f3ff958770215d70c8d82d311f82791b78/unzip.py python3 unzip.py audio1M<n<10M0 likes489 downloads1y agoHugging Face23Jayem-11 /mozilla_commonvoice_hackathon_preprocessed_train_batch_3 Dataset Card for "mozilla_commonvoice_hackathon_preprocessed_train_batch_3" More Information needed audio10K<n<100K0 likes486 downloads3y agoHugging Face24mikelalda /common_voice_17_0_eu Common Voice 17.0 - Euskera with Audio Features This dataset contains the Euskera (Basque) subset of the Mozilla Common Voice 17.0 dataset. It includes audio clips and corresponding transcriptions, prepared with audio features readily available for use with libraries like Hugging Face datasets. For the preparation of this dataset, the following notebook was used: Preparar_dataset_euskera.ipynb Source: The original data is from the Mozilla Common Voice 17.0 dataset. Language: Euskera… See the full description on the dataset page: https://huggingface.co/datasets/mikelalda/common_voice_17_0_eu.audio100K<n<1M0 likes485 downloads1y agoHugging Face25pymmdrza /Common-Voice-Speech-26.0-Persian-Clean Persian Common Voice Clean Dataset This dataset is a cleaned and prepared subset of the Persian (فارسی - fa) portion of Mozilla Common Voice Scripted Speech, based on cv-corpus-26.0-2026-06-12. The cleaned release contains 34,134 audio clips, representing approximately 43.105 hours of speech, equal to 2,586.303 minutes. The clips are associated with approximately 34,134 validated Persian sentences and come from 3,791 speakers. The original Persian Common Voice release contains… See the full description on the dataset page: https://huggingface.co/datasets/pymmdrza/Common-Voice-Speech-26.0-Persian-Clean.audiotext-to-speech10K<n<100K2 likes478 downloads1mo agoHugging Face26MohammadGholizadeh /common-voice-17-farsiaudio100K<n<1M3 likes477 downloads1y agoHugging Face27MohamedRashad /common-voice-18-arabic Dataset Card for Common Voice 18 – Arabic Edition Dataset Summary This dataset is an unofficial Arabic-only extraction of Mozilla Common Voice Corpus 18.0, prepared for Automatic Speech Recognition (ASR) research and development. It is derived from the original Common Voice 18 release and filtered to include Arabic (ar) speech data only, while preserving the original dataset structure, splits, and metadata fields. The dataset consists of validated, unvalidated, and… See the full description on the dataset page: https://huggingface.co/datasets/MohamedRashad/common-voice-18-arabic.audioautomatic-speech-recognition100K<n<1M5 likes467 downloads9mo agoHugging Face28saeedzou /common-voice-17-en-age-genderaudio100K<n<1M0 likes459 downloads2mo agoHugging Face29parandl /common_voice_16_1_hi_pseudo_labelledaudio100K<n<1M0 likes405 downloads2y agoHugging Face30NathanRoll /commonvoice_train_gender_accent_16k Dataset Card for "commonvoice_train_gender_accent_16k" More Information needed audio100K<n<1M1 likes391 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.