CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01fixie-ai /common_voice_17_0audio10M<n<100M18 likes202k downloads2y agoHugging Face02anke01 /uyghur-common-voice-tts Uyghur Common Voice TTS Dataset A cleaned and processed Text-to-Speech (TTS) dataset for the Uyghur language, derived from Mozilla Common Voice. Dataset Summary Property Value Language Uyghur (ug) Total Samples 43,054 Train Samples 40,901 Validation Samples 2,153 Audio Format WAV Source Mozilla Common Voice License CC0-1.0 Dataset Structure / ├── train.jsonl # Training data (40,901 samples) ├── val.jsonl #… See the full description on the dataset page: https://huggingface.co/datasets/anke01/uyghur-common-voice-tts.audiotext-to-speech10K<n<100K0 likes3.9k downloads7mo agoHugging Face03SpeechTest /common_voice_16_0audio100K<n<1M0 likes3.5k downloads8mo agoHugging Face04hezarai /common-voice-13-faThe Persian portion of the original CommonVoice 13 dataset at https://huggingface.co/datasets/mozilla-foundation/common_voice_13_0 Load # Using HF Datasets from datasets import load_dataset dataset = load_dataset("hezarai/common-voice-13-fa", split="train") # Using Hezar from hezar.data import Dataset dataset = Dataset.load("hezarai/common-voice-13-fa", split="train") audioautomatic-speech-recognition10K<n<100K1 likes3k downloads2y agoHugging Face05willcai /wav2vec2_common_voice_accents_3tabular100K<n<1M0 likes2.2k downloads5y agoHugging Face06mteb /common_voice_21_0_miniaudio10K<n<100K0 likes2.1k downloads9mo agoHugging Face07Peacockery /common-voice-scripted-speech-26 Common Voice Scripted Speech A row-normalized multilingual ASR dataset built from Mozilla Data Collective Common Voice Scripted Speech. Each upstream archive is converted to appendable parquet shards under data/<upstream_split>/, one shard per source archive and split, with audio bytes embedded in an audio struct column. Status Manifest languages: 60 Languages uploaded: 18 Columns audio (bytes, path) sentence, locale, language, upstream_split… See the full description on the dataset page: https://huggingface.co/datasets/Peacockery/common-voice-scripted-speech-26.tabularautomatic-speech-recognition100K<n<1M0 likes1.9k downloads3mo agoHugging Face08sarulab-speech /commonvoice22_sidongated CV22-Sidon Overview This dataset hosts a release of Mozilla Common Voice 22 restored with the Sidon speech restoration model. Source: Mozilla Common Voice 22.0 Processing: Sidon denoising (sarulab-speech/sidon-v0.1) with 21 s chunks and 48 kHz reconstruction Format: WebDataset shards (.tar.gz) Manifest: paths.yaml enumerates every shard path for Hugging Face–style loading License: Original Common Voice license (CC0 1.0) Languages 137 language folders are… See the full description on the dataset page: https://huggingface.co/datasets/sarulab-speech/commonvoice22_sidon.audiotext-to-speech10M<n<100M30 likes1.9k downloads1y agoHugging Face09DylanonWic /common_voice_10_1_th_augmented_pitch Dataset Card for "common_voice_10_1_th_augmented_pitch" More Information needed text10K<n<100K0 likes1.5k downloads4y agoHugging Face10AstraMindAI /CommonVoice-POSTPROCESS-f0ecfc0ctext100K<n<1M0 likes1.1k downloads1y agoHugging Face11TTS-AGI /commonvoice22-sidon-dacvae CommonVoice 22 (Sidon-enhanced) converted to DAC VAE latents Source sarulab-speech/commonvoice22_sidon Format Each tar shard (~2GB) contains samples with three files per sample: {sample_key}.audio.flac # Original audio (FLAC, original sample rate) {sample_key}.dacvae.npy # DAC VAE latent [T_latent, 128] numpy float32 {sample_key}.metadata.json # All metadata + duration_seconds + chars_per_second DAC VAE Latent Format Model:… See the full description on the dataset page: https://huggingface.co/datasets/TTS-AGI/commonvoice22-sidon-dacvae.audioautomatic-speech-recognition1M<n<10M1 likes1.1k downloads6mo agoHugging Face12Scicom-intl /CommonVoice22-Sidon-HQ CommonVoice22-Sidon-HQ The top-DNSMOS slice of sarulab-speech/commonvoice22_sidon — Mozilla Common Voice 22.0 restored to 48 kHz by sarulab-speech/sidon-v0.1, then scored clip-by-clip with DNSMOS P.835 and cut down to only the cleanest utterances. All 14,946,932 source clips (20,746 h, 2.61 TB of FLAC) were scored; 613,305 (4.1%, 1,021 h) passed and are published here. Filter DNSMOS P.835 (speechmos, ONNX) on a single centred 10 s window at 16 kHz… See the full description on the dataset page: https://huggingface.co/datasets/Scicom-intl/CommonVoice22-Sidon-HQ.audio100K<n<1M0 likes1k downloads2mo agoHugging Face13tuanmanh28 /VIVOS_CommonVoice_FOSD_Control_processed_dataset Dataset Card for "VIVOS_CommonVoice_FOSD_Control_processed_dataset" More Information needed audio10K<n<100K2 likes977 downloads3y agoHugging Face14Aniket-Tathe-08 /Custom_common_voice_dataset_using_RVC Custom Data Augmentation for low resource ASR using Bark and Retrieval-Based Voice Conversion Custom common_voice_v11 corpus with a custom voice was was created using RVC(Retrieval-Based Voice Conversion) The model underwent 200 epochs of training, utilizing a total of 1 hour of audio clips. The data was scraped from Youtube. The audio in the custom generated dataset is of a YouTuber named Ajay Pandey Description license: cc0-1.0 language: - hi… See the full description on the dataset page: https://huggingface.co/datasets/Aniket-Tathe-08/Custom_common_voice_dataset_using_RVC.tabular10K<n<100K0 likes807 downloads3y agoHugging Face15DylanonWic /common_voice_10_1_th_clean_split_0_old Dataset Card for "common_voice_10_1_th_clean_split_0" More Information needed text10K<n<100K0 likes715 downloads4y agoHugging Face16saeedzou /common-voice-17-en-age-gender-accentaudio100K<n<1M0 likes712 downloads2mo agoHugging Face17fixie-ai /common_voice_17_0_timestampsaudio1M<n<10M2 likes704 downloads2y agoHugging Face18dmnph /common_voice_16_1_hi_pseudo_labelledaudio100K<n<1M0 likes679 downloads2y agoHugging Face19DylanonWic /common_voice_10_1_th_clean_split_1 Dataset Card for "common_voice_10_1_th_clean_split_1_fix_spacial_char" More Information needed text10K<n<100K0 likes678 downloads3y agoHugging Face20BrunoHays /mixed_multilingual_commonvoice_all_languages_100kBuild from mozilla commonvoice 13 using the script commited in this repo. Used to teach a model to ignore languages that are not french audio100K<n<1M0 likes668 downloads2y agoHugging Face21kawsarahmd /common_voice_13_0_bn_multi_splitaudio1M<n<10M0 likes633 downloads2y agoHugging Face22dmnph /common_voice_17_0_en_pseudo_labelledaudio100K<n<1M0 likes627 downloads2y agoHugging Face23vumichien /preprocessed_jsut_jsss_css10_common_voice_11 Dataset Card for "preprocessed_jsut_jsss_css10_common_voice_11" More Information needed text10K<n<100K1 likes623 downloads4y agoHugging Face24DylanonWic /common_voice_10_1_th_clean_split_0 Dataset Card for "common_voice_10_1_th_clean_split_0_fix_spacial_char" More Information needed text10K<n<100K0 likes618 downloads3y agoHugging Face25JacobLinCool /common_voice_19_0_zh-TW Common Voice Corpus 19.0 Chinese (Taiwan) The test set is the same as the original test set, while validated_without_test includes all validated examples except those with sentence IDs that appear in the test set. validated_without_test has about 50,000 examples in total, equivalent to approximately 44 hours, and is intended for use as the training set. test has about 5,000 examples, which is approximately 5 hours. audioautomatic-speech-recognition10K<n<100K3 likes614 downloads2y agoHugging Face26malaysia-ai /common_voice_22_0 Common Voice Corpus 22.0 Originally from https://huggingface.co/datasets/fsicoli/common_voice_22_0, we mirror using multiple zip files also trimmed the silents. How to prepare the dataset huggingface-cli download --repo-type dataset \ --include '*.zip' \ --local-dir './' \ --max-workers 20 \ malaysia-ai/common_voice_22_0 wget https://gist.githubusercontent.com/huseinzol05/2e26de4f3b29d99e993b349864ab6c10/raw/9b2251f3ff958770215d70c8d82d311f82791b78/unzip.py python3… See the full description on the dataset page: https://huggingface.co/datasets/malaysia-ai/common_voice_22_0.audio10M<n<100M1 likes603 downloads1y agoHugging Face27JackyHoCL /common_voice_22_yue2025-08-03 Update: Use MP3 instead of WAV All Right reserved by mozilla-foundation audio100K<n<1M1 likes553 downloads1y agoHugging Face28MohammadGholizadeh /common-voice-17-farsiaudio100K<n<1M3 likes531 downloads1y agoHugging Face29DylanonWic /common_voice_10_1_th_clean_split_3_augment_old Dataset Card for "common_voice_10_1_th_clean_split_3_augment" More Information needed text10K<n<100K0 likes528 downloads3y agoHugging Face30laion /common-voice-subset-for-clapaudion<1K1 likes528 downloads9mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.