CoolFace
16 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01RidheshBhati /Complete_Data_Source_100K_HOURS Multi-Language Audio Collection (100K Hours) This repository is physically reorganized for Absolute 100% Data Visibility. 🏗️ Global Consolidator Select your language subset to listen to high-quality waveform audio. All shards from legacy and modern pipelines are automatically routed here. audio1M<n<10M4 likes17k downloads5mo agoHugging Face02BrunoHays /mixed_multilingual_commonvoice_all_languages_100kBuild from mozilla commonvoice 13 using the script commited in this repo. Used to teach a model to ignore languages that are not french audio100K<n<1M0 likes668 downloads2y agoHugging Face03AndreasXi /FineLAP-100kaudio100K<n<1M3 likes540 downloads6mo agoHugging Face04seonglae /vls-100k VLS 100K 74,936 MS-COCO images paired with a long written description, a one-sentence spoken summary of that description, the spoken audio, and that audio pre-encoded with a neural codec. Images and audio are embedded in the parquet, so the viewer renders them and one call opens the set: from datasets import load_dataset ds = load_dataset("seonglae/vls-100k", split="train") ds[0]["image"] # PIL image ds[0]["audio"] # decoded waveform ds[0]["sst"] # the sentence that was… See the full description on the dataset page: https://huggingface.co/datasets/seonglae/vls-100k.audiotext-to-speech10K<n<100K0 likes268 downloads29d agoHugging Face05nguyensu27 /VITS_DATASET_60k_100Kaudio100K<n<1M0 likes203 downloads9mo agoHugging Face06AhmedBadawy11 /UAE_100Kaudio10K<n<100K5 likes135 downloads2y agoHugging Face07ittailup /peruvian_speech_more_100kaudio100K<n<1M0 likes60 downloads2y agoHugging Face08ismailelsayedeltanja /Egyptian-dialect-100kaudio100K<n<1M0 likes53 downloads2mo agoHugging Face09OzLabs /qwen3-asr-hebrew-100kaudio100K<n<1M0 likes51 downloads7mo agoHugging Face10abhinav-spidey /IndicVoices-ML-100kaudio100K<n<1M0 likes42 downloads2mo agoHugging Face11Rabe3 /saudi-tts-synthetic-100kaudio1K<n<10K0 likes20 downloads3mo agoHugging Face12krishnakalyan3 /mj_emo_embed_100k Dataset Card for "mj_emo_embed_100k" More Information needed audio100K<n<1M0 likes11 downloads2y agoHugging Face13another-blue /FineLAP-100kaudio100K<n<1M0 likes10 downloads2mo agoHugging Face14RidheshBhati /Complete_100k_Dataaudio1K<n<10K0 likes9 downloads6mo agoHugging Face15MikeHonkers /SOVA-audiobooks-100kaudio100K<n<1M0 likes2 downloads11mo agoHugging Face16AdoCleanCode /en_asr_mls_sub0_al_50-100kgatedaudio10K<n<100K0 likes1 downloads9mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.