CoolFace
26 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01sarulab-speech /commonvoice22_sidongated CV22-Sidon Overview This dataset hosts a release of Mozilla Common Voice 22 restored with the Sidon speech restoration model. Source: Mozilla Common Voice 22.0 Processing: Sidon denoising (sarulab-speech/sidon-v0.1) with 21 s chunks and 48 kHz reconstruction Format: WebDataset shards (.tar.gz) Manifest: paths.yaml enumerates every shard path for Hugging Face–style loading License: Original Common Voice license (CC0 1.0) Languages 137 language folders are… See the full description on the dataset page: https://huggingface.co/datasets/sarulab-speech/commonvoice22_sidon.audiotext-to-speech10M<n<100M30 likes1.9k downloads1y agoHugging Face02TTS-AGI /commonvoice22-sidon-dacvae CommonVoice 22 (Sidon-enhanced) converted to DAC VAE latents Source sarulab-speech/commonvoice22_sidon Format Each tar shard (~2GB) contains samples with three files per sample: {sample_key}.audio.flac # Original audio (FLAC, original sample rate) {sample_key}.dacvae.npy # DAC VAE latent [T_latent, 128] numpy float32 {sample_key}.metadata.json # All metadata + duration_seconds + chars_per_second DAC VAE Latent Format Model:… See the full description on the dataset page: https://huggingface.co/datasets/TTS-AGI/commonvoice22-sidon-dacvae.audioautomatic-speech-recognition1M<n<10M1 likes1.3k downloads6mo agoHugging Face03laion /common-voice-subset-for-clapaudion<1K1 likes509 downloads9mo agoHugging Face04aiintelligentsystems /vel_commons_wikidata Visual Entity Linking: Wikimedia Commons & Wikidata This dataset allows to train and evaluate ML models that link Wikimedia Commons images to the Wikidata items they depict. Disclaimer: All images contained in this dataset are generally assumed to be freely usable (as intended for Wikimedia Commons). Each image's license and author/ uploader is - to the best of our ability - reported in its metadata (see section Dataset Structure). If you want your image's attribution changed or the… See the full description on the dataset page: https://huggingface.co/datasets/aiintelligentsystems/vel_commons_wikidata.image100K<n<1M5 likes355 downloads2y agoHugging Face05Sh1man /common_voice_21_ru Dataset Description Набор данных validated.tsv отфильтрованный по down_votes = 0 📊 Статистика датасета Информация по сплитам 🔹 Тренировочный набор (train) Метрика Значение Количество семплов 93,531 Общая продолжительность 132.25 часов (476,089.70 секунд) Средняя продолжительность семпла 5.09 секунд 🔹 Валидационный набор (validate) Метрика Значение Количество семплов 38,836 Общая продолжительность 55.21… See the full description on the dataset page: https://huggingface.co/datasets/Sh1man/common_voice_21_ru.audio100K<n<1M5 likes192 downloads1y agoHugging Face06RiddleHe /commonsense-baselineimage1K<n<10K0 likes117 downloads3y agoHugging Face07OpenVideo /YouTube-Commons-5G-Rawtextn<1K1 likes95 downloads2y agoHugging Face08ming030890 /common_voice_21_0_yuecantonese only audio10K<n<100K0 likes94 downloads1y agoHugging Face09OpenVideo /Youtube-Common-First-600textn<1K0 likes81 downloads2y agoHugging Face10Scralius /common_voice_16_1_fr_smallaudio100K<n<1M1 likes68 downloads3y agoHugging Face11humanify /common_voice_englishaudio1M<n<10M1 likes66 downloads6mo agoHugging Face12wusize /common_poolimage1M<n<10M0 likes60 downloads2y agoHugging Face13Peacockery /mozilla-common-voice-spontaneous-speech-asr-shared-task Mozilla Common Voice Spontaneous Speech ASR Shared Task This repository combines the Mozilla Data Collective Common Voice spontaneous speech ASR shared-task train/dev and test archives in one place. Locales present across the combined train/dev and test packages: ady, aln, bas, bew, bxk, cgg, el-CY, hch, kbd, kcn, koo, led, lke, lth, meh, mmc, pne, qxp, ruc, rwm, sco, tob, top, ttj, ukv, ush. Split package Mozilla Data Collective dataset ID Hub archive Original MDC archive… See the full description on the dataset page: https://huggingface.co/datasets/Peacockery/mozilla-common-voice-spontaneous-speech-asr-shared-task.textautomatic-speech-recognition10K<n<100K0 likes43 downloads3mo agoHugging Face14fsicoli /common_voice_16_1audio10K<n<100K0 likes34 downloads3y agoHugging Face15onlysainaa /common-voice-mn-24 🇲🇳 Common Voice Mongolian 24.0 Dataset This repository hosts the latest release (v24.0) of the Mozilla Common Voice Scripted Speech dataset for Mongolian (mn). This dataset is a vital resource for training robust Automatic Speech Recognition (ASR) and Text-to-Speech (TTS) systems for the Mongolian language. 📊 Dataset Statistics Metric Value Total Clips 96,308 Total Duration 140.56 Hours Validated Duration 49.19 Hours Total Speakers 606 Format… See the full description on the dataset page: https://huggingface.co/datasets/onlysainaa/common-voice-mn-24.text10K<n<100K0 likes29 downloads8mo agoHugging Face16onlysainaa /common-voice-scripted-speech-24.0-mongoliantext10K<n<100K0 likes27 downloads8mo agoHugging Face17TheSeriousProgrammer /Common-Screensimage1K<n<10K0 likes27 downloads6mo agoHugging Face18archivartaunik /commonvoice22_sidon_be_rawaudio100K<n<1M0 likes19 downloads6mo agoHugging Face19fcanercan /common_voice_14audio10K<n<100K0 likes16 downloads3y agoHugging Face20humair025 /UrduSpeech-CommonVoice22-SIDONaudio10K<n<100K0 likes16 downloads10mo agoHugging Face21Yuyang2022 /Common_Voice_Delta_Segment_11.0audion<1K1 likes14 downloads4y agoHugging Face22leungtianle /huawei-common-senseaudio10K<n<100K0 likes10 downloads11mo agoHugging Face23SEIEZ /Common_Voice_Corpus_21.0text100K<n<1M0 likes8 downloads1y agoHugging Face24OpenSound /CapSpeech-CommonVoicegated CapSpeech-CommonVoice Audio DataSet used for the paper: CapSpeech: Enabling Downstream Applications in Style-Captioned Text-to-Speech Please refer to 🤗CapSpeech for the whole dataset and 🚀CapSpeech repo for more details. Overview 🔥 CapSpeech is a new benchmark designed for style-captioned TTS (CapTTS) tasks, including style-captioned text-to-speech synthesis with sound effects (CapTTS-SE), accent-captioned TTS (AccCapTTS), emotion-captioned TTS (EmoCapTTS) and… See the full description on the dataset page: https://huggingface.co/datasets/OpenSound/CapSpeech-CommonVoice.audio10K<n<100K1 likes7 downloads1y agoHugging Face25Flutra /common_voice_sq_20_localtext1K<n<10K0 likes5 downloads2y agoHugging Face26xiaojiao123 /common_6text10K<n<100K0 likes3 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.