CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01UncovAI /FOR-normaudio0 likes2.4k downloads1y agoHugging Face02JackismyShephard /nst-da-norm Dataset Card for NST-da Normalized Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): da License: cc0-1.0 Dataset Sources [optional] Repository: [More Information Needed] Paper [optional]: [More Information Needed] Demo [optional]: [More Information Needed] Uses Direct Use… See the full description on the dataset page: https://huggingface.co/datasets/JackismyShephard/nst-da-norm.audioautomatic-speech-recognition100K<n<1M2 likes364 downloads3y agoHugging Face03Twelve2five /igbo_tts_normalizedaudio100K<n<1M2 likes240 downloads1y agoHugging Face04SPEAK-ASR /openslr-sinhala-asr-normaudio10K<n<100K0 likes200 downloads7mo agoHugging Face05simpra /xh-tts-vixsd-norm isiXhosa TTS clips (ViXSD, segmented) 3,861 clips, 22,050 Hz mono, 8 speakers, cut from long-form recordings by CTC forced alignment. Derived from ViXSD (Vuk'uzenzele isiXhosa Speech Dataset) by Lelapa AI / Way With Words, under the Esethu License — see https://huggingface.co/datasets/lelapa/Vukuzenzele_isiXhosa_Speech_Dataset_ViXSD Pipeline vixsd_extract.py — parquet to mono 22,050 Hz. Source is heterogeneous: rates 16k/22.05k/44.1k/48k/96k, depths 16/24/32, PCM… See the full description on the dataset page: https://huggingface.co/datasets/simpra/xh-tts-vixsd-norm.audiotext-to-speech1K<n<10K0 likes165 downloads12d agoHugging Face06Eimhin03 /Fleurs_Irish_normalizedaudio1K<n<10K0 likes144 downloads6mo agoHugging Face07ArissBandoss /fake_or_real_dataset_for_normaudio10K<n<100K0 likes124 downloads2y agoHugging Face08simpra /xh-tts-slr32-norm isiXhosa TTS — SLR32 prepared for VITS Multi-speaker isiXhosa speech, resampled and text-normalised for VITS training. Each audio file is paired with its transcript in metadata.csv. Attribution (required by the licence) Derived from OpenSLR SLR32, "High quality TTS data for four South African languages (af, st, tn, xh)", created by North West University and Google (2017), released under CC BY-SA 4.0. Source: https://openslr.org/32/ This derivative is likewise CC… See the full description on the dataset page: https://huggingface.co/datasets/simpra/xh-tts-slr32-norm.audiotext-to-speech1K<n<10K0 likes122 downloads12d agoHugging Face09hana92 /Arabic-Diacritized-TTS-Normalized Arabic-Diacritized-TTS Dataset Overview The Arabic-Diacritized-TTS dataset contains Arabic audio samples and their corresponding text with full diacritization. This dataset is designed to support research in Arabic speech processing, text-to-speech (TTS) synthesis, automatic diacritization, and other natural language processing (NLP) tasks. Dataset Contents Audio Samples: High-quality Arabic speech recordings. Text Transcriptions: Fully diacritized Arabic text… See the full description on the dataset page: https://huggingface.co/datasets/hana92/Arabic-Diacritized-TTS-Normalized.audio1K<n<10K0 likes119 downloads7mo agoHugging Face10MoaazTalab /ASVspoof_2021_DF_Balanced_Normalizedaudio100K<n<1M5 likes118 downloads2y agoHugging Face11warmestman /common-voice-20-mn-normalized Common Voice 20.0 Mongolian Dataset This dataset is a subset of Mozilla's Common Voice project, containing Mongolian speech data. It's part of Common Voice 20.0 release. Dataset Structure The dataset contains: Audio clips in .mp3 format Transcriptions for each audio clip Train/test/dev splits Additional metadata including speaker demographics Usage This dataset can be used for: Speech Recognition Voice Analysis Linguistic Research Speech Processing… See the full description on the dataset page: https://huggingface.co/datasets/warmestman/common-voice-20-mn-normalized.audioautomatic-speech-recognition10K<n<100K4 likes109 downloads2y agoHugging Face12omarabb315 /masc_filtered_normalizedaudio100K<n<1M0 likes98 downloads1y agoHugging Face13ghananlpcommunity /new-twi-tts-aligned_normalised This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. audio100K<n<1M0 likes95 downloads3mo agoHugging Face14MoaazTalab /ASVspoof_2021_LA_Balanced_Normalizedaudio100K<n<1M2 likes83 downloads2y agoHugging Face15groxaxo /google-latam-spanish-boundary-normalized Google LATAM Spanish Boundary-Normalized Audio Female Spanish speech from the following upstream datasets: Argentina: ylacombe/google-argentinian-spanish Chile: ylacombe/google-chilean-spanish Colombia: ylacombe/google-colombian-spanish Attribution and Thanks Many thanks to ylacombe for publishing and maintaining the original Argentinian, Chilean, and Colombian Spanish datasets. The recordings, transcripts, speaker labels, and original dataset structure come… See the full description on the dataset page: https://huggingface.co/datasets/groxaxo/google-latam-spanish-boundary-normalized.audiotext-to-speech1K<n<10K1 likes79 downloads2mo agoHugging Face16sanchit-gandhi /expresso-concatenated-half-normalaudio1K<n<10K0 likes66 downloads2y agoHugging Face17luigisaetta /atco2_normalized_augmentedaudio1K<n<10K0 likes65 downloads4y agoHugging Face18SAadettin-BERber /normalized_train_ATC_datasetaudio1K<n<10K0 likes60 downloads1y agoHugging Face19Eimhin03 /MCV_Fleurs_Combined_Irish_normalizedaudio10K<n<100K0 likes55 downloads6mo agoHugging Face20Yehor /my-voice-normalizedaudion<1K0 likes38 downloads1y agoHugging Face21SAadettin-BERber /normalized_test_ATC_datasetaudio1K<n<10K0 likes37 downloads1y agoHugging Face22BaseLayer /uzbek-normal-speech-10haudio1K<n<10K0 likes36 downloads3d agoHugging Face23Farhan2000 /UrduTTS-normalizedaudio10K<n<100K0 likes30 downloads1d agoHugging Face24Eimhin03 /MCV25_Irish_normalizedaudio10K<n<100K0 likes29 downloads6mo agoHugging Face25greentechapps /iqra_curated_normalised_1s_20s_finalaudio10K<n<100K0 likes23 downloads9mo agoHugging Face26Farhan2000 /UrduTTS-normalized-shortaudio10K<n<100K0 likes23 downloads1d agoHugging Face27mohammed-bahumaish /vocalsound-normalizedaudio1K<n<10K0 likes19 downloads6mo agoHugging Face28MoaazTalab /ASVspoof_2021_DF1_Balanced_Normalizedaudio100K<n<1M0 likes18 downloads2y agoHugging Face29vietnhat /orpheus-synthetic-dataset-normalizedaudion<1K0 likes14 downloads1y agoHugging Face30SPEAK-ASR /openslr-sinhala-asr-norm-noise-remaudio10K<n<100K0 likes14 downloads7mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.