CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01sanchit-gandhi /tedlium-data Dataset Card for "tedlium-data" More Information needed audio100K<n<1M3 likes2.8k downloads3y agoHugging Face02sanchit-gandhi /vctk Dataset Card for "vctk" More Information needed audio10K<n<100K2 likes2.7k downloads3y agoHugging Face03sanchit-gandhi /librispeech-data Dataset Card for "librispeech-data" More Information needed audio100K<n<1M2 likes1.1k downloads3y agoHugging Face04Reza2kn /ganjoor-recitations Ganjoor Persian Poetry Recitations (Full) Every published audio recitation on Ganjoor / AVA paired with its transcription — 30,133 clips, 1,276 hours of audio. Audio is stored full-length and unchunked, and every clip carries a single clean transcription in text, so it's ready for ASR / TTS training as-is. Columns column description audio full-length mp3 (native sample rate), embedded and playable text full transcription of the clip… See the full description on the dataset page: https://huggingface.co/datasets/Reza2kn/ganjoor-recitations.audioautomatic-speech-recognition10K<n<100K3 likes506 downloads3mo agoHugging Face05sanchit-gandhi /gtzan Dataset Card for "gtzan" More Information needed audion<1K1 likes447 downloads3y agoHugging Face06sanchit-gandhi /earnings22_robust_splitfrom datasets import load_dataset, DatasetDict ds = load_dataset("anton-l/earnings22_robust", split="test") print(ds) print("\n", "Split to ==>", "\n") # split train 90%/ dev 5% / test 5% # split twice and combine train_devtest = ds.train_test_split(shuffle=True, seed=1, test_size=0.1) dev_test = train_devtest['test'].train_test_split(shuffle=True, seed=1, test_size=0.5) ds_train_dev_test = DatasetDict({'train': train_devtest['train'], 'validation': dev_test['train'], 'test':… See the full description on the dataset page: https://huggingface.co/datasets/sanchit-gandhi/earnings22_robust_split.audio10K<n<100K0 likes381 downloads4y agoHugging Face07sanchit-gandhi /earnings22_splitWe partition the earnings22 dataset at https://huggingface.co/datasets/anton-l/earnings22_baseline_5_gram by source_id: Validation: 4420696 4448760 4461799 4469836 4473238 4482110 Test: 4432298 4450488 4470290 4479741 4483338 4485244 Train: remainder Official script for processing these splits will be released shortly. audio10K<n<100K0 likes315 downloads4y agoHugging Face08Reza2kn /ganjoor-recitations-chunked 🗂️ ganjoor-recitations-chunked English + فارسی · Part of Shenava 1.0 · Project hub · SLT paper submission 🌟 At a glance | معرفی سریع English فارسی 🎯 Purpose Ganjoor recitation chunked ASR dataset. قطعه‌های تلاوت و خوانش گنجور برای آموزش و ارزیابی گفتار ادبی، شعر و خوانش رسمی فارسی. 🧩 Role Persian speech dataset مجموعه‌دادهٔ گفتار فارسی 📦 Snapshot 64 files; approximately 118.09 GB 64 فایل؛ حدود 118.09 GB 🧱 Packaging 61 Parquet files and 0… See the full description on the dataset page: https://huggingface.co/datasets/Reza2kn/ganjoor-recitations-chunked.audioautomatic-speech-recognition100K<n<1M1 likes307 downloads2mo agoHugging Face09sanchit-gandhi /librispeech_asr_dummy Dataset Card for librispeech_asr_dummy Dataset Summary This is a truncated version of the LibriSpeech dataset. It contains 20 samples from each of the splits. To view the full dataset, visit: https://huggingface.co/datasets/librispeech_asr LibriSpeech is a corpus of approximately 1000 hours of 16kHz read English speech, prepared by Vassil Panayotov with the assistance of Daniel Povey. The data is derived from read audiobooks from the LibriVox project, and has been… See the full description on the dataset page: https://huggingface.co/datasets/sanchit-gandhi/librispeech_asr_dummy.audioautomatic-speech-recognitionn<1K0 likes204 downloads3y agoHugging Face10sanchit-gandhi /audioldm-readme-samplesaudion<1K0 likes183 downloads3y agoHugging Face11sanchit-gandhi /earnings22_split_resampledWe partition the earnings22 dataset at https://huggingface.co/datasets/anton-l/earnings22_baseline_5_gram by source_id: Validation: 4420696 4448760 4461799 4469836 4473238 4482110 Test: 4432298 4450488 4470290 4479741 4483338 4485244 Train: remainder Official script for processing these splits will be released shortly. audio10K<n<100K0 likes176 downloads4y agoHugging Face12farsi-asr /ganjoor-chunked-asr-datasetaudio100K<n<1M2 likes153 downloads2y agoHugging Face13sanchit-gandhi /rev16_csvaudion<1K0 likes123 downloads3y agoHugging Face14sanchit-gandhi /libritts_r_testaudio1K<n<10K1 likes102 downloads2y agoHugging Face15AbhijeetK3 /ganapati-atharvashirsha-chantingaudion<1K0 likes102 downloads12d agoHugging Face16farsi-asr /ganjoor-datasetaudio10K<n<100K0 likes93 downloads2y agoHugging Face17sanchit-gandhi /edacc Draft conversion of EdAcc Final dataset will be moved to the edinburghcstr organisation. audio10K<n<100K1 likes92 downloads3y agoHugging Face18Alirezav99 /ganjoor Ganjoor Persian Speech Dataset Dataset Description This dataset contains Persian speech recordings from Ganjoor.ir, segmented based on Persian poetry verses. The audio files have been processed and segmented into manageable chunks suitable for speech-to-text (STT) training and evaluation. Dataset Summary Language: Persian (Farsi) Domain: Persian poetry (classical and contemporary) Task: Automatic Speech Recognition (ASR) Format: MP3 audio files with… See the full description on the dataset page: https://huggingface.co/datasets/Alirezav99/ganjoor.audion<1K0 likes89 downloads8mo agoHugging Face19sanchit-gandhi /expresso-concatenated-half-normalaudio1K<n<10K0 likes65 downloads2y agoHugging Face20svjack /genshin_impact_ganyu_audio_sample audion<1K0 likes48 downloads1y agoHugging Face21sanchit-gandhi /common_voice_16_1_hi_pseudo_labelled Common Voice 16.1 Hindi Pseudo-Labelled This is the Common Voice 16.1 Hindi split pseudo-labelled using the Whisper large-v3 model, according to the instructions detailed in the Distil-Whisper repository. To reproduce this pseudo-labelling run, follow the instructions detailed here. audio1K<n<10K0 likes47 downloads2y agoHugging Face22Reza2kn /ganjoor-chunk-smoke Ganjoor Recitations — Clean-Cut Chunks Training-ready ~15-20 s segments derived from Reza2kn/ganjoor-recitations (full-length Persian poetry recitations from ganjoor.net). Two columns only: audio (16 kHz mono) and text — same schema as the source, just many more rows of shorter clips + matching labels. How it was chunked (never mid-word) Per recitation (tools/ganjoor_chunk_job.py): Forced-align gold text to audio with torchaudio MMS_FA (uroman -> per-word times +… See the full description on the dataset page: https://huggingface.co/datasets/Reza2kn/ganjoor-chunk-smoke.audioautomatic-speech-recognitionn<1K0 likes41 downloads3mo agoHugging Face23sanchit-gandhi /common_voice_13_0_hi_pseudo_labelled Dataset Card for "common_voice_13_0_hi_pseudo_labelled" More Information needed audio1K<n<10K0 likes37 downloads3y agoHugging Face24AbhijeetK3 /ganapati-atharvashirsha-chanting-wadkaraudion<1K0 likes33 downloads12d agoHugging Face25sanchit-gandhi /librispeech_asr_dummy_pseudo_labelledaudion<1K0 likes29 downloads3y agoHugging Face26sanchit-gandhi /concatenated-datasetaudio10K<n<100K0 likes29 downloads3y agoHugging Face27sanchit-gandhi /whisper-jax-test-files Dataset Card for "whisper-jax-test-files" More Information needed audion<1K5 likes27 downloads3y agoHugging Face28Ganaa0614 /mongolian-commonvoice-stt-translated-fullaudio10K<n<100K1 likes25 downloads5mo agoHugging Face29sanchit-gandhi /librispeech_asr_dummy_noise-noise Dataset Card for "librispeech_asr_dummy_noise-noise" More Information needed audion<1K1 likes24 downloads3y agoHugging Face30sanchit-gandhi /voxpopuli_dummyaudion<1K0 likes24 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.