CoolFace
26 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01INo0121 /low_quality_call_voice Dataset Card for "low_quality_call_voice" More Information needed audio100K<n<1M0 likes252 downloads3y agoHugging Face02dennohpeter /low-germanaudio1K<n<10K0 likes90 downloads10mo agoHugging Face03malaysia-ai /Low-Language-TTS Low-Language-TTS A held-out TTS/ASR test set for the long tail of malaysia-ai/Multilingual-TTS: 50 of the lowest-resource languages in that corpus, 25 utterances each (1250 rows, 2.56 hours). Every row carries the three things needed to score a NeuCodec speech-token model without touching the parent corpus: column what audio the original clip, exactly as stored upstream (mostly mp3) tokens NeuCodec speech tokens at 50 tokens/s — the <|s_N|> ids the Multilingual-TTS… See the full description on the dataset page: https://huggingface.co/datasets/malaysia-ai/Low-Language-TTS.audio1K<n<10K0 likes63 downloads8d agoHugging Face04lowres /sukasuka-anime-vocal-datasetthis dataset is the parquet version of the dataset that was created by mio original dataset link : https://huggingface.co/datasets/mio/sukasuka-anime-vocal-dataset please make sure to follow and heart react the original author (≧∇≦)ノ audioaudio-classification1K<n<10K2 likes60 downloads3y agoHugging Face05shun31y /uaspeech_tts_very_lowaudio10K<n<100K0 likes55 downloads8mo agoHugging Face06lowo /vertin_train_datasetaudio1K<n<10K6 likes50 downloads3y agoHugging Face07Reubencf /Adaption-low-resource-audio Adaption Low-Resource Audio A low-resource-language subset of Reubencf/PolyglotAudio, remastered with Adaption's Adaptive Data platform. Each row carries the original Tatoeba-derived audio clip alongside sharpened enhanced_prompt / enhanced_completion columns so the data is ready for speech-model fine-tuning and evaluation on languages that are typically under-represented in open ASR/TTS corpora. Dataset size 3,704 rows of paired audio + text, spanning 10 languages… See the full description on the dataset page: https://huggingface.co/datasets/Reubencf/Adaption-low-resource-audio.audioautomatic-speech-recognition1K<n<10K1 likes34 downloads5mo agoHugging Face08AhmedBadawy11 /alsallom_update_para_UAE_transcription_low_chunk_by_elevenlabaudion<1K0 likes26 downloads1y agoHugging Face09AIGenLab /high-sound-and-low-musicaudio10K<n<100K0 likes22 downloads10mo agoHugging Face10lca0503 /interleaving_mmau_lowest_pitchaudio1K<n<10K0 likes21 downloads1y agoHugging Face11nhatminh /hmong_low_bleuaudio1K<n<10K0 likes20 downloads11mo agoHugging Face12shun31y /uaspeech_tts_lowaudio10K<n<100K0 likes20 downloads8mo agoHugging Face13Surpem /low-decoder low-decoding Author: Surpem This dataset contains 1200 unique, clean synthetic audio signals representing decoded text commands. The audio signals represent synthesized Morse Code message blocks. Dataset Structure id: A unique UUID string. audio: The audio wav bytes (16kHz Mono). text: The decoded string transcription. Dataset Level: LOW Low: Slow WPM (~12 WPM), high signal-to-noise ratio (clean), short command strings. Medium: Fast WPM (~24 WPM), background… See the full description on the dataset page: https://huggingface.co/datasets/Surpem/low-decoder.audioautomatic-speech-recognition1K<n<10K4 likes16 downloads4mo agoHugging Face14rikeshsilwalekg /43-143-phase2-appconv-ime-low_pitchaudio10K<n<100K0 likes15 downloads2y agoHugging Face15IAMCB /Laila_low_pauseaudion<1K0 likes13 downloads1y agoHugging Face16BophaAI /low_quality_khmer_speechaudion<1K0 likes12 downloads5mo agoHugging Face17Lowan /Dubeningaudion<1K0 likes8 downloads3y agoHugging Face18nguyenhieu3205xt /VietMuong-LowResourceaudio1K<n<10K0 likes8 downloads6mo agoHugging Face19Srijith-rkr /Low_Resource_Arabic_Adaptationaudio1 likes6 downloads4y agoHugging Face20lca0503 /speech_mmau_lowest_pitchaudio1K<n<10K0 likes6 downloads1y agoHugging Face21LasseRogers2111 /stt_lowrank_finetuningaudion<1K0 likes6 downloads1y agoHugging Face22Trelis /eval-whisper-small-pilotgpt-unified-all-data-lowercase-new-rewritten-20260223-2259gated Training Evaluation: whisper-small-pilotgpt-unified-all-data-lowercase-new-rewritten Evaluation results comparing base model vs fine-tuned model. Summary Model WER openai/whisper-small (base) 46.81% Trelis/whisper-small-pilotgpt-unified-all-data-lowercase-new-rewritten (fine-tuned) 34.22% Improvement: 12.59% WER reduction (lower is better) Source Data Evaluation Dataset: Trelis/pilotgpt-test-0.5s-rewritten Base Model: openai/whisper-small… See the full description on the dataset page: https://huggingface.co/datasets/Trelis/eval-whisper-small-pilotgpt-unified-all-data-lowercase-new-rewritten-20260223-2259.audion<1K0 likes4 downloads7mo agoHugging Face23Trelis /pilotgpt-unified-all-data-lowercase-new-rewrittengatedaudio1K<n<10K0 likes3 downloads7mo agoHugging Face24Trelis /pilotgpt-unified-all-data-lowercase-data-prepgatedaudio1K<n<10K0 likes2 downloads7mo agoHugging Face25Trelis /eval-whisper-small-pilotgpt-unified-all-data-lowercase-data-prep-6772-20260219-1448gated Training Evaluation: whisper-small-pilotgpt-unified-all-data-lowercase-data-prep-6772 Evaluation results comparing base model vs fine-tuned model. Summary Model WER openai/whisper-small (base) 53.69% Trelis/whisper-small-pilotgpt-unified-all-data-lowercase-data-prep-6772 (fine-tuned) 27.54% Improvement: 26.15% WER reduction (lower is better) Source Data Evaluation Dataset: Trelis/pilotgpt-test-0.5s Base Model: openai/whisper-small… See the full description on the dataset page: https://huggingface.co/datasets/Trelis/eval-whisper-small-pilotgpt-unified-all-data-lowercase-data-prep-6772-20260219-1448.audion<1K0 likes2 downloads7mo agoHugging Face26procit002 /SimpleScript_HouseNumberSpeechgenDataset_lowercasegatedaudion<1K0 likes1 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.