CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01asdrty123 /stream-data-newaudion<1K1 likes6.7k downloads1h agoHugging Face02RidheshBhati /Codemixed_New Codemixed ASR Dataset Unified collection of code-mixed ASR datasets. audioautomatic-speech-recognition100K<n<1M2 likes5.5k downloads5mo agoHugging Face03Ayesha758 /Emotion_new_collected_datasetaudio10K<n<100K0 likes5.1k downloads4mo agoHugging Face04cm2435-new /gdpval_preference_rubricsaudion<1K0 likes1.6k downloads5mo agoHugging Face05titasmallick96 /daily-bio-newsaudion<1K0 likes1.5k downloads5h agoHugging Face06ghanaopenai /new-twi-tts-aligned-ipa new-twi-tts-aligned + IPA phonemes ghanaopendata/new-twi-tts-aligned with a machine-generated IPA phoneme transcription for every clip, produced with ghananlpcommunity/ghana-speech-phoneme-asr. Audio included — this is self-contained, no join with the source dataset needed. Contents split clips hours phoneme units mean units/clip test 16,140 17.24 663,140 41.1 train 145,258 155.21 5,945,389 40.9 Columns column type meaning… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/new-twi-tts-aligned-ipa.audioautomatic-speech-recognition100K<n<1M0 likes1.3k downloads2mo agoHugging Face07RidheshBhati /Indic-total-New-TTS-Merge Indic Total TTS Merge Merged TTS dataset with 13 Indic languages. All audio clips are >= 3.0 seconds duration. Languages assamese, bengali, english, gujarati, hindi, kannada, malayalam, marathi, nepali, odia, punjabi, tamil, telugu Columns audio: Audio data text: Transcript text duration: Duration in seconds (all >= 3.0s) language: Language name audio100K<n<1M1 likes872 downloads7mo agoHugging Face08VillaLabs /voice_ds_newaudio1M<n<10M0 likes747 downloads1y agoHugging Face09NathanRoll /global-news-radio-30s Global News Radio Dataset Multilingual news radio recordings from 51 languages across 42 countries. Recordings 51 Total audio 1500 min (25.0 h) Format MP3 16kHz mono 64kbps Parquet shards 11 Languages 51 Countries 42 Size 687 MB Languages Amharic, Arabic, Bashkir, Basque, Belarusian, Bengali, Brazilian Portuguese,Portugues Do Brasil,Português Brasil, Catalan, Croatian, Czech, Danish, Dutch, English, Estonian, Faroese, Finnish, Flemish… See the full description on the dataset page: https://huggingface.co/datasets/NathanRoll/global-news-radio-30s.audioautomatic-speech-recognitionn<1K0 likes631 downloads6mo agoHugging Face10ghanaopenai /new-twi-tts-aligned This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. Twi TTS Dataset A speech dataset of Twi (Akan) extracted from Ghanaian news media broadcasts, designed for training and fine-tuning Text-To-Speech (TTS) models. 📂 Dataset Structure Column Type Description… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/new-twi-tts-aligned.audio100K<n<1M0 likes631 downloads3mo agoHugging Face11tony-pitchblack /news-segmentationaudio0 likes540 downloads2y agoHugging Face12islomov /news_youtube_uzbek_speech_dataset News Youtube Uzbek Speech Dataset Dataset Description This dataset contains audio clips and their corresponding transcriptions in the Uzbek language with differenent dialects. The data was collected from publicly available news videos on YouTube. It is designed for training and evaluating Automatic Speech Recognition (ASR) models. Most of the content comes from the Kunuz, Qalampir YouTube channels. The data was transcribed using Gemini 2.5 Pro and was intelligently… See the full description on the dataset page: https://huggingface.co/datasets/islomov/news_youtube_uzbek_speech_dataset.audioautomatic-speech-recognition10K<n<100K10 likes497 downloads1y agoHugging Face13ghananlpcommunity /new-twi-tts-aligned-ipa new-twi-tts-aligned + IPA phonemes ghanaopendata/new-twi-tts-aligned with a machine-generated IPA phoneme transcription for every clip, produced with ghananlpcommunity/ghana-speech-phoneme-asr. Audio included — this is self-contained, no join with the source dataset needed. Contents split clips hours phoneme units mean units/clip test 16,140 17.24 663,140 41.1 train 145,258 155.21 5,945,389 40.9 Columns column type meaning… See the full description on the dataset page: https://huggingface.co/datasets/ghananlpcommunity/new-twi-tts-aligned-ipa.audioautomatic-speech-recognition100K<n<1M0 likes350 downloads2mo agoHugging Face14Serrano31 /new_trainaudio100K<n<1M0 likes255 downloads2mo agoHugging Face15meowmeowcatskill /bible-new-testamentaudion<1K0 likes251 downloads2y agoHugging Face16leungtianle /new-rl-vitaaudio100K<n<1M2 likes206 downloads9mo agoHugging Face17VillaLabs /voice_ds_new_200audio100K<n<1M0 likes204 downloads1y agoHugging Face18azimislom /news_youtube_uzbek_speech_dataset News Youtube Uzbek Speech Dataset Dataset Description This dataset contains audio clips and their corresponding transcriptions in the Uzbek language with differenent dialects. The data was collected from publicly available news videos on YouTube. It is designed for training and evaluating Automatic Speech Recognition (ASR) models. Most of the content comes from the Kunuz, Qalampir YouTube channels. The data was transcribed using Gemini 2.5 Pro and was intelligently… See the full description on the dataset page: https://huggingface.co/datasets/azimislom/news_youtube_uzbek_speech_dataset.audioautomatic-speech-recognition10K<n<100K0 likes182 downloads7mo agoHugging Face19BoburAmirov /news_youtube_uzbek_speech_dataset News Youtube Uzbek Speech Dataset Dataset Description This dataset contains audio clips and their corresponding transcriptions in the Uzbek language with differenent dialects. The data was collected from publicly available news videos on YouTube. It is designed for training and evaluating Automatic Speech Recognition (ASR) models. Most of the content comes from the Kunuz, Qalampir YouTube channels. The data was transcribed using Gemini 2.5 Pro and was intelligently… See the full description on the dataset page: https://huggingface.co/datasets/BoburAmirov/news_youtube_uzbek_speech_dataset.audioautomatic-speech-recognition10K<n<100K0 likes172 downloads10mo agoHugging Face20lilgoose777 /tibetan-speech-english-text-dataset-new-updatedaudio1K<n<10K0 likes166 downloads8mo agoHugging Face21hostbot77 /news_youtube_uzbek_speech_dataset News Youtube Uzbek Speech Dataset Dataset Description This dataset contains audio clips and their corresponding transcriptions in the Uzbek language with differenent dialects. The data was collected from publicly available news videos on YouTube. It is designed for training and evaluating Automatic Speech Recognition (ASR) models. Most of the content comes from the Kunuz, Qalampir YouTube channels. The data was transcribed using Gemini 2.5 Pro and was intelligently… See the full description on the dataset page: https://huggingface.co/datasets/hostbot77/news_youtube_uzbek_speech_dataset.audioautomatic-speech-recognition10K<n<100K0 likes162 downloads6mo agoHugging Face22Yettiesoft /voice_medical_newaudio10K<n<100K0 likes141 downloads2y agoHugging Face23NathanRoll /global-news-radio-debug Global News Radio Dataset (1 hour per station) Every news radio station from the Radio Browser API, recorded for 1 hour each. Attempted 3037 Successful 2553 Failed 484 Total audio 21 hours Parquet shards 256 Size 0.6 GB Format MP3 16kHz mono 64kbps Usage from datasets import load_dataset ds = load_dataset("NathanRoll/global-news-radio-debug", streaming=True) for sample in ds["train"]: print(sample["station_name"], sample["language"]… See the full description on the dataset page: https://huggingface.co/datasets/NathanRoll/global-news-radio-debug.audioautomatic-speech-recognition1K<n<10K0 likes127 downloads6mo agoHugging Face24NathanRoll /global-news-radio-fanoutaudion<1K0 likes125 downloads6mo agoHugging Face25AnonymousContinuousBench /News AnonymousContinuousBench — News A news-grounded QA benchmark built from Common Crawl News (CC-NEWS) articles crawled in September 2025. QAs are generated by Gemini 2.5 from clusters of related articles, then filtered for answerability and grounded with a retrieval-based set of supporting articles drawn from the corpus. What's inside Config Splits Size What it's for qa (default) val (1,189), test (1,415) 233 MB Evaluate QA on news, post-event corpus_large… See the full description on the dataset page: https://huggingface.co/datasets/AnonymousContinuousBench/News.audio1 likes109 downloads5mo agoHugging Face26cdactvm /kannada_new_dataaudio100K<n<1M0 likes104 downloads2y agoHugging Face27zuhri025 /munch-1-latent-NEW-parquet 🎙️ Urdu TTS Latent Dataset — munch-1-latent-NEW-parquet Pre-computed DACVAE latent representations for 51,021 Urdu utterances, ready for TTS model training. No audio decoding required at training time — load the dataset, reshape the binary blob, and train. Source Field Value Source audio Humair332/Urdu-munch-1 Codec Aratako/Semantic-DACVAE-Japanese-32dim Codec sample rate 48,000 Hz Encoder hop size 1,920 samples Latent frame rate 25.0 Hz Latent dim… See the full description on the dataset page: https://huggingface.co/datasets/zuhri025/munch-1-latent-NEW-parquet.tabulartext-to-speech10K<n<100K1 likes104 downloads5mo agoHugging Face28meandyou200175 /new1000dtsaudio0 likes100 downloads11mo agoHugging Face29ghananlpcommunity /new-twi-tts-aligned_normalised This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. audio100K<n<1M0 likes94 downloads3mo agoHugging Face30yvonne66 /BZNSYP_newaudio10K<n<100K0 likes90 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.