CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01freds0 /TAGARELA TAGARELA: A Portuguese Speech Dataset From Podcasts TAGARELA is a large-scale Portuguese speech dataset built from podcast audio and curated for speech technology research, especially Automatic Speech Recognition (ASR) and Text-to-Speech (TTS). The dataset contains more than 8,972 hours of Portuguese speech derived from the Cem Mil Podcasts collection. It includes Brazilian Portuguese and European Portuguese speech, processed through a pipeline involving audio standardization… See the full description on the dataset page: https://huggingface.co/datasets/freds0/TAGARELA.audioautomatic-speech-recognition1M<n<10M10 likes8.7k downloads2mo agoHugging Face02freds0 /BRSpeech BRSpeech BRSpeech is a single-speaker Brazilian Portuguese speech dataset extracted and curated specifically for Text-to-Speech (TTS) and voice modeling tasks. It corresponds directly to speaker 2961 from the multi-speaker BRSpeech-TTS dataset, which represents the speaker with the highest volume of recorded audio/hours in the entire corpus. Dataset Summary Language: Portuguese (pt-BR) Speaker ID: 2961 (from BRSpeech-TTS) Task: Single-speaker Text-to-Speech… See the full description on the dataset page: https://huggingface.co/datasets/freds0/BRSpeech.audiotext-to-speech10K<n<100K1 likes510 downloads28d agoHugging Face03freds0 /cml_tts_dataset_spanishaudio100K<n<1M3 likes466 downloads2y agoHugging Face04freds0 /cml_tts_dataset_frenchaudio100K<n<1M2 likes393 downloads2y agoHugging Face05freds0 /cml_tts_dataset_portugueseaudio10K<n<100K3 likes381 downloads2y agoHugging Face06freds0 /cml_tts_dataset_germanaudio100K<n<1M3 likes285 downloads2y agoHugging Face07freds0 /cml_tts_dataset_italianaudio10K<n<100K4 likes258 downloads2y agoHugging Face08freddyaboulton /common-voice-english-audioaudio1K<n<10K1 likes211 downloads1y agoHugging Face09freds0 /cml_tts_dataset_dutchaudio100K<n<1M1 likes187 downloads2y agoHugging Face10freds0 /BRSpeech-TTSaudio10K<n<100K0 likes141 downloads1y agoHugging Face11FreddyFazbear0209 /muong_voice_textaudio1K<n<10K0 likes82 downloads10mo agoHugging Face12FreddyNom /swahili-asr-zindiaudio10K<n<100K0 likes39 downloads7mo agoHugging Face13freds0 /cml_tts_dataset_polishaudio10K<n<100K1 likes30 downloads2y agoHugging Face14FreddyFazbear0209 /viet_muong_50_trimmed_samples_100ms_silent_150ms_stoptoken_refined_text_labelaudion<1K0 likes10 downloads9mo agoHugging Face15Eididkd /Freddyaudion<1K0 likes9 downloads3y agoHugging Face16PLS442 /Fredaudion<1K0 likes9 downloads2y agoHugging Face17FreddyFazbear0209 /viet_muong_100_denoised_clean_silent_segments_customaudio1K<n<10K0 likes9 downloads10mo agoHugging Face18FreddyFazbear0209 /viet_muong_50_original_1_labeled_samples_with_refined_text_labelaudion<1K0 likes7 downloads8mo agoHugging Face19Soaresuz /freddyaudion<1K0 likes6 downloads3y agoHugging Face20Obreyer /freddyaudion<1K0 likes5 downloads3y agoHugging Face21MestreSilvio /Fredericaudion<1K0 likes5 downloads2y agoHugging Face22FreddyFazbear0209 /viet_muong_100_denoised_clean_silent_segments_libraryaudio1K<n<10K0 likes5 downloads9mo agoHugging Face23FreddyFazbear0209 /viet_muong_50_1_labeled_samples_for_smoothing_testingaudion<1K0 likes5 downloads8mo agoHugging Face24FreddyFazbear0209 /viet_muong_50_1_labeled_samples_merged_0.1s_silence_0.15s_eos_15dB_thresholdaudion<1K0 likes4 downloads8mo agoHugging Face25FreddyFazbear0209 /viet_muong_50_1_labeled_samples_merged_0.1s_silence_0.15s_eos_30dB_thresholdaudion<1K0 likes4 downloads8mo agoHugging Face26Freddyjanson /sabriaudion<1K0 likes3 downloads3y agoHugging Face27TH78 /freddiekingaudion<1K0 likes3 downloads2y agoHugging Face28FreddyFazbear0209 /viet_muong_50_trimmed_samples_100ms_silent_150ms_stoptokenaudion<1K0 likes3 downloads9mo agoHugging Face29archivartaunik /frederyk-braun-arena Арэна Metadata Author: Фрэдэрык Браўн Title: Арэна Narrator: Source Group: Аўдыёкнігі Source: Notes The original audio files are preserved as-is: no conversion; no re-encoding; no filename changes inside each split folder, except removing one common top-level archive folder when present. To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders. Target maximum split size: about 250 MB. Each split folder… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/frederyk-braun-arena.audion<1K0 likes2 downloads4mo agoHugging Face30fredsteve /yunliaudion<1K0 likes1 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.