CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01RidheshBhati /Indic-total-New-TTS-Merge Indic Total TTS Merge Merged TTS dataset with 13 Indic languages. All audio clips are >= 3.0 seconds duration. Languages assamese, bengali, english, gujarati, hindi, kannada, malayalam, marathi, nepali, odia, punjabi, tamil, telugu Columns audio: Audio data text: Transcript text duration: Duration in seconds (all >= 3.0s) language: Language name audio100K<n<1M1 likes885 downloads7mo agoHugging Face02amanuelbyte /merged_speech_datasetaudio1M<n<10M0 likes760 downloads7mo agoHugging Face03Ritwika03 /syspin_hindi_mergedaudio10K<n<100K1 likes561 downloads1y agoHugging Face04badrex /anv_data_ke_kikuyu_mergedaudio100K<n<1M0 likes410 downloads1y agoHugging Face05Ritwika03 /hindi_karya_mergedaudio100K<n<1M0 likes398 downloads1y agoHugging Face06om-ai /merged-data-250924gatedTest test Hello 1234 audio100K<n<1M0 likes373 downloads2y agoHugging Face07bagasshw /indo-merged-dataset-2 Dataset Card for "indo-merged-dataset-2" More Information needed audio10K<n<100K0 likes215 downloads1y agoHugging Face08madoss /merged-bambara-dioula-datasetaudio10K<n<100K0 likes210 downloads2mo agoHugging Face09equal-ai /merged-hindi-audio-datasetaudio10K<n<100K0 likes172 downloads8mo agoHugging Face10mshojaei77 /persian_tts_merged Merged Persian TTS Dataset Overview This dataset is a comprehensive collection of Persian speech data, merged from several high-quality sources to facilitate Text-to-Speech (TTS) research and development for the Persian language. It combines audio recordings with their corresponding transcriptions, providing a rich resource for training and evaluating Persian TTS systems. Dataset Details Language: Persian (Farsi) Total Samples: [Insert total number of samples… See the full description on the dataset page: https://huggingface.co/datasets/mshojaei77/persian_tts_merged.audio10K<n<100K6 likes158 downloads2y agoHugging Face11ntariklk /darija-merged-asraudio10K<n<100K0 likes155 downloads5mo agoHugging Face12om-ai /oct21_merged_finalgatedaudio100K<n<1M0 likes151 downloads2y agoHugging Face13SayantanJoker /Shrutilipi_Hindi_resampled_44100_merged_10audio10K<n<100K0 likes147 downloads1y agoHugging Face14ilyes25 /wjbmattingly_xhosa_merged_audio Xhosa Merged Audio This dataset was cultivated from Beijuka/xhosa_parakeet_50hr. This dataset orginally came from NCHLT isiXhosa Speech Corpus (see below). The original corpus contained audio and transcription in 3-5 word segments. This meant that the majority of the dataset was ~5 seconds long. Whisper can receive an input of 30 seconds. This meant that the dataset required substantial padding. To reduce the amount of padding, the audio segments were merged together sequentially… See the full description on the dataset page: https://huggingface.co/datasets/ilyes25/wjbmattingly_xhosa_merged_audio.audioautomatic-speech-recognition1K<n<10K0 likes137 downloads1y agoHugging Face15BrunoHays /muscat-merged-samples MUSCAT — Merged Long-Form Samples This dataset is a merged, long-form reformatting of goodpiku/muscat-eval (MUSCAT: A Multi-Device Dataset for Code-Switching ASR and Segmentation Evaluation). The original MUSCAT release stores each conversation as many short, single-language segments. Here those segments are concatenated back into one continuous recording per conversation, so each row is a single long-form code-switching audio with inline language/timing markers. The layout… See the full description on the dataset page: https://huggingface.co/datasets/BrunoHays/muscat-merged-samples.audioautomatic-speech-recognitionn<1K0 likes134 downloads18d agoHugging Face16Ankesh1234 /dysarthria-asr-mergedaudio10K<n<100K1 likes133 downloads1y agoHugging Face17Aynursusuz /Turkish-Podcast-Merge-v1audio100K<n<1M2 likes123 downloads7mo agoHugging Face18OBY632 /merged-bambara-dioula-datasetaudio10K<n<100K0 likes110 downloads6mo agoHugging Face19PThi35 /S2T_Korean_Merge_2_fixed4audio10K<n<100K0 likes106 downloads5mo agoHugging Face20SayantanJoker /Shrutilipi_Hindi_resampled_44100_merged_1audio10K<n<100K0 likes100 downloads1y agoHugging Face21PharynxAI /merged_english_accent_datasetaudio10K<n<100K0 likes100 downloads1y agoHugging Face22om-ai /241022_merged_data_fixed_distributiongatedaudio100K<n<1M0 likes94 downloads2y agoHugging Face23wjbmattingly /xhosa_merged_audio Xhosa Merged Audio This dataset was cultivated from Beijuka/xhosa_parakeet_50hr. This dataset orginally came from NCHLT isiXhosa Speech Corpus (see below). The original corpus contained audio and transcription in 3-5 word segments. This meant that the majority of the dataset was ~5 seconds long. Whisper can receive an input of 30 seconds. This meant that the dataset required substantial padding. To reduce the amount of padding, the audio segments were merged together sequentially… See the full description on the dataset page: https://huggingface.co/datasets/wjbmattingly/xhosa_merged_audio.audioautomatic-speech-recognition1K<n<10K2 likes90 downloads2y agoHugging Face24Ritwika03 /syspin_merged_with_descaudio10K<n<100K0 likes90 downloads1y agoHugging Face25KYAGABA /Merged_Luo_Datasetaudio10K<n<100K0 likes89 downloads2y agoHugging Face26chiyuanhsiao /TTS_merge-ties_ls960-testaudio1K<n<10K0 likes89 downloads1y agoHugging Face27SayantanJoker /IndicVoices_Hindi_audio_44100_mergedaudio10K<n<100K1 likes84 downloads1y agoHugging Face28chiyuanhsiao /TTS_merge-dare_ls960-testaudio1K<n<10K0 likes81 downloads1y agoHugging Face292XIth /ViMD_Dataset_Merged2026audio10K<n<100K0 likes79 downloads4mo agoHugging Face30Harbidel /tigrinya-asr-mergedgated tigrinya-asr-merged A merged Tigrinya speech-recognition dataset, combining and deduplicating: badrex/tigrinya-speech (train pool) google/WaxalNLP config tir_asr (train pool) UBC-NLP/SimbaBench_dataset config asr_test_tir (held-out benchmark test set) Processing Standardized to audio (16kHz mono) and text columns, with a source column tracking origin Unicode NFC-normalized transcripts, empty transcripts dropped Exact-duplicate transcripts removed from the train… See the full description on the dataset page: https://huggingface.co/datasets/Harbidel/tigrinya-asr-merged.audioautomatic-speech-recognition10K<n<100K0 likes74 downloads24d agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.