CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01kohei0209 /mls_hq_urgent_track1audio100K<n<1M0 likes1.9k downloads2y agoHugging Face02krishnakalyan3 /emo_webds_2audio10K<n<100K7 likes1.6k downloads2y agoHugging Face03krishnakalyan3 /emo_parleraudio1M<n<10M2 likes1.4k downloads2y agoHugging Face04krishnakalyan3 /emo_webdsaudio10K<n<100K5 likes1.3k downloads2y agoHugging Face05k2-fsa /OpenDialog OpenDialog OpenDialog is a 6.8k hours spoken dialogue dataset, introduced in the paper ZipVoice-Dialog: Non-Autoregressive Spoken Dialogue Generation with Flow Matching. Paper: https://arxiv.org/abs/2507.09318 GitHub: https://github.com/k2-fsa/ZipVoice Project Page: https://zipvoice-dialog.github.io OpenDialog is the first large-scale (6.8k hours) open-source spoken dialogue dataset derived from in-the-wild speech data. It consists of: English data: 5074 hours Chinese data: 1759… See the full description on the dataset page: https://huggingface.co/datasets/k2-fsa/OpenDialog.audiotext-to-speech100K<n<1M24 likes550 downloads5mo agoHugging Face06KBlueLeaf /laion-coco-13m-taraudio10M<n<100M1 likes342 downloads1y agoHugging Face07kohei0209 /mls_hqaudio10M<n<100M0 likes221 downloads2y agoHugging Face08krishnakalyan3 /emo_speech_filtered_v12 second filtered emotional speech in webdataset format https://huggingface.co/datasets/EQ4You/Emotional_Speech audio10K<n<100K0 likes181 downloads2y agoHugging Face09krishnakalyan3 /vocal_bursts_taxonomy_100_clean_wdsaudio10K<n<100K0 likes112 downloads1y agoHugging Face10krishnakalyan3 /laion-audio-preview-splitaudio1M<n<10M2 likes88 downloads2y agoHugging Face11seastar105 /Emilia-YODAS-KO-filteredaudio100K<n<1M0 likes68 downloads2y agoHugging Face12freococo /khit_thit_news_voices Khit Thit News Voices In the fight for truth, these are the voices that refuse to be silenced. Khit Thit News Voices is a focused collection of 15,841 audio segments (≈14.7 hours total) from Khit Thit News, one of Myanmar's most vital and trusted independent media outlets. Founded by renowned journalist Mr. Thar Lun Zaung Htet, Khit Thit News stands as a pillar of reliable information and a primary voice for democratic forces within the country. This dataset primarily features the… See the full description on the dataset page: https://huggingface.co/datasets/freococo/khit_thit_news_voices.audioautomatic-speech-recognition10K<n<100K1 likes63 downloads1y agoHugging Face13Reord-AI /kenya-philippines-twospeaker-english-dialoguegated Kenya/Philippines English Dialogue Two-speaker dialogues in English, recorded on split tracks. Changelog Jan 2026: v1 release - vad-segmented WebRTC tracks Specs Speakers: >150; ~15 PH, remaining KE Total duration: ~65 hours Files sample rate: 48kHz Actual sample rate: TBD Language: English (PH, KE accents) Topics: day-to-day conversation Collection method The dataset is built to capture the variety in the Kenyan accent. The Philippino interviewers… See the full description on the dataset page: https://huggingface.co/datasets/Reord-AI/kenya-philippines-twospeaker-english-dialogue.audio10K<n<100K2 likes58 downloads8mo agoHugging Face14krishnakalyan3 /earsaudio10K<n<100K0 likes54 downloads2y agoHugging Face15freococo /sagaw_karen_asrThis is the first public Sagaw Karen language ASR dataset in AI history. Sagaw Karen ASR This dataset contains audio recordings and aligned metadata in the Sagaw Karen language (ISO 639-3: ksw), a major Sgaw Karenic language spoken throughout southern and eastern Myanmar. The language is sometimes also referred to as Sgaw Karen or Sakaw Karen in English transliterations. All audio segments in this dataset were sourced from publicly available news broadcasts published by PVTV… See the full description on the dataset page: https://huggingface.co/datasets/freococo/sagaw_karen_asr.audioautomatic-speech-recognition1K<n<10K0 likes38 downloads1y agoHugging Face16freococo /karenni_language_asr_audio RFA Karenni (Kayah) Language Voices This dataset contains 17 hours of audio in the Karenni (Kayah) language, sourced from news broadcasts by Radio Free Asia (RFA) Burmese. This is one of the largest publicly accessible audio resources for the Karenni language family, designed to support research in low-resource automatic speech recognition (ASR), voice activity detection, and other speech-related tasks. This dataset was created by freococo. The audio has been automatically segmented… See the full description on the dataset page: https://huggingface.co/datasets/freococo/karenni_language_asr_audio.audioautomatic-speech-recognition1K<n<10K0 likes37 downloads1y agoHugging Face17freococo /kachin_asr_audio Dataset Summary This is the first public Kachin language ASR dataset in history. Kachin ASR Audio is a collection of speech data in the Kachin (Jinghpaw) language, sourced entirely from publicly available PVTV (People’s Voice Television) broadcasts. The dataset includes narration, interviews, and spoken reports intended to support the development of automatic speech recognition (ASR) systems for rare-resource indigenous languages in Myanmar. Each audio file is paired with metadata… See the full description on the dataset page: https://huggingface.co/datasets/freococo/kachin_asr_audio.audioautomatic-speech-recognition1K<n<10K0 likes30 downloads1y agoHugging Face18kehanlu /Speech-IFEvalaudio1K<n<10K0 likes29 downloads2y agoHugging Face19freococo /western_poe_karen_asrThis is the first public Western Poe Karen language ASR dataset in AI history. Western Poe Karen ASR This dataset contains audio recordings and aligned transcriptions in the Western Poe Karen language (also known in linguistic literature as Western Pwo or Delta Pwo, ISO 639-3: pwo), a Karenic language spoken primarily in the Ayeyarwady Delta region of Myanmar. Although linguists commonly refer to this language as Western Pwo Karen, the community and this project prefer the spelling… See the full description on the dataset page: https://huggingface.co/datasets/freococo/western_poe_karen_asr.audioautomatic-speech-recognition1K<n<10K0 likes29 downloads1y agoHugging Face20leungtianle /kimi-audio-dpoaudio100K<n<1M1 likes28 downloads9mo agoHugging Face21freococo /eastern_poe_karen_asrThis is the first public Eastern Poe Karen language ASR dataset in AI history. Eastern Poe Karen ASR This dataset contains audio recordings and aligned metadata in the Eastern Poe Karen language (a regional variety of Eastern Pwo, ISO 639-3: pwo), a Karenic language spoken primarily in Mon State and Kayin State in southeastern Myanmar. While linguistically described as Eastern Pwo Karen, the community and this project prefer the term Poe as a community-endorsed spelling. All audio… See the full description on the dataset page: https://huggingface.co/datasets/freococo/eastern_poe_karen_asr.audioautomatic-speech-recognition1K<n<10K0 likes25 downloads1y agoHugging Face22ReopenAI /COIG-Kun-Aug-Audio本数据集基于https://huggingface.co/datasets/m-a-p/COIG-Kun作为种子问题,使用Qwen2.5-72B-Instruct-GPTQ-Int4继续生成更多轮次的问题。然后使用Qwen2.5-72B-Instruct-GPTQ-Int4生成问题的答案(每轮答案生成都会将之前的问题和答案当作上下文,确保当前的答案和历史相关)。见sharegpt.json文件。问题使用cosyvoice生成对应音频。audio_part1-3.tar.gz分别是三部分压缩包,分别解压后合并成一个audio文件夹,也可只下载一部分使用。 audio100K<n<1M1 likes22 downloads1y agoHugging Face23krishnakalyan3 /wds_vocal_burst_100audio10K<n<100K0 likes22 downloads1y agoHugging Face24Vyvo-Research /Emilia-YODAS-KOaudio1M<n<10M0 likes21 downloads11mo agoHugging Face25kehanlu /dynamic-superb-train-noise-reverbaudio1K<n<10K0 likes19 downloads1y agoHugging Face26cmeraki /youtube_ka_rajesh_raw_tempaudion<1K0 likes18 downloads2y agoHugging Face27cmeraki /youtube_kn_rajesh_raw_tempaudion<1K0 likes16 downloads2y agoHugging Face28Vyvo-Research /Emilia-KOaudio10K<n<100K0 likes15 downloads11mo agoHugging Face29cmeraki /youtube_ka_bookbrahma_raw_tempaudio1K<n<10K0 likes14 downloads2y agoHugging Face30kuanhuggingface /BEAF-Audioaudio1K<n<10K0 likes13 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.