CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Kukedlc /suno-ai-music-dataset Suno AI Music Dataset (Multi-Genre Curated) A human-curated, multi-genre audio dataset generated with Suno V5.5 (chirp-fenix), covering 100+ sub-sub-genres across electronic, hip-hop, Latin, jazz, world, rock, ambient, pop, reggae, and classical music. Each track ships with full audio (MP3), cover art, the original generation prompt, and a 32-column metadata schema designed for downstream audio-ML research. This is not a "scrape everything Suno produces" dump. It is a… See the full description on the dataset page: https://huggingface.co/datasets/Kukedlc/suno-ai-music-dataset.audioaudio-classificationn<1K29 likes2.3k downloads4mo agoHugging Face02KRAFTON /Raon-OpenTTS-Eval Raon-OpenTTS-Eval Technical Report A robustness-oriented evaluation benchmark for zero-shot text-to-speech, covering 4 acoustic regimes (Clean, Noisy, Wild, Expressive) across 12 datasets with 6,000 prompt–text pairs. Existing zero-shot TTS benchmarks typically evaluate models using prompts drawn from a single read-speech dataset, providing an incomplete view of robustness under realistic and challenging recording scenarios. Raon-OpenTTS-Eval… See the full description on the dataset page: https://huggingface.co/datasets/KRAFTON/Raon-OpenTTS-Eval.audiotext-to-speech1K<n<10K9 likes1.9k downloads4mo agoHugging Face03khaledalganem /sada2022 Dataset Card for SADA صدى Dataset Summary يعتبر توفر البيانات من أهم ممكنات تطوير نماذج ذكاء اصطناعي متفوقة إن لم يكن أهمها، ولكن لا تزال البيانات الصوتية المفتوحة وخصوصاً باللغة العربية ولهجاتها المختلفة شحيحة المصدر. ومن هذا المنطلق وحرصًا على إطلاق القيمة الكامنة للبيانات وتمكين تطوير منتجات مبنية على الذكاء الاصطناعي، قام المركز الوطني للذكاء الاصطناعي في سدايا (الهيئة الوطنية للبيانات والذكاء الاصطناعي) بالتعاون مع الهيئة السعودية للإذاعة والتلفزيون بنشر مجموعة… See the full description on the dataset page: https://huggingface.co/datasets/khaledalganem/sada2022.audio100K<n<1M4 likes187 downloads2y agoHugging Face04aranemini /central-kurdish-tts4all TTS4All Central Kurdish Speech Dataset Dataset Summary The TTS4All Central Kurdish Speech Dataset is a multi-speaker speech corpus developed for speech synthesis and speech technology research in Central Kurdish (Sorani Kurdish). The dataset was created within the TTS4All initiative during the JSALT 2025 Workshop and provides more than 35 hours of transcribed speech from three native Central Kurdish speakers. The corpus was designed to support: Text-to-Speech… See the full description on the dataset page: https://huggingface.co/datasets/aranemini/central-kurdish-tts4all.audiotext-to-speech10K<n<100K3 likes147 downloads3mo agoHugging Face05ivrit-ai /knesset-plenumsgated About This dataset is derived from raw a/v recordings and human-generated protocols of the Knesset (the Israeli house of representatives) plenums as part of the ivrit.ai project. Consider visiting the preview space for this dataset here Method Data dumps from the Knesset contain A/V recordings, alongside proprietary protocols with timestamps. We extract the audio stream, and clean up timestamp mistakes (such as backward jumps, or out-of-order timestamp artifacts). The… See the full description on the dataset page: https://huggingface.co/datasets/ivrit-ai/knesset-plenums.audioautomatic-speech-recognition1K<n<10K3 likes108 downloads10mo agoHugging Face06Sundus246 /SADA_khaledalganemsada2022_Rawdate Dataset Card for SADA صدى Dataset Summary يعتبر توفر البيانات من أهم ممكنات تطوير نماذج ذكاء اصطناعي متفوقة إن لم يكن أهمها، ولكن لا تزال البيانات الصوتية المفتوحة وخصوصاً باللغة العربية ولهجاتها المختلفة شحيحة المصدر. ومن هذا المنطلق وحرصًا على إطلاق القيمة الكامنة للبيانات وتمكين تطوير منتجات مبنية على الذكاء الاصطناعي، قام المركز الوطني للذكاء الاصطناعي في سدايا (الهيئة الوطنية للبيانات والذكاء الاصطناعي) بالتعاون مع الهيئة السعودية للإذاعة والتلفزيون بنشر… See the full description on the dataset page: https://huggingface.co/datasets/Sundus246/SADA_khaledalganemsada2022_Rawdate.audio100K<n<1M0 likes82 downloads2mo agoHugging Face07kibaraki /Shinekhen-BuryatAudio collected by Yamakoshi (Tokyo University of Foreign Studies), originally uploaded here (CC BY-SA 4.0). start_time and end_time are from the original audio clips; the audio uploaded here are already converted into per-sentence audio clips. Used in [paper] [GitHub] audioautomatic-speech-recognition1K<n<10K0 likes80 downloads1y agoHugging Face08holimon /kreyol-tts-623 license: cc-by-nc-nd-4.0 audiotext-to-speechn<1K0 likes58 downloads1y agoHugging Face09ud-nlp /human-robot-conversation-korean Human-Robot Conversation Dataset (Korean) - 660+ Hours Dataset (Korean) contains 660+ hours of audio featuring dialogues between AI and a human in German across 20,000 recordings. The dataset supports conversational AI, speech recognition, and human-robot interaction research, with short M4A audio files (up to 2 minutes) and structured metadata for model training. - Get the data Dataset characteristics: Characteristic Data Description Audio of dialogues between AI… See the full description on the dataset page: https://huggingface.co/datasets/ud-nlp/human-robot-conversation-korean.audioautomatic-speech-recognitionn<1K1 likes49 downloads6mo agoHugging Face10UniDataPro /human-robot-conversation-korean Human-Robot Dataset The dataset comprises 660+ hours of audio recordings across 20,000+ files for human-robot interactions in the Korean language. It captures authentic dialogues between humans and artificial conversational agents, specifically designed for training language models and advancing speech recognition systems. By utilizing this dataset, researchers and developers can advance their understanding and capabilities in robotic systems and conversational AI technologies.… See the full description on the dataset page: https://huggingface.co/datasets/UniDataPro/human-robot-conversation-korean.audioautomatic-speech-recognitionn<1K1 likes48 downloads1mo agoHugging Face11Speech-data /Korean-Speech-Dataset 🎧 Korean Speech Dataset The Korean Speech Dataset is a large-scale speech audio dataset designed to provide high-quality and structured audio data for advanced AI and machine learning systems. It includes 192 hours of audio data across 628 files, delivered in MP3 and WAV formats, with a total size of 447 MB. This well-balanced audio dataset ensures diverse and representative voice data, with 52% female and 48% male speakers, and an age distribution ranging from 18 to 50+ years. The… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Korean-Speech-Dataset.audioautomatic-speech-recognitionn<1K0 likes46 downloads6mo agoHugging Face12kabir5297 /OldDumpDataaudio10K<n<100K0 likes42 downloads2y agoHugging Face13UniDataPro /korean-speech-recognition Korean Speech Dataset Dataset comprises 10+ hours of audio recordings from 20+ speakers, featuring telephone-quality speech data from native korean speakers. It provides a diverse collection of spoken language for automatic speech recognition tasks and serves as essential training data for model training in NLP and speech detection research. By utilizing this dataset, researchers and developers can advance their understanding and capabilities in automatic speech recognition… See the full description on the dataset page: https://huggingface.co/datasets/UniDataPro/korean-speech-recognition.audioautomatic-speech-recognitionn<1K1 likes37 downloads1mo agoHugging Face14Mariya987 /DDD-Cambodia-khmer-speech-dataset-f-adt2-0002audio1K<n<10K0 likes28 downloads3mo agoHugging Face15dusen0528 /kws-recordings-soundailabel kws-recordings-soundailabel This repository serves as the KWS audio dataset repo for label_kws training. Layout audio/ manifests/ docs/ README.md dataset_summary.json Source of Truth manifests/full_manifest.csv manifests/kws_multiclass_manifest.csv manifests/binary_emergency_manifest.csv manifests/review_queue.csv For compatibility, the same derived files are also available under data/derived/. Current Summary rows with local audio present: 796… See the full description on the dataset page: https://huggingface.co/datasets/dusen0528/kws-recordings-soundailabel.audion<1K0 likes25 downloads6mo agoHugging Face16ud-nlp /korean-speech-recognition Korean Speech Recognition Dataset - 10+ hours Dataset comprises 10 hours of high-quality telephone audio recordings in Korean, featuring 20 native speakers. Designed for advancing speech recognition models and language processing, this extensive speech data corpus covers diverse topics and domains, making it ideal for training robust automatic speech recognition (ASR) systems. - Get the data Dataset characteristics: Characteristic Data Description Audio of… See the full description on the dataset page: https://huggingface.co/datasets/ud-nlp/korean-speech-recognition.audioautomatic-speech-recognitionn<1K0 likes20 downloads10mo agoHugging Face17archivartaunik /dzhordzh-oruel-1984-kupalautsy 1984 Metadata Author: Джордж Оруэл Title: 1984 Narrator: купалаўцы Source Group: Аўдыёкнігі Source: Notes The original audio files are preserved as-is: no conversion; no re-encoding; no filename changes inside each split folder, except removing one common top-level archive folder when present. To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders. Target maximum split size: about 250 MB. Each split… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/dzhordzh-oruel-1984-kupalautsy.audion<1K0 likes20 downloads4mo agoHugging Face18Speech-data /Kyrgyz-Speech-Datasetaudioautomatic-speech-recognitionn<1K0 likes18 downloads2mo agoHugging Face19Speech-data /Kannada-Speech-Dataset 🎧 Kannada Speech Dataset The Kannada Speech Dataset is a high-quality speech audio dataset designed to deliver structured and reliable audio data for AI and machine learning workflows. It includes 90 hours of audio data across 651 files, available in MP3 and WAV formats, with a total size of 220 MB. This well-organized audio dataset provides balanced and representative voice data, with 48% female and 52% male speakers, and an age range spanning from 18 to 50+ years. The dataset… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Kannada-Speech-Dataset.audioautomatic-speech-recognitionn<1K0 likes17 downloads6mo agoHugging Face20s479246 /dwesui-grupa-2-kulinarna G2-Polish-Culinary-ASR-Evaluation-Corpus Korpus do ewaluacji systemow ASR jezyka polskiego (domena kulinarna) stworzony w ramach warsztatow Ewaluacja Systemow Rozpoznawania Mowy (UAM WMI, edycja 2026, zespol 2). Publikowany podzbior to mowa naturalna z wideo kulinarnych YouTube (licencja CC-BY) - sluzy do badania odpornosci ASR na szum kuchenny oraz dopasowania domenowego do specjalistycznego slownictwa (zapozyczenia, miary, liczby). Pelny eksperyment ewaluacyjny zespolu… See the full description on the dataset page: https://huggingface.co/datasets/s479246/dwesui-grupa-2-kulinarna.audioautomatic-speech-recognitionn<1K0 likes14 downloads3mo agoHugging Face21archivartaunik /ianka-kupala-raskidanae-gniazdo Раскіданае гняздо Metadata Author: Янка Купала Title: Раскіданае гняздо Narrator: Source Group: Аўдыёкнігі Source: https://www.youtube.com/channel/UCbS-jD6aM11LW_KtrqFZVWQ Notes The original audio files are preserved as-is: no conversion; no re-encoding; no filename changes inside each split folder, except removing one common top-level archive folder when present. To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/ianka-kupala-raskidanae-gniazdo.audion<1K0 likes13 downloads4mo agoHugging Face22K-University-AIED /hallym_AI_OpenDataset Hallym Adult and Child Speech Dataset This dataset contains speech recordings and transcriptions collected from adult and child speakers for AI-based speech and language research. Dataset Overview Total Records: 2,714 Speakers: 49 (adult: 25, child: 24) Groups: adult, child File Format: WAV (audio) + TXT (transcription) Speaker Statistics Group Count Gender Age Range Adult 25명 남/여 50~78세 Child 24명 남/여 3~8세 Dataset Fields… See the full description on the dataset page: https://huggingface.co/datasets/K-University-AIED/hallym_AI_OpenDataset.audio1K<n<10K0 likes12 downloads7mo agoHugging Face23Speech-data /Kazakh-Speech-Dataset 🎧 Kazakh Speech Dataset The Kazakh Speech Dataset is a high-quality speech audio dataset developed to provide structured and scalable audio data for AI and machine learning applications. It includes 130 hours of audio data distributed across 672 files, delivered in MP3 and WAV formats, with a total size of 123 MB. This well-balanced audio dataset ensures diverse and representative voice data, featuring 54% female and 46% male speakers, with an age range spanning from 18 to 50+… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Kazakh-Speech-Dataset.audioautomatic-speech-recognitionn<1K0 likes11 downloads6mo agoHugging Face24uam-wmi-asr-eval-labs /2026-dwesui-g02-kulinarna DWESUI 2026 - Grupa 2 - kulinarna (PIEROGA) Robocza/archiwalna kopia zbioru ewaluacyjnego ASR zbudowanego przez studentow kursu Warsztaty z ewaluacji systemow rozpoznawania mowy (UAM WMI), edycja 2026, tryb dzienny. Zespol (atrybucja): Grupa 2 (DWESUI 2026) Zrodlo oryginalne: https://huggingface.co/datasets/s479246/dwesui-grupa-2-kulinarna Domena: kulinarna Licencja zrodla: nagrania YouTube CC-BY/CC-BY-SA + TTS Status: kopia publiczna w organizacji kursowej (zespół opublikował… See the full description on the dataset page: https://huggingface.co/datasets/uam-wmi-asr-eval-labs/2026-dwesui-g02-kulinarna.audioautomatic-speech-recognitionn<1K0 likes11 downloads1mo agoHugging Face25koichi12 /l2_small_datasetsaudion<1K0 likes8 downloads1y agoHugging Face26kazeric /DVoice-VoxLingua107-trialrun1audio1K<n<10K0 likes4 downloads2y agoHugging Face27kaarthu2003 /CommonVoice17-Cloneaudion<1K0 likes3 downloads1y agoHugging Face28archivartaunik /ivan-bunin-kazimir-stanislavavich-uladzimir-ragautsou Казімір Станіслававіч Metadata Author: Іван Бунін Title: Казімір Станіслававіч Narrator: Уладзімір Рагаўцоў Source Group: Аўдыёкнігі Source: БЛР#аўдыякніга Notes The original audio files are preserved as-is: no conversion; no re-encoding; no filename changes inside each split folder, except removing one common top-level archive folder when present. To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders.… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/ivan-bunin-kazimir-stanislavavich-uladzimir-ragautsou.audion<1K0 likes3 downloads4mo agoHugging Face29archivartaunik /ivan-bunin-klopat-uladzimir-ragautsou Клопат Metadata Author: Іван Бунін Title: Клопат Narrator: Уладзімір Рагаўцоў Source Group: Аўдыёкнігі Source: БЛР#аўдыякніга Notes The original audio files are preserved as-is: no conversion; no re-encoding; no filename changes inside each split folder, except removing one common top-level archive folder when present. To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders. Target maximum split size:… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/ivan-bunin-klopat-uladzimir-ragautsou.audion<1K0 likes3 downloads4mo agoHugging Face30Frostie08 /kreyol-tts-623 license: cc-by-nc-nd-4.0 audiotext-to-speechn<1K0 likes2 downloads7mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.