CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01UniDataPro /american-speech-recognition-dataset American Speech Dataset for recognition task Dataset comprises 1,136 hours of telephone dialogues in American, collected from 1,416 native speakers across various topics and domains, achieving an impressive 95% Sentence Accuracy Rate. It is designed for research in automatic speech recognition (ASR) systems. By utilizing this dataset, researchers and developers can advance their understanding and capabilities in natural language processing (NLP), speech recognition, and machine… See the full description on the dataset page: https://huggingface.co/datasets/UniDataPro/american-speech-recognition-dataset.textn<1K5 likes137 downloads1mo agoHugging Face02UniDataPro /russian-speech-recognition-dataset Russian Speech Dataset for recognition task Dataset comprises 338 hours of telephone dialogues in Russian, collected from 460 native speakers across various topics and domains, with an impressive 98% Word Accuracy Rate. It is designed for research in speech recognition, focusing on various recognition models, primarily aimed at meeting the requirements for automatic speech recognition (ASR) systems. By utilizing this dataset, researchers and developers can advance their… See the full description on the dataset page: https://huggingface.co/datasets/UniDataPro/russian-speech-recognition-dataset.textn<1K7 likes94 downloads1mo agoHugging Face03UniDataPro /vietnamese-speech-recognition Vietnamese Speech Dataset Dataset comprises 10+ hours of telephone dialogues in Vietnamese, collected from 20 native speakers across various topics and domains. It is designed for research in speech recognition, focusing on various recognition models, primarily aimed at meeting the requirements for automatic speech recognition (ASR) systems. By utilizing this dataset, researchers and developers can advance their understanding and capabilities in transcribing audio, and natural… See the full description on the dataset page: https://huggingface.co/datasets/UniDataPro/vietnamese-speech-recognition.audioautomatic-speech-recognitionn<1K4 likes87 downloads1mo agoHugging Face04UniDataPro /slovenian-speech-recognition Slovenian Speech Dataset Dataset comprises 10+ hours of audio recordings featuring 20+ speakers engaged in telephone dialogues in the Slovenian language. It contains speech data designed for training robust language models and automatic speech recognition systems in real-world conversational scenarios. By utilizing this dataset, researchers and developers can advance their understanding and capabilities in natural language processing (NLP), speech recognition, and machine… See the full description on the dataset page: https://huggingface.co/datasets/UniDataPro/slovenian-speech-recognition.audioautomatic-speech-recognitionn<1K3 likes86 downloads1mo agoHugging Face05IbrahimSalah /The_Arabic_News_speech_Corpus_Dataset Arabic News Speech Corpus Dataset This dataset is an Arabic speech corpus that supports the development of syllable-based Arabic speech recognition using Wav2Vec-2 architecture and a 5-gram language model. It consists of Modern Standard Arabic (MSA) syllables extracted from TV news broadcasts, annotated with diacritics. Dataset Details Dataset Description This corpus contains 15 hours of WAV audio recordings transcribed into diacritized Modern Standard Arabic… See the full description on the dataset page: https://huggingface.co/datasets/IbrahimSalah/The_Arabic_News_speech_Corpus_Dataset.audioautomatic-speech-recognition1K<n<10K6 likes83 downloads2y agoHugging Face06Speech-data /Chinese-Speech-Dataset 🎧 Chinese (Simplified) Speech Dataset The Chinese (Simplified) speech dataset is a high-quality speech audio dataset developed to support scalable AI and machine learning solutions with diverse and structured audio data. It contains 105 hours of speech data across 700 audio files, delivered in MP3 and WAV formats, with a total size of 229 MB. This well-balanced audio dataset provides reliable voice data, featuring 54% female and 46% male speakers, with age distribution ranging from… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Chinese-Speech-Dataset.audioautomatic-speech-recognitionn<1K0 likes63 downloads6mo agoHugging Face07UniDataPro /japanese-speech-recognition-dataset Japanese Speech Dataset for recognition task Dataset comprises 10+ hours of telephone dialogues in Japanese, collected from 10 native speakers across various topics and domains. It is designed for research in speech recognition, focusing on various recognition models, primarily aimed at meeting the requirements for automatic speech recognition (ASR) systems. By utilizing this dataset, researchers and developers can advance their understanding and capabilities in automatic speech… See the full description on the dataset page: https://huggingface.co/datasets/UniDataPro/japanese-speech-recognition-dataset.audion<1K1 likes60 downloads1mo agoHugging Face08maristombayeva /lost-in-speechgated Lost in Speech A trilingual benchmark for reference-free classification of synthetically introduced factual and contextual alterations in English, Russian, and Kazakh. It contains 12,013 samples derived from news articles, with text, synthesized speech, and ASR transcript representations used in the study. Altered samples are LLM-generated rewrites with a controlled alteration type—contradiction, fabrication, or context inconsistency—and severity level—mild, moderate, or severe.… See the full description on the dataset page: https://huggingface.co/datasets/maristombayeva/lost-in-speech.audiotext-classification1K<n<10K0 likes49 downloads2d agoHugging Face09pervasiveaidataresearchlab2025 /VN-SpeechMix_Datasetgated VN-SpeechMix: A Large-Scale Multi-Dialect Vietnamese Speech Mixture Dataset VN-SpeechMix is a large-scale, multi-dialect Vietnamese speech mixture dataset for two-speaker speech separation research. It is built from the ViMD corpus (Van Dinh et al., EMNLP 2024) using a loudness-aware mixing pipeline (LUFS normalization + two-stage anti-clipping) and a dialect-aware pairing strategy across Vietnam's three macro-dialect regions (North / Central / South). 26,000 two-speaker… See the full description on the dataset page: https://huggingface.co/datasets/pervasiveaidataresearchlab2025/VN-SpeechMix_Dataset.tabularaudio-to-audio10K<n<100K0 likes47 downloads21d agoHugging Face10Speech-data /Korean-Speech-Dataset 🎧 Korean Speech Dataset The Korean Speech Dataset is a large-scale speech audio dataset designed to provide high-quality and structured audio data for advanced AI and machine learning systems. It includes 192 hours of audio data across 628 files, delivered in MP3 and WAV formats, with a total size of 447 MB. This well-balanced audio dataset ensures diverse and representative voice data, with 52% female and 48% male speakers, and an age distribution ranging from 18 to 50+ years. The… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Korean-Speech-Dataset.audioautomatic-speech-recognitionn<1K0 likes46 downloads6mo agoHugging Face11FatimahEmadEldin /Yemeni-Speech-Emotion-Dataset YSED — Yemeni Speech Emotion Dataset (audio-classification repackaging) A clean repackaging of YSED with a metadata.csv and stratified train/validation/test splits, for emotion classification on Yemeni Arabic. Original dataset: Derhem, S., AL-Mekhlafi, E., AL-Majmar, N. A., & AL-Makhlafi, M. (2025). YSED: Yemeni Speech Emotion Dataset. Data in Brief. DOI: 10.1016/j.dib.2025.112233. Zenodo: https://zenodo.org/records/15227219. What's in here 1432 audio clips across… See the full description on the dataset page: https://huggingface.co/datasets/FatimahEmadEldin/Yemeni-Speech-Emotion-Dataset.audioaudio-classification1K<n<10K1 likes44 downloads5mo agoHugging Face12ofc-its-phyla /amharic-speech-dataset-2026 Amharic Speech Dataset 2026 Overview This dataset contains Amharic speech recordings collected using the Leyu Platform for the Leyu Platform Competition 2026. Language Amharic (am) Dialect Standard Addis Ababa Amharic Speaker Information Number of Speakers: 1 Speaker IDs: SPK001 Audio Format Format: M4A Duration: 10–60 seconds per recording Directory Structure audio/ metadata.csv… See the full description on the dataset page: https://huggingface.co/datasets/ofc-its-phyla/amharic-speech-dataset-2026.audion<1K0 likes42 downloads2mo agoHugging Face13Speech-data /russian-speech-dataset Russian Speech Dataset The Russian Speech Dataset is a structured speech audio dataset designed to deliver high-quality audio data for machine learning and AI-driven voice systems. It includes 91 hours of audio data distributed across 641 files, provided in MP3 and WAV formats with a total size of 307 MB. This well-organized audio dataset ensures balanced voice data, with 50% female and 50% male speakers, and a broad age distribution from 18 to 50+ years. The dataset language is… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/russian-speech-dataset.audioautomatic-speech-recognitionn<1K0 likes39 downloads6mo agoHugging Face14Speech-data /French-Speech-Dataset 🎧 French Speech Dataset The French Speech Dataset is a comprehensive speech audio dataset designed to deliver high-quality and diverse audio data for advanced AI and machine learning applications. It includes 198 hours of audio data across 912 files, provided in MP3 and WAV formats, with a total size of 445 MB. This well-structured audio dataset ensures balanced and representative voice data, with 51% female and 49% male speakers, and a wide age distribution from 18 to 50+ years.… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/French-Speech-Dataset.audioautomatic-speech-recognitionn<1K0 likes38 downloads6mo agoHugging Face15Speech-data /Indonesian-Speech-Dataset 🎧 Indonesian Speech Dataset The Indonesian Speech Dataset is a high-quality speech audio dataset designed to deliver structured and scalable audio data for AI-powered voice systems. It contains 162 hours of audio data across 821 files, provided in MP3 and WAV formats, with a total size of 210 MB. This well-curated audio dataset ensures balanced and representative voice data, with 51% female and 49% male speakers, and a broad age distribution from 18 to 50+ years. The dataset… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Indonesian-Speech-Dataset.audioautomatic-speech-recognitionn<1K0 likes37 downloads6mo agoHugging Face16Speech-data /arabic-speech-dataset Field Value License cc-by-nc-nd-4.0 Task Categories Automatic Speech Recognition Language Arabic (ar) Tags Arabic, Speech, Audio, Speech Recognition, Machine Learning Size Category 1K < n < 10K 🎧 Arabic Speech Dataset 📘 Overview The Arabic Speech Dataset is a high-quality speech audio dataset built for developing, training, and evaluating advanced AI voice systems. It provides 76 hours of audio data distributed across 558 files, available… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/arabic-speech-dataset.audioautomatic-speech-recognitionn<1K0 likes36 downloads6mo agoHugging Face17Speech-data /Swedish-Speech-Dataset 🎧 Swedish Speech Dataset The Swedish Speech Dataset is a high-quality speech audio dataset designed to support advanced AI and machine learning workflows with structured and diverse audio data. It contains 162 hours of voice recordings distributed across 558 files, stored in MP3 and WAV formats, with a total size of 446 MB. This carefully curated audio dataset provides rich and balanced voice data, featuring 55% female and 45% male speakers, and an age distribution ranging from 18… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Swedish-Speech-Dataset.audioautomatic-speech-recognitionn<1K1 likes36 downloads6mo agoHugging Face18Speech-data /Japanese-Speech-Dataset 🎧 Japanese Speech Dataset The Japanese Speech Dataset is a production-ready speech audio dataset designed to provide high-quality, structured audio data for AI and machine learning applications. It includes 132 hours of audio data distributed across 733 files, available in MP3 and WAV formats, with a total size of 272 MB. This well-balanced audio dataset delivers diverse voice data, with 54% female and 46% male speakers, and an age distribution spanning from 18 to 50+ years. The… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Japanese-Speech-Dataset.audioautomatic-speech-recognitionn<1K0 likes34 downloads6mo agoHugging Face19Mariya987 /DDD-Cambodia-khmer-speech-dataset-f-adt2-0002audio1K<n<10K0 likes34 downloads3mo agoHugging Face20Speech-data /Hebrew-Speech-Dataset 🎧 Hausa Speech Dataset The Hausa Speech Dataset is a structured and high-quality speech audio dataset designed to support modern AI systems that require diverse audio data and reliable voice data for multilingual model training. It contains 160 hours of recordings across 849 files, stored in MP3 and WAV formats, with a total size of 270 MB. This carefully engineered audio dataset ensures balanced representation with 48% female and 52% male speakers, covering an age range from 18 to… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Hebrew-Speech-Dataset.audioautomatic-speech-recognitionn<1K0 likes33 downloads6mo agoHugging Face21UniDataPro /korean-speech-recognition Korean Speech Dataset Dataset comprises 10+ hours of audio recordings from 20+ speakers, featuring telephone-quality speech data from native korean speakers. It provides a diverse collection of spoken language for automatic speech recognition tasks and serves as essential training data for model training in NLP and speech detection research. By utilizing this dataset, researchers and developers can advance their understanding and capabilities in automatic speech recognition… See the full description on the dataset page: https://huggingface.co/datasets/UniDataPro/korean-speech-recognition.audioautomatic-speech-recognitionn<1K1 likes32 downloads1mo agoHugging Face22Speech-data /Hungarian-Speech-Dataset 🎧 Hungarian Speech Dataset The Hungarian Speech Dataset is a high-quality speech audio dataset designed to support advanced AI systems that depend on diverse audio data and reliable voice data for multilingual model training. It comprises 169 hours of recordings across 743 files, provided in MP3 and WAV formats, with a total size of 134 MB. This structured audio dataset ensures balanced speaker representation, featuring 46% female and 54% male speakers, and an age distribution… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Hungarian-Speech-Dataset.audioautomatic-speech-recognitionn<1K0 likes32 downloads6mo agoHugging Face23Speech-data /Portuguese-Speech-Dataset 🎧 Portuguese Speech Dataset The Portuguese Speech Dataset is a large-scale speech audio dataset designed to provide structured and high-quality audio data for modern AI and machine learning systems. It contains 195 hours of recorded speech data distributed across 894 files, available in MP3 and WAV formats, with a total size of 437 MB. This carefully curated audio dataset delivers diverse and representative voice data, with a balanced speaker distribution of 52% female and 48% male… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Portuguese-Speech-Dataset.audioautomatic-speech-recognitionn<1K0 likes31 downloads6mo agoHugging Face24UniDataPro /british-english-speech-recognition-dataset British English Speech Dataset for recognition task Dataset comprises 200 hours of high-quality audio recordings featuring 310 speakers, achieving an impressive 95% Sentence Accuracy Rate. This extensive collection of speech data is designed for NLP tasks such as speech recognition, dialogue systems, and language understanding. By utilizing this dataset, developers and researchers can advance their work in automatic speech recognition and improve recognition systems. - Get the… See the full description on the dataset page: https://huggingface.co/datasets/UniDataPro/british-english-speech-recognition-dataset.textautomatic-speech-recognitionn<1K1 likes30 downloads1mo agoHugging Face25Speech-data /Irish-Speech-Dataset Irish Dataset Metadata Field Value 📜 License CC BY-NC-ND 4.0 🎯 Task Categories Automatic Speech Recognition 🌍 Language Irish (ga) 🏷️ Tags Audio, ML, Machine, Machine Learning, Speech, Speech Recognition, Irish 📦 Size Category n < 1K audioautomatic-speech-recognitionn<1K0 likes30 downloads6mo agoHugging Face26Speech-data /Latvian-Speech-Dataset Latvian Dataset Metadata Field Value 📜 License CC BY-NC-ND 4.0 🎯 Task Categories Automatic Speech Recognition 🌍 Language Latvian (la) 🏷️ Tags Audio, Speech, Speech Recognition, ML, Machine, Machine Learning, Latvian 📦 Size Category n < 1K audioautomatic-speech-recognitionn<1K0 likes29 downloads6mo agoHugging Face27Speech-data /Cebuano-Speech-Dataset 🎧 Cebuano Speech Dataset The Cebuano Speech Dataset is a high-quality speech audio dataset designed to deliver structured and diverse audio data for AI-powered voice applications. It includes 108 hours of audio data distributed across 807 files, provided in MP3 and WAV formats, with a total size of 135 MB. This well-organized audio dataset ensures balanced voice data, with 49% female and 51% male speakers, and a broad age range from 18 to 50+ years. The dataset language is Cebuano… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Cebuano-Speech-Dataset.audioautomatic-speech-recognitionn<1K1 likes28 downloads6mo agoHugging Face28Speech-data /Dutch-Speech-Dataset 🎧 Dutch Speech Dataset The Dutch Speech Dataset is a high-quality speech audio dataset designed to provide structured and diverse audio data for modern AI and machine learning applications. It includes 179 hours of audio data across 548 files, delivered in MP3 and WAV formats, with a total size of 190 MB. This well-organized audio dataset ensures balanced and representative voice data, with 51% female and 49% male speakers, and a wide age distribution from 18 to 50+ years. The… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Dutch-Speech-Dataset.audioautomatic-speech-recognitionn<1K0 likes28 downloads6mo agoHugging Face29Speech-data /Hindi-Speech-Dataset 🎧 Hindi Speech Dataset The Hindi Speech Dataset is a high-quality and structured speech audio dataset developed to support modern AI systems that rely on diverse audio data and scalable voice data. It contains 132 hours of recordings distributed across 565 files, available in MP3 and WAV formats, with a total size of 101 MB. This carefully curated audio dataset provides balanced speaker representation with 49% female and 51% male contributors, covering an age range from 18 to 50+… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Hindi-Speech-Dataset.audioautomatic-speech-recognitionn<1K0 likes28 downloads6mo agoHugging Face30Speech-data /Filipino-Tagalog-Speech-Dataset 🎧 Filipino Speech Dataset The Filipino Speech Dataset is a high-quality speech audio dataset designed to provide structured and diverse audio data for AI-driven voice technologies. It includes 75 hours of audio data across 639 files, delivered in MP3 and WAV formats, with a total size of 212 MB. This well-prepared audio dataset ensures reliable and representative voice data, with 55% male and 45% female speakers, and a wide age distribution from 18 to 50+ years. The dataset… See the full description on the dataset page: https://huggingface.co/datasets/Speech-data/Filipino-Tagalog-Speech-Dataset.audion<1K1 likes27 downloads6mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.