CoolFace
23 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01TigreGotico /not-wake-words-speech-en not-wake-words-speech-en Negative (non-wake-word) speech clips, used to measure false accepts for OVOS wake-word plugins. Derived from the Multilingual Spoken Words Corpus (MLCommons), which is built from Mozilla Common Voice and licensed CC-BY-4.0. This derivative keeps the same licence and attribution requirement. Produced with support from the NGI0 Commons Fund. audioaudio-classification10K<n<100K0 likes553 downloads21d agoHugging Face02michsethowusu /twi-words-speech-text-parallel-400k Twi Words Speech-Text Parallel Dataset Dataset Description This dataset contains 413463 parallel speech-text pairs for Twi (Akan), a language spoken primarily in Ghana. The dataset consists of audio recordings paired with their corresponding text transcriptions, making it suitable for automatic speech recognition (ASR) and text-to-speech (TTS) tasks. Dataset Summary Language: Twi (Akan) - tw Task: Speech Recognition, Text-to-Speech Size: 413463 audio files >… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/twi-words-speech-text-parallel-400k.audioautomatic-speech-recognition100K<n<1M1 likes251 downloads1y agoHugging Face03Buraaq /quran-md-words Quran-MD - Word Level This dataset is part of the complete Quran-MD dataset available here: Complete Quran-MD Paper Quran-MD: A Fine-Grained Multimodal Dataset of the Quran (Accepted at: 5th Muslims in ML Workshop co-located with NeurIPS 2025) 📄 Paper Link: quran-md-paper Abstract We present Quran-MD, a comprehensive multimodal dataset of the Qur’an that integrates textual, linguistic, and audio dimensions at the verse and word levels. For each verse (ayah)… See the full description on the dataset page: https://huggingface.co/datasets/Buraaq/quran-md-words.audio10K<n<100K13 likes230 downloads8mo agoHugging Face04michsethowusu /swahili-words-speech-text-parallel Swahili Words Speech-Text Parallel Dataset Dataset Description This dataset contains 411048 parallel speech-text pairs for Swahili, a widely spoken language in East Africa. The dataset consists of audio recordings paired with corresponding text transcriptions, making it suitable for automatic speech recognition (ASR) and text-to-speech (TTS) tasks. Dataset Summary Language: Swahili - sw Task: Speech Recognition, Text-to-Speech Size: 411048 audio files > 1KB… See the full description on the dataset page: https://huggingface.co/datasets/michsethowusu/swahili-words-speech-text-parallel.audioautomatic-speech-recognition100K<n<1M1 likes206 downloads1y agoHugging Face05AlienKevin /wordshk_cantonese_speechaudio100K<n<1M0 likes170 downloads2y agoHugging Face06farbodbij /persian-words Persian Words This is a dataset of approximately 5K words, read aloud by a variety of native speakers. The dataset has been directly redistributed from this URL. It can be used as a valuable resource for evaluating/training ASR engines or speech synthesis engines. P.S.: I'm not the original creator of this dataset, for crediting or ownership change you can contact audioautomatic-speech-recognition1K<n<10K2 likes69 downloads1y agoHugging Face07mcamara /all-words-in-english-with-pink-trombone Dataset Card for Pink Trombone English Phonetic & Landmark Dataset Repository: mcamara/all-words-in-english-with-pink-trombone Modality: Audio + time-aligned events (landmarks) + articulatory keyframes Language: English (IPA) Sampling rate: 44,100 Hz (mono) Voices: two synthetic voices — M (male) and F (female) Summary A large-scale, clean synthetic speech dataset generated with the Pink Trombone articulatory synthesizer. Every English dictionary word is… See the full description on the dataset page: https://huggingface.co/datasets/mcamara/all-words-in-english-with-pink-trombone.audio100K<n<1M0 likes68 downloads4mo agoHugging Face08TheSeriousProgrammer /spoken_words_en_ml_commons_filtered_splittextaudio-classification100K<n<1M0 likes50 downloads4y agoHugging Face09WatsonNT /wordshk_cantonese_speechaudio100K<n<1M0 likes43 downloads28d agoHugging Face10ghananlpcommunity /twi-stitched-words-asr This dataset is shared under CC BY-NC 4.0, which means you are free to use, share, and adapt it for non-commercial research and educational purposes with attribution. You can read the full license at https://creativecommons.org/licenses/by-nc/4.0/. audio10K<n<100K0 likes26 downloads3mo agoHugging Face11speech31 /PhonemeSegmentCounting_Librispeech-wordsaudio1K<n<10K0 likes19 downloads2y agoHugging Face12Ankesh1234 /torgo-dysarthria-male-words-100audion<1K0 likes13 downloads11mo agoHugging Face13arbml /Merged_Arabic_Corpus_of_Isolated_Words Dataset Card for Merged_Arabic_Corpus_of_Isolated_Words Dataset Summary [More Information Needed] Supported Tasks and Leaderboards [More Information Needed] Languages [More Information Needed] Dataset Structure Data Instances [More Information Needed] Data Fields [More Information Needed] Data Splits [More Information Needed] Dataset Creation Curation Rationale [More Information… See the full description on the dataset page: https://huggingface.co/datasets/arbml/Merged_Arabic_Corpus_of_Isolated_Words.audio1K<n<10K0 likes11 downloads2y agoHugging Face14DynamicSuperb /PhonemeSegmentCounting_Librispeech-wordsaudio1K<n<10K1 likes10 downloads2y agoHugging Face15aharalambieva /target-words-geminitts Target-Word (TW) Evaluation Set Synthetic speech clips for 101 rare drug terms, intended for evaluation only — measuring how well an ASR system recognises rare / out-of-vocabulary medical vocabulary (target-word WER / CER / recall). Each clip reads a real DailyMed sentence containing one target drug name, synthesised with Google Gemini TTS across multiple voices. This is the frozen evaluation set from the master's thesis "Audio-free lexical adaptation of Whisper's decoder"… See the full description on the dataset page: https://huggingface.co/datasets/aharalambieva/target-words-geminitts.audioautomatic-speech-recognition1K<n<10K0 likes10 downloads2mo agoHugging Face16arbml /Speech_Corpus_for_Isolated_Words Dataset Card for Speech_Corpus_for_Isolated_Words Dataset Summary [More Information Needed] Supported Tasks and Leaderboards [More Information Needed] Languages [More Information Needed] Dataset Structure Data Instances [More Information Needed] Data Fields [More Information Needed] Data Splits [More Information Needed] Dataset Creation Curation Rationale [More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/arbml/Speech_Corpus_for_Isolated_Words.audio1K<n<10K0 likes9 downloads2y agoHugging Face17deepinfinityai /50_words_datasetaudion<1K0 likes8 downloads2y agoHugging Face18melakio /quran-md-words Quran-MD - Word Level This dataset is part of the complete Quran-MD dataset available here: Complete Quran-MD Paper Quran-MD: A Fine-Grained Multimodal Dataset of the Quran (Accepted at: 5th Muslims in ML Workshop co-located with NeurIPS 2025) 📄 Paper Link: quran-md-paper Abstract We present Quran-MD, a comprehensive multimodal dataset of the Qur’an that integrates textual, linguistic, and audio dimensions at the verse and word levels. For each verse (ayah)… See the full description on the dataset page: https://huggingface.co/datasets/melakio/quran-md-words.audio10K<n<100K0 likes7 downloads7mo agoHugging Face19abuelnasr /eg-ADI-wordsaudio10K<n<100K0 likes6 downloads2y agoHugging Face20aharalambieva /target_wordsgatedaudio10K<n<100K0 likes5 downloads6mo agoHugging Face21aharalambieva /target_words_lexical_adaptationgatedaudio10K<n<100K0 likes3 downloads6mo agoHugging Face22smfreeze /50-words-stephen-fryDataset of Stephen Fry from his singing on 50 Words For Snow by kate bush. audion<1K0 likes2 downloads2y agoHugging Face23Qurat17 /kash_wordsgatedaudion<1K0 likes2 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.