CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Digital-Divide-Data /Luhya-ASR-Data-subset-642H Luhya ASR Data Subset 642H Luhya speech dataset for automatic speech recognition. audioautomatic-speech-recognition100K<n<1M1 likes8.4k downloads1mo agoHugging Face02ziggylott /tlott-digital-products T. Lott Digital Products Digital product files for T. Lott's online store. Products Audiobooks (MP3) eBooks (PDF) Software (ZIP) Cover images (PNG) Download URLs Files can be downloaded directly: https://huggingface.co/datasets/ziggylott/tlott-digital-products/resolve/main/{filepath} audion<1K0 likes5k downloads19d agoHugging Face03Digital-Divide-Data /Somali-ASR-Subset-68H Somali ASR Subset 68H Somali speech dataset for automatic speech recognition. audioautomatic-speech-recognition100K<n<1M3 likes3.9k downloads1mo agoHugging Face04Digital-Divide-Data /khmer-speech-dataset Khmer ASR Cultural Dataset 727.94 hours of manually curated speech-text pairs by native speakers in the Khmer language about Cambodian cultural topics. On average, each recording is 8 seconds. Speaker metadata (gender, age group, and origin city) is provided. Language: Khmer (khm). Source(s): Native speakers from Cambodia (5 females, 7 males). The utterances were manually generated based on topics and subtopics listed in metadata. Domain(s): Cultural domain, with a total of 61… See the full description on the dataset page: https://huggingface.co/datasets/Digital-Divide-Data/khmer-speech-dataset.audioautomatic-speech-recognition100K<n<1M26 likes3.8k downloads3mo agoHugging Face05Digital-Divide-Data /Kamba-ASR-Data-Subset-484H Kamba ASR Data Subset 484H Kamba speech dataset for automatic speech recognition. audioautomatic-speech-recognition100K<n<1M0 likes2.8k downloads1mo agoHugging Face06Digital-Divide-Data /Gusii-ASR-Data-Subset-470H Gusii ASR Data Subset 470H Gusii speech dataset for automatic speech recognition. audioautomatic-speech-recognition100K<n<1M0 likes2k downloads1mo agoHugging Face07Digital-Divide-Data /khm-asr-cultural Khmer ASR Cultural Dataset 134.6 hours manually curated speech-text pairs by native speakers in Khmer language about Cambodian cultural topics. On average, each recording is 8.54 seconds with the standard deviation of 3.37. Speaker metadata (gender, age group, and origin city) is provided. Language: Khmer (khm). Source(s): Native speakers from Cambodia (4 females, 4 males). The utterances were manually generated based on topics and subtopics listed in metadata. Domain(s):… See the full description on the dataset page: https://huggingface.co/datasets/Digital-Divide-Data/khm-asr-cultural.audioautomatic-speech-recognition10K<n<100K9 likes1.3k downloads5mo agoHugging Face08mazkooleg /digit_mask_false_positive_cv12_rawaudio1M<n<10M0 likes619 downloads2y agoHugging Face09DigitalUmuganda /afrispeak_kinyarwanda_male_tts_datasetgatedaudio0 likes531 downloads2y agoHugging Face10mazkooleg /digit_mask_augmented_raw Dataset Card for "digit_mask_augmented_raw" More Information needed audio1M<n<10M0 likes487 downloads3y agoHugging Face11origami-digital /in-the-grooveCompiled from several different sets of songs: (ITG) In the Groove (ITG) In the Groove 2 Songs were downloaded from https://search.stepmaniaonline.net/packs/in+the+groove and are stored here for persistence. In The Groove/ITG typically refers to DDR beatmaps done with an eye towards pad play. Dataset info: https://paperswithcode.com/dataset/itg audioaudio-classificationn<1K0 likes475 downloads3y agoHugging Face12Digital-Divide-Data /Luhya-ASR-Data-subset-50haudio10K<n<100K0 likes329 downloads11mo agoHugging Face13hoangbang /speak-the-digit Speak the Digit: Spoken Digit Recognition Dataset Summary A public, viewer-ready educational challenge dataset. Host-only scoring data and hidden targets are excluded. Splits Split Examples Description train 2,400 Labeled training data test 600 Public inputs with withheld target labels or annotations Data Fields Field Type audio Audio id string label string (test sentinel: unlabeled)… See the full description on the dataset page: https://huggingface.co/datasets/hoangbang/speak-the-digit.audioaudio-classification1K<n<10K0 likes323 downloads2mo agoHugging Face14Congo-digital-service /audios-lingala-annotatees Annotated Lingala Dataset – Full Version Description This dataset gathers annotated Lingala audio data, intended for open-source automatic speech recognition (ASR) research and for fine-tuning Whisper-type models. It includes: the original audio files (viewable directly in the Hugging Face viewer) text transcriptions Mel spectrograms tokenized labels Overall statistics Metric Value Total volume 5 h 0 min 18 s Number of audio segments… See the full description on the dataset page: https://huggingface.co/datasets/Congo-digital-service/audios-lingala-annotatees.audioautomatic-speech-recognition10K<n<100K0 likes318 downloads16d agoHugging Face15Congo-digital-service /audios-lingala-annotatees-v2 Annotated Lingala Audio — canonical corpus Annotated Lingala speech for open automatic speech recognition research and for fine-tuning speech models. This release is a full reconstruction of the corpus from its source recordings and annotations. It supersedes Congo-digital-service/audios-lingala-annotatees, which is deprecated — see Relationship to the previous release below. What this dataset contains Each row is one annotated speech segment, carrying the audio… See the full description on the dataset page: https://huggingface.co/datasets/Congo-digital-service/audios-lingala-annotatees-v2.audioautomatic-speech-recognition10K<n<100K0 likes155 downloads12d agoHugging Face16Digital-Divide-Data /Luhya-ASR-Data-subset-LWaudio1K<n<10K0 likes153 downloads11mo agoHugging Face17Digital-Divide-Data /Luhya-ASR-Data-subset-CAaudio1K<n<10K0 likes139 downloads11mo agoHugging Face18DigitalUmuganda /Afrivoice_Swahili-Voice_Instruct_Formataudio100K<n<1M0 likes107 downloads11mo agoHugging Face19mteb /free-spoken-digit-datasetaudio1K<n<10K0 likes99 downloads1y agoHugging Face20Digital-Divide-Data /Luhya-ASR-Data-subset-TOaudio1K<n<10K0 likes99 downloads11mo agoHugging Face21Digital-Divide-Data /Luhya-ASR-Data-subsetaudio1K<n<10K0 likes98 downloads11mo agoHugging Face22DigitalUmuganda /afrispeak_kinyarwanda_female_tts_datasetaudio0 likes84 downloads2y agoHugging Face23Beijuka /DigitalUmuganda_AfriVoice_shonaaudio10K<n<100K2 likes60 downloads2y agoHugging Face24Digital-Divide-Data /Luhya-ASR-Data-subset-LAaudio1K<n<10K0 likes58 downloads11mo agoHugging Face25madoss /faso-speech-dioula-digitsgated Faso Speech Dioula Digits A spoken-digit classification dataset in Dioula (Jula/Bambara), built from the Zenodo record 8320370 archive (DOI: 10.5281/zenodo.8320370). Each clip is one speaker saying a single digit, 1 through 4, in Dioula. Recordings vary in speaker, accent, and recording environment. Dataset Summary Split Rows Duration Per-class rows train 1,532 01:34:14.9 383 / 383 / 383 / 383 validation 168 00:10:08.2 42 / 42 / 42 / 42 The split… See the full description on the dataset page: https://huggingface.co/datasets/madoss/faso-speech-dioula-digits.audioaudio-classification1K<n<10K0 likes45 downloads1mo agoHugging Face26mohnasgbr /spoken-arabic-digits Overview This dataset contains spoken Arabic digits from 40 speakers from multiple Arab communities and local dialects. It is augmented using various techniques to increase the size of the dataset and improve its diversity. The recordings went through a number of pre-processors to evaluate and process the sound quality using Audacity app. Dataset Creation The dataset was created by collecting recordings of the digits 0-9 from 40 speakers from different Arab communities… See the full description on the dataset page: https://huggingface.co/datasets/mohnasgbr/spoken-arabic-digits.audion<1K2 likes27 downloads3y agoHugging Face27silky1708 /Free-Spoken-Digit-Datasetaudio1K<n<10K0 likes18 downloads2y agoHugging Face28herbiel /indonesia-earlymedia-digits-v1 Dataset Card for "indonesia-earlymedia-digits-v1" More Information needed audion<1K0 likes14 downloads1y agoHugging Face29mazkobot /train_valid_digit_mask_augmented_raw Dataset Card for "train_valid_digit_mask_augmented_raw" More Information needed audio10K<n<100K0 likes13 downloads3y agoHugging Face30BradyM14 /digital-love-dance Digital love dance • Reachy Mini Moves Community-contributed Marionette recordings captured on Reachy Mini. Files live under data/, each move ships as a JSON trajectory plus an optional WAV. Recorded with the Marionette web app. Reuse Cite this dataset as BradyM14/digital-love-dance. Keep the reachy_mini_community_moves tag when sharing derivatives so the community can discover related sets. audioroboticsn<1K0 likes10 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.