CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01m-hamza-mughal /beat2-additional-annotations BEAT2 Official Release + Additional Annotations This is a fork of H-Liu1997/BEAT2 that adds annotations contributed by the RAG-Gesture (CVPR 2025) and MIBURI (CVPR 2026) projects. The base BEAT2-English data (motion, audio, TextGrids, semantic labels, pretrained motion-autoencoder weights) is inherited verbatim from upstream; the additional annotations from RAG-Gesture and MIBURI are pushed on top. Citations If you use only the original BEAT2 dataset, please cite… See the full description on the dataset page: https://huggingface.co/datasets/m-hamza-mughal/beat2-additional-annotations.audio1K<n<10K0 likes2.6k downloads3mo agoHugging Face02Hamed744 /mp3audion<1K0 likes2.5k downloads8mo agoHugging Face03hammoualiyoucef20 /quran-audioaudion<1K0 likes601 downloads12d agoHugging Face04hamsaai /Recorrected_Classification_Data_filtered_trainaudio10K<n<100K0 likes256 downloads3mo agoHugging Face05hamaada /dataset_ashaudio10K<n<100K0 likes222 downloads1y agoHugging Face06hamedtu /eeg-motor-imagery-dataaudion<1K0 likes177 downloads1y agoHugging Face07hamsaai /Recorrected_Classification_Data_filtered_train_22audio10K<n<100K0 likes136 downloads3mo agoHugging Face08Hamozwa /RepeatAudio Dataset Card for RepeatAudio Audio datasets referenced in Class-Agnostic Audio Repetition Counting. Contains two main sections: RS, RSN and RVN: Synthetic datasets containing varying levels of noise. Uniformly 10 seconds long, with 0-8 repetition events contained in each sample. Clocks, Heartbeats and Dolphins: Real-world derived samples with variable length across mechanical, ecological and medical domains. Relevant code can be found in this repo. Dataset Sources… See the full description on the dataset page: https://huggingface.co/datasets/Hamozwa/RepeatAudio.audio1K<n<10K1 likes109 downloads4mo agoHugging Face09hamayawayuri /common-voice-26-en-audio English — English (en) This datasheet is for cv-corpus-26.0-2026-06-12 of the Mozilla Common Voice Scripted Speech dataset for English [English - en]. The dataset contains 2583051 clips representing 3780.95 hours of recorded speech (2784.88 hours validated) from 100172 speakers, recorded from a text corpus of 1,721,897 sentences. Language English is a West Germanic language with origins in England. There are an estimated 1.5 billion English speakers, making it the… See the full description on the dataset page: https://huggingface.co/datasets/hamayawayuri/common-voice-26-en-audio.audio1K<n<10K0 likes105 downloads21d agoHugging Face10hamezksm /machine_noise_dataset Machine Sound Doctor Dataset Labeled audio clips of running machinery, used to train the classifier behind Machine Sound Doctor — a phone-based predictive maintenance tool: call in, hold your phone near a running machine, get an SMS back diagnosing the sound. Classes Class Clips Description normal 33 Machine running normally bearing_fault 33 Bearing fault sound signature belt_slip 33 Belt slip sound signature 99 clips total. Most are 16-bit PCM… See the full description on the dataset page: https://huggingface.co/datasets/hamezksm/machine_noise_dataset.audioaudio-classificationn<1K0 likes97 downloads23d agoHugging Face11hamzasibous /darija-stt-dataset Dataset Card for "darija-stt-dataset" More Information needed audio1K<n<10K0 likes59 downloads1y agoHugging Face12hamza11111 /persian-tts-dataset-maleaudio10K<n<100K2 likes52 downloads1y agoHugging Face13hamsaai /Recorrected_Classification_Data_filtered_syr_validatedaudio1K<n<10K0 likes42 downloads2mo agoHugging Face14hamsaai /Recorrected_Classification_Data_filtered_train2audio1K<n<10K0 likes40 downloads3mo agoHugging Face15AnanOmri /hamza-belloumi-tunisian-tts Hamza Belloumi Tunisian TTS Dataset A Tunisian Arabic speech dataset for TTS model training. audiotext-to-speech1K<n<10K0 likes39 downloads6mo agoHugging Face16HamdanXI /uclass_clipped_labeled Dataset Card for "uclass_clipped_labeled" More Information needed audio1K<n<10K0 likes27 downloads2y agoHugging Face17hamees /asr-taskaudio1K<n<10K0 likes27 downloads2y agoHugging Face18hamsaai /Recorrected_Classification_Data_samplesaudion<1K0 likes27 downloads2mo agoHugging Face19hamsaai /Test_Data_filtered_samplesaudion<1K0 likes23 downloads3mo agoHugging Face20hamedfrogh /StethoBench StethoBench StethoBench is a comprehensive benchmark for cardiopulmonary auscultation, comprising 77,027 instruction–response pairs synthesized from 16,125 labeled recordings across 11 public datasets. It is the training and evaluation benchmark for StethoLM, published in the Transactions on Machine Learning Research (TMLR). Dataset Description StethoBench was constructed by synthesizing instruction–response pairs from existing labeled cardiopulmonary audio datasets… See the full description on the dataset page: https://huggingface.co/datasets/hamedfrogh/StethoBench.textaudio-classification10K<n<100K0 likes20 downloads6mo agoHugging Face21hamza11111 /jalandhary_asr_enhancedaudio10K<n<100K0 likes16 downloads1y agoHugging Face22HamdanXI /fb_labeled_v4audio1K<n<10K0 likes15 downloads2y agoHugging Face23hamzabennz /8dretna_daridjaaudio1K<n<10K0 likes14 downloads2y agoHugging Face24hamzabouajila /msa_law_asr_testaudion<1K0 likes14 downloads11mo agoHugging Face25hamzahanif /memoni-audio Memoni Audio Dataset Cleaned, voice-activity-detected speech segments for Arabic/regional dialect ASR. Usage from datasets import load_dataset ds = load_dataset("hamzahanif/memoni-audio", split="train", streaming=True) sample = next(iter(ds)) print(sample["audio"]) print(sample["unique_id"]) print(sample["channel_name"]) audioautomatic-speech-recognition10K<n<100K0 likes13 downloads4mo agoHugging Face26hamees /asr-task-testaudio1K<n<10K0 likes12 downloads2y agoHugging Face27hamzasibous /daraudio10K<n<100K0 likes12 downloads1y agoHugging Face28HamdanXI /myst_single_and_long_utt_balancedgatedaudio10K<n<100K0 likes12 downloads10mo agoHugging Face29HamdanXI /fb_labeledaudio1K<n<10K0 likes11 downloads2y agoHugging Face30HamdanXI /fb_labeled_v6_w2v2audio1K<n<10K0 likes11 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.