CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01anonymous-user-592888 /vocalgrad VocalGrad VocalGrad is an audio benchmark for evaluating whether a model can detect the direction of gradual perceptual change in speech. This public release contains the test split only. Each example contains one audio clip and one target attribute. The task is to answer whether that attribute increases or decreases over time. Task Given an audio clip and an attribute name, predict one of two labels: increase decrease The ground-truth label is derived from the metadata… See the full description on the dataset page: https://huggingface.co/datasets/anonymous-user-592888/vocalgrad.audio10K<n<100K0 likes1.3k downloads5mo agoHugging Face02Scicom-intl /Synthetic-User-Turn-TTS Synthetic Malaysian Telco Call-Centre Speech Synthetic Malaysian call-centre customer utterances, as text and as speech. The text is fully synthetic dialogue styled after real Malaysian ISP/telco ("Unifi") call-centre recordings, containing no real customer data. The audio subsets take customer (user) turns and voice them with a voice-conversion model, keeping only clips an ASR round-trip confirms are accurate. Subsets subset rows content default 4,260… See the full description on the dataset page: https://huggingface.co/datasets/Scicom-intl/Synthetic-User-Turn-TTS.audio100K<n<1M0 likes1.1k downloads28d agoHugging Face03AnhP /Mir-1k-use-DJCM-trainingaudio1K<n<10K2 likes425 downloads1y agoHugging Face04AudioLLMs /MMAU-mini-do-not-useWARNING: The original dataset is revised and pleased refer to new data source. Please refer to: MMAU-v05.15.25: https://github.com/Sakshi113/MMAU @misc{sakshi2024mmaumassivemultitaskaudio, title={MMAU: A Massive Multi-Task Audio Understanding and Reasoning Benchmark}, author={S Sakshi and Utkarsh Tyagi and Sonal Kumar and Ashish Seth and Ramaneswaran Selvakumar and Oriol Nieto and Ramani Duraiswami and Sreyan Ghosh and Dinesh Manocha}, year={2024}, eprint={2410.19168}… See the full description on the dataset page: https://huggingface.co/datasets/AudioLLMs/MMAU-mini-do-not-use.audio1K<n<10K2 likes256 downloads1y agoHugging Face05adarshxs /voxcpm2-native-generated-audio-user-ref VoxCPM2 Native Generated Audio (User Ref) Raw audio files generated from the native VoxCPM2 path in sglang-omni using a user-provided reference clip. Contents 9 generated .wav files metadata.json with prompt text, mode, status, size, and latency Source Reference Audio Reference clip used for the reference-mode generations: https://huggingface.co/datasets/adarshxs/voxcpm2-native-test-samples/resolve/main/data/audio.wav Files ref_expressive.wav… See the full description on the dataset page: https://huggingface.co/datasets/adarshxs/voxcpm2-native-generated-audio-user-ref.audiotext-to-speechn<1K0 likes135 downloads5mo agoHugging Face06jjjiaozi /usev-vox2-preprocessedaudio10K<n<100K0 likes83 downloads11d agoHugging Face07Pullo-Africa-Protagonist /Usethisaudio10K<n<100K0 likes58 downloads10mo agoHugging Face08AlienKevin /guangzhou-daily-use-speechASR-SCCantDuSC: A Scripted Chinese Cantonese (Canton) Daily-use Speech Corpus This open-source dataset consists of 4.06 hours of transcribed Guangzhou Cantonese scripted speech focusing on daily use sentences, where 4,060 utterances contributed by ten speakers were contained. Source: https://magichub.com/datasets/guangzhou-cantonese-scripted-speech-corpus-daily-use-sentence/ audio1K<n<10K1 likes57 downloads2y agoHugging Face09User1115 /singleWordaudion<1K0 likes33 downloads3y agoHugging Face10mali6 /audio-user-studyaudion<1K0 likes28 downloads2y agoHugging Face11UsergyAI /Global-Conversational-Speechgated Global Conversational Speech Dataset 305 hours. 18 locales. Real conversations. Not scraped from YouTube. Not recorded by anonymous crowds who don't speak the language. Every conversation in this dataset traces back to verified native speakers we know by name. The [Human] Standard Most speech datasets are built the same way: scrape the internet, hire anonymous contractors, run it through automated QC, ship it. The result? Models that are confidently wrong. We… See the full description on the dataset page: https://huggingface.co/datasets/UsergyAI/Global-Conversational-Speech.audioautomatic-speech-recognition100K<n<1M2 likes24 downloads2mo agoHugging Face12User1115 /singleWordSmallaudion<1K0 likes16 downloads3y agoHugging Face13malaysia-ai /scripted-malay-daily-use-speech-corpus scripted-malay-daily-use-speech-corpus Mirror for https://magichub.com/datasets/malay-scripted-speech-corpus-daily-use-sentence/, license is Creative Commons Attribution-NonCommercial-NoDerivatives 4.0 International License audio1K<n<10K0 likes13 downloads2y agoHugging Face14kiranpantha /raw-donot-use-NepaliParliamentDSaudio1K<n<10K0 likes12 downloads1y agoHugging Face15polinaeterna /test-userLorem ipsumaudion<1K0 likes11 downloads4y agoHugging Face16userdata /malaya-speech-malay-stt-4kaudio1K<n<10K1 likes11 downloads2y agoHugging Face17vietnhat /us-erica-higgs-metadata1-v1audion<1K0 likes11 downloads1y agoHugging Face18User1115 /SingleComaudion<1K0 likes10 downloads3y agoHugging Face19userdata /merged-dataset-4kOnline-500Manualaudio1K<n<10K0 likes10 downloads2y agoHugging Face20malaysia-ai /scripted-malay-daily-use-speech-corpus-whisper-formataudio1K<n<10K0 likes10 downloads2y agoHugging Face21vietnhat /us-erica-higgs-metadata1-v2audion<1K0 likes9 downloads1y agoHugging Face22bojieli /pine-tau2-voice-gpt41-usersimaudio10K<n<100K0 likes9 downloads3mo agoHugging Face23teamsleeping /musicality_useraudion<1K0 likes9 downloads2mo agoHugging Face24hoangvanvietanh /user_5476d2c924204b6f9e38713118fdb9b2_datasetaudion<1K0 likes7 downloads3y agoHugging Face25hoangvanvietanh /user_03aa5df890b64866be4aef51a01c0a8a_datasetaudion<1K0 likes7 downloads3y agoHugging Face26userdata /malaya-speech-malay-stt-2kaudio1K<n<10K0 likes7 downloads2y agoHugging Face27huhu-user /totalaudion<1K0 likes6 downloads3y agoHugging Face28userdata /audio-labelled-ssaudion<1K0 likes6 downloads2y agoHugging Face29User1115 /SingleWordUSTSaudion<1K0 likes5 downloads3y agoHugging Face30hoangvanvietanh /user_35621758bf084337aad673e1cc332d6f_datasetaudion<1K0 likes5 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.