CoolFace
26 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01stepfun-ai /StepEval-Audio-Paralinguistic StepEval-Audio-Paralinguistic Dataset Paper: Step-Audio 2 Technical ReportCode: https://github.com/stepfun-ai/Step-Audio2Project Page: https://www.stepfun.com/docs/en/step-audio2 Overview StepEval-Audio-Paralinguistic is a speech-to-speech benchmark designed to evaluate AI models' understanding of paralinguistic information in speech across 11 distinct dimensions. The dataset contains 550 carefully curated and annotated speech samples for assessing capabilities beyond… See the full description on the dataset page: https://huggingface.co/datasets/stepfun-ai/StepEval-Audio-Paralinguistic.audion<1K12 likes282 downloads1y agoHugging Face02jacobrrak /ace-step-multilingual ACE-Step Multilingual Prompt Language Dataset Dataset accompanying: "Does Prompt Language Affect AI-Generated Music? A Multilingual Evaluation of ACE-Step 1.5 Turbo" Dataset description 250 instrumental tracks generated using ACE-Step 1.5 Turbo. 5 prompt languages 10 musical prompt families 5 matched random seeds 30 seconds per track 48 kHz identical generation settings across language conditions Languages English Swedish Spanish German French… See the full description on the dataset page: https://huggingface.co/datasets/jacobrrak/ace-step-multilingual.audiotext-to-audion<1K1 likes258 downloads24d agoHugging Face03TashaSkyUp /audio-quality-dataset-nfe4-30-step2 Audio Quality Dataset: NFE 4-30 Step 2 Overview This dataset publishes synthetic speech artifacts and derived spectrograms used for repo-local audio-quality experiments. At a glance: 2800 synthetic runs 200 short English prompt sentences 14 NFE settings: 4, 6, 8, ..., 30 fixed seed 1024 Each row represents one synthetic run and includes: prompt text raw synthetic WAV processed synthetic WAV spectrogram PNG NFE value procedural weak label Here, NFE means the number of… See the full description on the dataset page: https://huggingface.co/datasets/TashaSkyUp/audio-quality-dataset-nfe4-30-step2.audio1K<n<10K0 likes84 downloads6mo agoHugging Face04svjack /StepAudio2_Lu_Yin_BaiShiXi_TTSaudion<1K0 likes17 downloads1y agoHugging Face05vietnhat /mapalo-stepfun-metadata-real-2audion<1K0 likes16 downloads10mo agoHugging Face06StephaneBah /AfroRadVoice-FRgated AfroRadVoice-FR Dataset Description AfroRadVoice-FR is a French speech dataset composed of radiology report recordings, designed to support research in Automatic Speech Recognition (ASR) for African-accented French in medical contexts. The dataset combines real recordings, synthetic speech, and augmented audio to address data scarcity and improve acoustic diversity in a specialized domain. Motivation Current ASR systems show strong performance in… See the full description on the dataset page: https://huggingface.co/datasets/StephaneBah/AfroRadVoice-FR.audioautomatic-speech-recognitionn<1K1 likes16 downloads3mo agoHugging Face07Stephen-Lee /FormulaEval_datasets FormulaEval Datasets This repository provides the official datasets for FormulaEval, a benchmark for evaluating scientific formula vocalization in large speech language models toward accessible learning. Included Subsets The dataset repository contains three subsets: Subset Domain Language Physics700 Physics formulas and equations Chinese & English (bilingual) ChemEquation Chemical equations and formulas Chinese & English (bilingual) MixMath Mathematical… See the full description on the dataset page: https://huggingface.co/datasets/Stephen-Lee/FormulaEval_datasets.audio1K<n<10K2 likes15 downloads6mo agoHugging Face08sw-voice /swamiji-stepaudio-editx-refsaudion<1K0 likes14 downloads2mo agoHugging Face09StephaneBah /shopwithvoice-yoruba-benchmarkaudion<1K1 likes14 downloads2mo agoHugging Face10vietnhat /mapalo-stepfun-metadata2audion<1K0 likes13 downloads10mo agoHugging Face11vietnhat /dee-stepfun-metadata-real-2audion<1K0 likes11 downloads10mo agoHugging Face12vietnhat /trevor-stepfun-metadata2-real-v1audion<1K0 likes10 downloads10mo agoHugging Face13Kongezi /step_chemaudion<1K1 likes8 downloads1y agoHugging Face14vietnhat /mapalo-stepfun-metadata-real-1audion<1K0 likes8 downloads10mo agoHugging Face15yufan /voice_Stephen_Fryaudio1K<n<10K0 likes6 downloads1y agoHugging Face16aangry-mouse /stepik_ml_ruaudio1K<n<10K1 likes5 downloads2y agoHugging Face17PossiblyGeb /stephenfryaudio1K<n<10K0 likes5 downloads2y agoHugging Face18beatbox1200 /stephannieaudion<1K0 likes4 downloads2y agoHugging Face19svjack /Dont_be_your_lover_StepAudio_Spoiled_CuteBoy_Spoken_Datasetaudion<1K0 likes4 downloads1y agoHugging Face20smfreeze /50-words-stephen-fryDataset of Stephen Fry from his singing on 50 Words For Snow by kate bush. audion<1K0 likes3 downloads2y agoHugging Face21captainfr4nk /StepperMotorSoundsWithLabelsaudion<1K0 likes3 downloads2y agoHugging Face22mmarron14 /CV17_16kHz_es_PROC50-STEPSgated Dataset Card for Dataset Name This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): [More Information Needed] License: [More Information Needed] Dataset Sources [optional] Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/mmarron14/CV17_16kHz_es_PROC50-STEPS.audioautomatic-speech-recognition10K<n<100K0 likes3 downloads1y agoHugging Face23svjack /genshin_StepAudio2_Lu_Yin_answer_Mavuika_audio_samplesaudion<1K0 likes3 downloads1y agoHugging Face24Isario111 /StephenSalvatoreaudion<1K0 likes2 downloads2y agoHugging Face25svjack /Origin_StepAudio_Spoiled_CuteBoy_Spoken_Datasetaudion<1K0 likes2 downloads1y agoHugging Face26Nhat1106 /final-stepaudio1K<n<10K0 likes1 downloads5mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.