CoolFace
6 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01datapointai /text-to-speech-human-preferences-315kgated Text-to-speech human preferences: 315K votes across 15 models This gated dataset contains the evaluation record behind Datapoint Audio Bench: 315,000 eligible pairwise votes comparing 15 text-to-speech models in a complete round-robin over 300 English prompts. The prompt set covers eight practical voice-agent categories, and every generated sample is included as a typed audio record. The source evaluation collected 357,651 completed responses. The published benchmark excluded… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/text-to-speech-human-preferences-315k.audiotext-to-speech100K<n<1M36 likes570 downloads21d agoHugging Face02vietkemmai /Data_voice_AI_human_scam Vietnamese Deepfake Voice Dataset Dataset Description The Vietnamese Deepfake Voice Dataset is a multimodal dataset designed for research on deepfake voice detection and scam call detection in Vietnamese. The dataset contains both authentic human speech and AI-generated speech collected from multiple speech synthesis and voice cloning systems. The dataset is intended for developing and evaluating machine learning and deep learning models for: Audio deepfake… See the full description on the dataset page: https://huggingface.co/datasets/vietkemmai/Data_voice_AI_human_scam.audioaudio-classificationn<1K1 likes51 downloads2mo agoHugging Face03datapointai /tts-human-preferences-largegated TTS Human Preferences (Large) Human preference dataset for text-to-speech (TTS) audio quality evaluation. Each row contains two TTS audio renderings of the same text prompt, along with 15 human preference annotations indicating which audio sounds more natural. This is the large (2,700-row) subset. See also: small (1,000 rows), medium (2,000 rows). Dataset Summary Metric Value Total rows 2,700 Annotations per row 15 Total annotations 40,500 Unique… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/tts-human-preferences-large.audioaudio-classification1K<n<10K7 likes28 downloads7mo agoHugging Face04datapointai /tts-human-preferences-smallgated TTS Human Preferences (Small) Human preference dataset for text-to-speech (TTS) audio quality evaluation. Each row contains two TTS audio renderings of the same text prompt, along with 15 human preference annotations indicating which audio sounds more natural. This is the small (1,000-row) subset. Larger versions will follow. Dataset Summary Metric Value Total rows 1,000 Annotations per row 15 Total annotations 15,000 Unique prompts 1,000 Audio format… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/tts-human-preferences-small.audioaudio-classification1K<n<10K9 likes13 downloads7mo agoHugging Face05datapointai /tts-human-preferences-mediumgated TTS Human Preferences (Medium) Human preference dataset for text-to-speech (TTS) audio quality evaluation. Each row contains two TTS audio renderings of the same text prompt, along with 15 human preference annotations indicating which audio sounds more natural. This is the medium (2,000-row) subset. See also: small (1,000 rows). Larger versions will follow. Dataset Summary Metric Value Total rows 2,000 Annotations per row 15 Total annotations 30,000… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/tts-human-preferences-medium.audioaudio-classification1K<n<10K7 likes11 downloads7mo agoHugging Face06datasetsANDmodels /Noise-from-HumanThis dataset includes noise from Human, such as BREATH MUNCHING audion<1K0 likes3 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.