CoolFace
11 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01laion /emolia-3k-speaker-clusters Emolia 3K Speaker Clusters A curated set of 3,000 diverse speaker clusters derived from the TTS-AGI/emolia-hq dataset, with up to 20 representative audio samples per cluster. Overview The original emolia-hq dataset contains hundreds of thousands of speech samples with 128-dimensional WavLM speaker timbre embeddings. These were first clustered into 10,000 centroids, then intelligently pruned to 3,000 using density-aware farthest-point sampling to ensure: Outlier… See the full description on the dataset page: https://huggingface.co/datasets/laion/emolia-3k-speaker-clusters.audioaudio-classification10K<n<100K0 likes90 downloads6mo agoHugging Face02caster97 /slurp_clustered_datasetaudio100K<n<1M1 likes58 downloads2y agoHugging Face03adaadig /jenny-tts-with-text-clustersaudio10K<n<100K0 likes40 downloads1y agoHugging Face04mmn3690 /voice-gender-clustering Dataset Details VoxCelebs Dataset separated by gender (https://dagshub.com/DagsHub/audio-datasets/src/main/voice_gender_detection) Dataset Description Celebrities voice recordings separated by their gender. Dataset Sources [optional] VoxCeleb dataset (https://www.robots.ox.ac.uk/~vgg/data/voxceleb/vox2.html) \Separation (https://dagshub.com/DagsHub/audio-datasets/src/main/voice_gender_detection) audio1K<n<10K0 likes38 downloads2y agoHugging Face05caster97 /slurp_clustered_split_dataset_fold1audio100K<n<1M0 likes34 downloads1y agoHugging Face06mteb /voxpopuli-accent-clusteringaudio1K<n<10K0 likes32 downloads1y agoHugging Face07aimonbc24 /Complexly-cluster-thresh0_75-conf-thresh0_9audio10K<n<100K0 likes19 downloads1y agoHugging Face08Thanarit /Thai-Voice-Test-Clustering Thanarit/Thai-Voice Combined Thai audio dataset from multiple sources Dataset Details Total samples: 20 Total duration: 0.02 hours Language: Thai (th) Audio format: 16kHz mono WAV Volume normalization: -20dB Sources Processed 1 datasets in streaming mode Source Datasets GigaSpeech2: Large-scale multilingual speech corpus Usage from datasets import load_dataset # Load with streaming to avoid downloading everything dataset =… See the full description on the dataset page: https://huggingface.co/datasets/Thanarit/Thai-Voice-Test-Clustering.audion<1K0 likes7 downloads1y agoHugging Face09Essam174 /sada_female_cluster2audion<1K0 likes6 downloads6mo agoHugging Face10Thanarit /Thai-Voice-Test-Clustering-batch Thanarit/Thai-Voice Combined Thai audio dataset from multiple sources Dataset Details Total samples: 100 Total duration: 0.11 hours Language: Thai (th) Audio format: 16kHz mono WAV Volume normalization: -20dB Sources Processed 1 datasets in streaming mode Source Datasets GigaSpeech2: Large-scale multilingual speech corpus Usage from datasets import load_dataset # Load with streaming to avoid downloading everything dataset =… See the full description on the dataset page: https://huggingface.co/datasets/Thanarit/Thai-Voice-Test-Clustering-batch.audion<1K0 likes3 downloads1y agoHugging Face11midralab /gol-dala-clustergated GOL voice clusters — audited repair This audit covers midralab/gol-dala-cluster revision 89e1c5982086207f0a8de22cdb880e5ea52789f6. The original repository has no dataset card and stores its files under Windows-style backslash paths. Its 89-byte cluster_statistics.json ends inside the cluster_statistics object and is invalid JSON. Verified source layout voice_clusters.csv: 7,362,684 strictly parsed rows, 596 folders, 19,245 speakers, 150 cluster IDs (0–149), and… See the full description on the dataset page: https://huggingface.co/datasets/midralab/gol-dala-cluster.audioaudio-classification1M<n<10M0 likes1 downloads19h agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.