CoolFace
6 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01VoiceNet /emolia-thinking Emolia-Thinking — a VoiceNet-annotated, balanced subset of Emolia Emolia-Thinking is a richly annotated speech dataset created for the VoiceNet project. It takes a balanced subset of the Emolia corpus — balanced across speaker-embedding clusters and emotion-embedding clusters so that speakers, voices and emotional states are evenly represented rather than dominated by the most common cases — and annotates every clip along the full VoiceNet Extended voice-performance taxonomy… See the full description on the dataset page: https://huggingface.co/datasets/VoiceNet/emolia-thinking.audioaudio-classification100K<n<1M0 likes4.2k downloads2mo agoHugging Face02laion /emolia-thinking-balanced-buckets Emolia-Thinking — Balanced Per-Dimension Bucket Subset A balanced, per-dimension bucket subset of VoiceNet/emolia-thinking, derived from that dataset's zero-shot VoiceNet-dimension labels. For every VoiceNet voice/prosody/timbre/style dimension, this subset draws a roughly equal number of clips from each ordinal bucket (0–6), so that downstream training / probing sees a balanced distribution along each axis instead of the strongly skewed natural distribution. How… See the full description on the dataset page: https://huggingface.co/datasets/laion/emolia-thinking-balanced-buckets.audioaudio-classification100K<n<1M0 likes990 downloads2mo agoHugging Face03VoiceNet /majestrino-thinkingaudio1M<n<10M0 likes25 downloads5mo agoHugging Face04VoiceNet /laions-got-talent-thinkingaudio1M<n<10M0 likes16 downloads5mo agoHugging Face05ThinkingMachinesDataScience /Ratchada-STTgated RATCHADA-STT Dataset Overview The dataset includes recordings from earnings calls of publicly traded companies in Thailand. Each audio file is accompanied by a transcription and metadata such as company name, reporting period, and other relevant details. Dataset Info Total Duration: Train: 26507.87 seconds (~ 7.36 hours) Test: 8376.28 seconds (~ 2.33 hours) File Count: Train: 10912 files Test: 2804 files Dataset Structure The dataset consists… See the full description on the dataset page: https://huggingface.co/datasets/ThinkingMachinesDataScience/Ratchada-STT.audioautomatic-speech-recognition10K<n<100K1 likes7 downloads2y agoHugging Face06VoiceNet /multilingual-in-the-wild-thinkingaudio100K<n<1M0 likes3 downloads5mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.