CoolFace
12 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01VoiceNet /emolia-thinking Emolia-Thinking — a VoiceNet-annotated, balanced subset of Emolia Emolia-Thinking is a richly annotated speech dataset created for the VoiceNet project. It takes a balanced subset of the Emolia corpus — balanced across speaker-embedding clusters and emotion-embedding clusters so that speakers, voices and emotional states are evenly represented rather than dominated by the most common cases — and annotates every clip along the full VoiceNet Extended voice-performance taxonomy… See the full description on the dataset page: https://huggingface.co/datasets/VoiceNet/emolia-thinking.audioaudio-classification100K<n<1M0 likes4.2k downloads2mo agoHugging Face02laion /emolia-thinking-balanced-buckets Emolia-Thinking — Balanced Per-Dimension Bucket Subset A balanced, per-dimension bucket subset of VoiceNet/emolia-thinking, derived from that dataset's zero-shot VoiceNet-dimension labels. For every VoiceNet voice/prosody/timbre/style dimension, this subset draws a roughly equal number of clips from each ordinal bucket (0–6), so that downstream training / probing sees a balanced distribution along each axis instead of the strongly skewed natural distribution. How… See the full description on the dataset page: https://huggingface.co/datasets/laion/emolia-thinking-balanced-buckets.audioaudio-classification100K<n<1M0 likes990 downloads2mo agoHugging Face03ThinkVoiceAI /thinkvoice-dataset-v3combinedgatedaudio100K<n<1M0 likes214 downloads5mo agoHugging Face04Splend1dchan /AF-Think-audiosaudio1K<n<10K0 likes110 downloads8mo agoHugging Face05Catalan258 /thinkomni_eval ThinkOmni Evaluation Dataset This repository contains the evaluation datasets for ThinkOmni, a training-free framework that lifts textual reasoning to omni-modal scenarios via guidance decoding. ThinkOmni enhances omni-modal large language models (OLLMs) with the reasoning capabilities of large reasoning models (LRMs) at decoding time, adaptively balancing perception and reasoning signals. Paper: ThinkOmni: Lifting Textual Reasoning to Omni-modal Scenarios via Guidance Decoding… See the full description on the dataset page: https://huggingface.co/datasets/Catalan258/thinkomni_eval.audioimage-text-to-text1K<n<10K0 likes58 downloads4mo agoHugging Face06VoiceNet /majestrino-thinkingaudio1M<n<10M0 likes25 downloads5mo agoHugging Face07VoiceNet /laions-got-talent-thinkingaudio1M<n<10M0 likes16 downloads5mo agoHugging Face08ThinkingMachinesDataScience /Ratchada-STTgated RATCHADA-STT Dataset Overview The dataset includes recordings from earnings calls of publicly traded companies in Thailand. Each audio file is accompanied by a transcription and metadata such as company name, reporting period, and other relevant details. Dataset Info Total Duration: Train: 26507.87 seconds (~ 7.36 hours) Test: 8376.28 seconds (~ 2.33 hours) File Count: Train: 10912 files Test: 2804 files Dataset Structure The dataset consists… See the full description on the dataset page: https://huggingface.co/datasets/ThinkingMachinesDataScience/Ratchada-STT.audioautomatic-speech-recognition10K<n<100K1 likes7 downloads2y agoHugging Face09EmirUcar /Thinkvoice_combined_v2gatedaudio100K<n<1M0 likes3 downloads6mo agoHugging Face10buxiaoqimizi /thinker-talker-datasetsgatedaudio0 likes3 downloads4mo agoHugging Face11VoiceNet /multilingual-in-the-wild-thinkingaudio100K<n<1M0 likes3 downloads5mo agoHugging Face12EmirUcar /Thinkvoice_combined_v1gatedaudio100K<n<1M0 likes1 downloads7mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.