CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01laion /emolia-thinking-balanced-buckets Emolia-Thinking — Balanced Per-Dimension Bucket Subset A balanced, per-dimension bucket subset of VoiceNet/emolia-thinking, derived from that dataset's zero-shot VoiceNet-dimension labels. For every VoiceNet voice/prosody/timbre/style dimension, this subset draws a roughly equal number of clips from each ordinal bucket (0–6), so that downstream training / probing sees a balanced distribution along each axis instead of the strongly skewed natural distribution. How… See the full description on the dataset page: https://huggingface.co/datasets/laion/emolia-thinking-balanced-buckets.audioaudio-classification100K<n<1M0 likes1k downloads2mo agoHugging Face02Milana /vctk_resampled_16k_balancedaudio10K<n<100K0 likes454 downloads2y agoHugging Face03greentechapps /everyayah_curated_1s_20s_balancedaudio10K<n<100K0 likes327 downloads1y agoHugging Face04greentechapps /everyayah_curated_1s_20s_balanced_largeaudio10K<n<100K0 likes261 downloads1y agoHugging Face05enyoukai /AudioSet-Strong-Balancedaudio10K<n<100K0 likes246 downloads10mo agoHugging Face06nixiieee /dusha_balanced Dataset Details Dataset 'Dusha' split into train, val and test. Half of original train was taken, test split in halfs for val and test, 'neutral' category was cut to make the label distribution more balanced audioaudio-classification10K<n<100K2 likes187 downloads1y agoHugging Face07greentechapps /everyayah_curated_1s_20s_balanced_mediumaudio10K<n<100K0 likes140 downloads1y agoHugging Face08greentechapps /everyayah_curated_1s_20s_balanced_tinyaudio1K<n<10K0 likes131 downloads1y agoHugging Face09MoaazTalab /ASVspoof_2021_DF_Balanced_Normalizedaudio100K<n<1M5 likes112 downloads2y agoHugging Face10MoaazTalab /ASVspoof_2021_LA_Balanced_Normalizedaudio100K<n<1M2 likes82 downloads2y agoHugging Face11Praha-Labs /malayalam-emotion-balancedaudio10K<n<100K0 likes71 downloads1y agoHugging Face12FidelOdok /DOA_dataset_6_classes_balanced Dataset Card for "DOA_dataset_6_classes_balanced" More Information needed audio10K<n<100K0 likes70 downloads3y agoHugging Face13danjacobellis /audioset_opus_24kbps_balancedaudio10K<n<100K1 likes68 downloads2y agoHugging Face14ciempiess /ciempiess_balance Dataset Card for ciempiess_balance Dataset Summary The CIEMPIESS BALANCE Corpus is designed to match with the CIEMPIESS LIGHT Corpus (LDC2017S23). So, "Balance" means that if the CIEMPIESS BALANCE is combined with the CIEMPIESS LIGHT, one will get a gender balanced corpus. To appreciate this, one need to know that the CIEMPIESS LIGHT is by itself, a gender unbalanced corpus of approximately 25% of female speakers and 75% of male speakers. So, the CIEMPIESS BALANCE is a… See the full description on the dataset page: https://huggingface.co/datasets/ciempiess/ciempiess_balance.audioautomatic-speech-recognition1K<n<10K1 likes64 downloads2y agoHugging Face15ittailup /ciempiess_balanceaudio1K<n<10K0 likes55 downloads2y agoHugging Face16SeifElden2342532 /parler-tts-dataset-balancedaudio10K<n<100K1 likes43 downloads7mo agoHugging Face17laion /emolia-balanced-5M-subset emolia-balanced-5M-subset A balanced ~5.26M-sample subset of laion/Emolia (80.5M speech samples), packaged as WebDataset-compatible tar shards for direct use in training pipelines. How this subset was filtered Samples were selected if they met either of two criteria: 1. Emotion thresholds Each sample carries 40 emotion annotation scores (from the Emonet taxonomy) in its metadata. A sample qualifies for an emotion bucket if its score for that emotion meets or… See the full description on the dataset page: https://huggingface.co/datasets/laion/emolia-balanced-5M-subset.audio1M<n<10M1 likes39 downloads5mo agoHugging Face18bobboyms /phoneme-ctc-english-60h-balanced Phoneme CTC — English 60h (Balanced & Normalized) A cleaned, normalized and phoneme-balanced version of bobboyms/phoneme-ctc-english-60h-noisy, for training phoneme recognition models (CTC) — e.g. as the native acoustic model behind pronunciation-feedback systems. What's different from the source dataset Label noise removed Roman numerals dropped — eSpeak reads ii/iv/… as "Roman two/four", producing labels that don't match the audio. Non-English phonemes dropped… See the full description on the dataset page: https://huggingface.co/datasets/bobboyms/phoneme-ctc-english-60h-balanced.audioautomatic-speech-recognition10K<n<100K0 likes39 downloads3mo agoHugging Face19danjacobellis /audioset_opus_24kbps_balanced_527audio10K<n<100K0 likes34 downloads10mo agoHugging Face20TTS-AGI /balanced-emotion-dataset-majestrino-withtemporal-detailed-captions Balanced Emotion Dataset — Majestrino with Temporal Detailed Captions An emotion-balanced subset of TTS-AGI/majestrino-unified-detailed-captions-temporal. Overview Total samples: 482,594 Samples per emotion category: 12,997 Number of emotion categories: 40 Format: WebDataset (tar files with FLAC audio + JSON metadata) Number of tar files: 483 Samples per tar: ~1000 Balancing Strategy Samples were selected from the source dataset using keyword matching on… See the full description on the dataset page: https://huggingface.co/datasets/TTS-AGI/balanced-emotion-dataset-majestrino-withtemporal-detailed-captions.audioaudio-classification100K<n<1M0 likes34 downloads6mo agoHugging Face21OscarGD6 /audio_bbox_balancedaudio10K<n<100K0 likes25 downloads1y agoHugging Face22b-brave /asr_bbrave_balancedaudio1K<n<10K0 likes22 downloads2y agoHugging Face23OscarGD6 /audio-prompt-coco-balanced-extendedaudio1K<n<10K0 likes18 downloads1y agoHugging Face24MoaazTalab /ASVspoof_2021_DF1_Balanced_Normalizedaudio100K<n<1M0 likes17 downloads2y agoHugging Face25duyan2803 /wake_work_detection_balancedaudio10K<n<100K0 likes16 downloads1y agoHugging Face26parkky21 /hindi_eng_balanced_v2audio10K<n<100K0 likes16 downloads1y agoHugging Face27OscarGD6 /audio-prompt-coco-balancedaudio1K<n<10K0 likes14 downloads1y agoHugging Face28inaam1995 /cv17_su_lu_balanced Common Voice 17 -- Single / Long Utterance experiment dataset Built from fixie-ai/common_voice_17_0 (English); the original CV splits are preserved and each is bucketed into single-utterance (1 word) and long-utterance (>= 3 words). Splits: train_single, test_single, train_long, test_without_single. audio1K<n<10K0 likes14 downloads3mo agoHugging Face29sulaimank /six_lg_cv_balancedaudio10K<n<100K0 likes13 downloads2y agoHugging Face30HamdanXI /myst_single_and_long_utt_balancedgatedaudio10K<n<100K0 likes11 downloads10mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.