datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
balanced-audio-snippets-40x3k-DACVAEenhanced-emo-snippets-balanced-DACVAE
Enhanced Emotion Snippets — Balanced DACVAE
A balanced, emotion-bucketed subset of TTS-AGI/enhanced-audiosnippets-DACVAE,
organized by Empathic Insight Voice+ emotion and voice attribute categories.
Overview
This dataset provides up to 100 samples per magnitude bucket for each of the
40 emotion categories and 15 voice attribute dimensions scored by
Empathic Insight Voice+.
Selection Criteria
Emotion Categories (40 dimensions)
For each emotion (e.g.… See the full description on the dataset page: https://huggingface.co/datasets/TTS-AGI/enhanced-emo-snippets-balanced-DACVAE.emolia-balanced-5M-subset
emolia-balanced-5M-subset
A balanced ~5.26M-sample subset of laion/Emolia (80.5M speech samples), packaged as WebDataset-compatible tar shards for direct use in training pipelines.
How this subset was filtered
Samples were selected if they met either of two criteria:
1. Emotion thresholds
Each sample carries 40 emotion annotation scores (from the Emonet taxonomy) in its metadata. A sample qualifies for an emotion bucket if its score for that emotion meets or… See the full description on the dataset page: https://huggingface.co/datasets/laion/emolia-balanced-5M-subset.balanced-audio-score-datasets-DACVAEOCT_binary_balanced
