datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
TTS-Voice-Direction-Benchmark
TTS Voice Direction Benchmark
🏆 Leaderboard | 🛠️ Evaluation Suite
TTS Voice Direction is a benchmark of 700 reference-conditioned speech
generation tasks. It evaluates whether a text-to-speech model can preserve a
reference speaker while following a natural-language direction that controls
how a new transcript is performed.
The benchmark emphasizes practical voice acting beyond basic emotion control.
It covers accent, acoustic delivery, vocal events, emotion, physiological… See the full description on the dataset page: https://huggingface.co/datasets/BreezeBlue/TTS-Voice-Direction-Benchmark.minimax-music3-communityminimax_music3_qlora_trainersimulated-binaural-speech-directivity
Simulated Binaural Speech Source Directivity Dataset
This dataset contains 96,000 simulated binaural speech recordings generated for the study of speech source directivity classification. The dataset is designed to support the binary classification of whether a speech source is oriented toward or away from a listener.
The recordings were generated under controlled acoustic and spatial conditions using an acoustic simulation workflow based on RAVEN… See the full description on the dataset page: https://huggingface.co/datasets/sguajardo799/simulated-binaural-speech-directivity.star1_direct_refusal_audiodirect_test_audio_datasetTBD-LLaMA-500M-Final-Direction-500M_correct_codebook
