datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
voxpopuli-mls-de-descriptions
Natural Language Voice Descriptions of the VoxPopuli and MLS German Datasets
German read and parliamentary speech paired with its transcript, acoustic
measurements, discrete German descriptor tags, and a free-text German
description of the speaker's voice and recording conditions. The dataset is intended
for training description-conditioned TTS models such as
Parler-TTS.
The data was built as part of research work. It is a random subset of the
pooled German portions of VoxPopuli… See the full description on the dataset page: https://huggingface.co/datasets/leonhard-behr/voxpopuli-mls-de-descriptions.hindi_dataset_stats_catagorical_description_audioCantonese-Radio-Description-Instructions
Cantonese-Radio-Description-Instructions
Originally from alvanlii/cantonese-radio, we use Qwen/Qwen2.5-72B-Instruct to generate description based on the transcription.
how to prepare the dataset
huggingface-cli download \
mesolitica/Cantonese-Radio-Description-Instructions \
--include '*.zip' \
--repo-type "dataset" \
--local-dir './'
wget https://gist.githubusercontent.com/huseinzol05/2e26de4f3b29d99e993b349864ab6c10/raw/9b2251f3ff958770215d70c8d82d311f82791b78/unzip.py… See the full description on the dataset page: https://huggingface.co/datasets/mesolitica/Cantonese-Radio-Description-Instructions.fma-music-descriptions
🎵 Free Music Archive with Full Music Flamingo Descriptions
A curated collection of 594 high-quality music tracks from the Free Music Archive, with complete semantic descriptions generated by NVIDIA's Music Flamingo model.
✨ What's New
This dataset includes the full Music Flamingo descriptions, not just extracted tags. Each track has:
📝 Complete textual description (mood, energy, instrumentation, production, use cases)
🏷️ Extracted semantic tags
🎵 High-quality audio… See the full description on the dataset page: https://huggingface.co/datasets/atoof/fma-music-descriptions.transcribed_description_samples_dialect_22transcribed_description_samples_dialect_2LSVSC_descriptionodia-tts-descriptionsaudio_flamingo_descriptionaudio_flamingo_description_4transcribed_description_samplesdeepspeech_with_qwen_description_exp1_score_with_emotion_and_weraudio_flamingo_description_3audio_flamingo_description_variationVideo-Annotation-and-Scene-Description-Samplesdeepspeech_with_qwen_descriptionfleurs-greek-with-descriptionsdeepspeech_with_qwen_description_exp1_score
