CoolFace
18 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01leonhard-behr /voxpopuli-mls-de-descriptionsgated Natural Language Voice Descriptions of the VoxPopuli and MLS German Datasets German read and parliamentary speech paired with its transcript, acoustic measurements, discrete German descriptor tags, and a free-text German description of the speaker's voice and recording conditions. The dataset is intended for training description-conditioned TTS models such as Parler-TTS. The data was built as part of research work. It is a random subset of the pooled German portions of VoxPopuli… See the full description on the dataset page: https://huggingface.co/datasets/leonhard-behr/voxpopuli-mls-de-descriptions.audiotext-to-speech100K<n<1M0 likes414 downloads1d agoHugging Face02RutwikShete /hindi_dataset_stats_catagorical_description_audioaudio100K<n<1M1 likes137 downloads2y agoHugging Face03mesolitica /Cantonese-Radio-Description-Instructions Cantonese-Radio-Description-Instructions Originally from alvanlii/cantonese-radio, we use Qwen/Qwen2.5-72B-Instruct to generate description based on the transcription. how to prepare the dataset huggingface-cli download \ mesolitica/Cantonese-Radio-Description-Instructions \ --include '*.zip' \ --repo-type "dataset" \ --local-dir './' wget https://gist.githubusercontent.com/huseinzol05/2e26de4f3b29d99e993b349864ab6c10/raw/9b2251f3ff958770215d70c8d82d311f82791b78/unzip.py… See the full description on the dataset page: https://huggingface.co/datasets/mesolitica/Cantonese-Radio-Description-Instructions.audio100K<n<1M0 likes110 downloads1y agoHugging Face04atoof /fma-music-descriptions 🎵 Free Music Archive with Full Music Flamingo Descriptions A curated collection of 594 high-quality music tracks from the Free Music Archive, with complete semantic descriptions generated by NVIDIA's Music Flamingo model. ✨ What's New This dataset includes the full Music Flamingo descriptions, not just extracted tags. Each track has: 📝 Complete textual description (mood, energy, instrumentation, production, use cases) 🏷️ Extracted semantic tags 🎵 High-quality audio… See the full description on the dataset page: https://huggingface.co/datasets/atoof/fma-music-descriptions.audioaudio-classificationn<1K0 likes42 downloads7mo agoHugging Face05nadsoft /transcribed_description_samples_dialect_22audion<1K0 likes37 downloads15d agoHugging Face06nadsoft /transcribed_description_samples_dialect_2audion<1K0 likes29 downloads1mo agoHugging Face07giangtranducts /LSVSC_descriptionaudio10K<n<100K0 likes16 downloads2y agoHugging Face08Arnold145 /odia-tts-descriptionsaudio1K<n<10K0 likes15 downloads8mo agoHugging Face09nadsoft /audio_flamingo_descriptionaudion<1K0 likes15 downloads2mo agoHugging Face10nadsoft /audio_flamingo_description_4audion<1K0 likes12 downloads1mo agoHugging Face11nadsoft /transcribed_description_samplesaudion<1K0 likes9 downloads1mo agoHugging Face12anonymousforemotion /deepspeech_with_qwen_description_exp1_score_with_emotion_and_weraudio10K<n<100K0 likes7 downloads2y agoHugging Face13nadsoft /audio_flamingo_description_3audion<1K0 likes7 downloads2mo agoHugging Face14nadsoft /audio_flamingo_description_variationaudion<1K0 likes7 downloads1mo agoHugging Face15beatpulse /Video-Annotation-and-Scene-Description-Samplesaudion<1K0 likes6 downloads2mo agoHugging Face16anonymousforemotion /deepspeech_with_qwen_descriptionaudio10K<n<100K0 likes5 downloads2y agoHugging Face17nvlachak /fleurs-greek-with-descriptionsaudio1K<n<10K0 likes5 downloads2mo agoHugging Face18anonymousforemotion /deepspeech_with_qwen_description_exp1_scoregatedaudio10K<n<100K0 likes2 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.