CoolFace
11 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01AutoAudioReasoningMOS /synthesized_audio Dataset Card for "synthesized_audio" More Information needed audio10K<n<100K0 likes377 downloads5mo agoHugging Face02AutoAudioReasoningMOS /synthesized_audio_part2audio10K<n<100K0 likes203 downloads5mo agoHugging Face03mrfakename /pdmx-multi-instrument-synthesizedSynthesized version of https://github.com/pnlong/PDMX/ (public domain) filtered to only songs with two or more instruments audio10K<n<100K1 likes156 downloads2y agoHugging Face04openmusic /pdmx-multi-instrument-synthesizedA synthesized subset of the PDMX dataset containing only MIDI files with multiple instruments. Captions are auto-generated with Qwen2 Audio. Audio is licensed under CC0 (from PDMX). Captions are licensed under CC-BY. audio10K<n<100K0 likes151 downloads2y agoHugging Face05ghanaopenai /ghana-twi-synthesized-speech Ghana Twi & Code-Switching Synthesized Speech Dataset Synthesized text-to-speech audio for Twi and English-Twi code-switching sentences. How this dataset was built 1. Source text The sentences come from two sources, combined and deduplicated: A sentence-subset sampled from ghananlpcommunity/pristine-twi-english (greedy set-cover over 3,000 articles so that every word appearing in the corpus is present in at least one selected sentence). The… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/ghana-twi-synthesized-speech.audiotext-to-speech10K<n<100K1 likes111 downloads11d agoHugging Face06ghananlpcommunity /ghana-twi-synthesized-speech Ghana Twi & Code-Switching Synthesized Speech Dataset Synthesized text-to-speech audio for Twi and English-Twi code-switching sentences. How this dataset was built 1. Source text The sentences come from two sources, combined and deduplicated: A sentence-subset sampled from ghananlpcommunity/pristine-twi-english (greedy set-cover over 3,000 articles so that every word appearing in the corpus is present in at least one selected sentence). The… See the full description on the dataset page: https://huggingface.co/datasets/ghananlpcommunity/ghana-twi-synthesized-speech.audiotext-to-speech10K<n<100K0 likes67 downloads7d agoHugging Face07DL-Project /hatespeech_synthesized_datasetaudio10K<n<100K1 likes57 downloads2y agoHugging Face08mrfakename /pdmx-multi-instrument-synthesized-miniaudio1K<n<10K0 likes28 downloads2y agoHugging Face09txya900619 /name_synthesizedaudio10K<n<100K0 likes16 downloads1y agoHugging Face10crosspad /asr_tts_synthesizedaudio1K<n<10K0 likes10 downloads6mo agoHugging Face11ReopenAI /Thai_synthesized_audio已思考若干秒 Use common Thai vocabulary from Kaikki to generate example sentences that simulate real-life scenarios with https://huggingface.co/google/gemma-4-31B-it, then use the https://huggingface.co/k2-fsa/OmniVoice model for TTS. Then use Qwen3-ASR to ASR. audio100K<n<1M1 likes6 downloads4mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.