datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
synthesized_audio
Dataset Card for "synthesized_audio"
More Information needed
synthesized_audio_part2pdmx-multi-instrument-synthesizedSynthesized version of https://github.com/pnlong/PDMX/ (public domain) filtered to only songs with two or more instruments
pdmx-multi-instrument-synthesizedA synthesized subset of the PDMX dataset containing only MIDI files with multiple instruments.
Captions are auto-generated with Qwen2 Audio.
Audio is licensed under CC0 (from PDMX). Captions are licensed under CC-BY.
ghana-twi-synthesized-speech
Ghana Twi & Code-Switching Synthesized Speech Dataset
Synthesized text-to-speech audio for Twi and English-Twi code-switching sentences.
How this dataset was built
1. Source text
The sentences come from two sources, combined and deduplicated:
A sentence-subset sampled from
ghananlpcommunity/pristine-twi-english
(greedy set-cover over 3,000 articles so that every word appearing in the corpus is present in at
least one selected sentence).
The… See the full description on the dataset page: https://huggingface.co/datasets/ghanaopenai/ghana-twi-synthesized-speech.ghana-twi-synthesized-speech
Ghana Twi & Code-Switching Synthesized Speech Dataset
Synthesized text-to-speech audio for Twi and English-Twi code-switching sentences.
How this dataset was built
1. Source text
The sentences come from two sources, combined and deduplicated:
A sentence-subset sampled from
ghananlpcommunity/pristine-twi-english
(greedy set-cover over 3,000 articles so that every word appearing in the corpus is present in at
least one selected sentence).
The… See the full description on the dataset page: https://huggingface.co/datasets/ghananlpcommunity/ghana-twi-synthesized-speech.hatespeech_synthesized_datasetpdmx-multi-instrument-synthesized-mininame_synthesizedasr_tts_synthesizedThai_synthesized_audio已思考若干秒
Use common Thai vocabulary from Kaikki to generate example sentences that simulate real-life scenarios with https://huggingface.co/google/gemma-4-31B-it, then use the https://huggingface.co/k2-fsa/OmniVoice model for TTS. Then use Qwen3-ASR to ASR.
