rlabz/quantum-tts-tokenized
Swahili (swa_spk3) SNAC-Tokenized Dataset for Orpheus-TTS Fine-Tuning Dataset Summary A single-speaker Kiswahili subset, resampled and tokenized for fine-tuning Orpheus-TTS. It is derived from rlabz/swa_lug_tts by: Filtering the train and validation splits down to speaker swa_spk3 only. Resampling all audio from its original 22,050 Hz to 24,000 Hz, the sample rate required by SNAC (snac_24khz), the neural audio codec Orpheus is trained on. Encoding each clip with… See the full description on the dataset page: https://huggingface.co/datasets/rlabz/quantum-tts-tokenized.
This repository belongs to rlabz on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
