CoolFace
Datasetpublic

simpra/xh-tts-vixsd

isiXhosa TTS clips (ViXSD, segmented) 3,861 clips, 22,050 Hz mono, 8 speakers, cut from long-form recordings by CTC forced alignment. Derived from ViXSD (Vuk'uzenzele isiXhosa Speech Dataset) by Lelapa AI / Way With Words, under the Esethu License — see https://huggingface.co/datasets/lelapa/Vukuzenzele_isiXhosa_Speech_Dataset_ViXSD Pipeline vixsd_extract.py — parquet to mono 22,050 Hz. Source is heterogeneous: rates 16k/22.05k/44.1k/48k/96k, depths 16/24/32, PCM… See the full description on the dataset page: https://huggingface.co/datasets/simpra/xh-tts-vixsd.

sourceHugging Faceotherupdated 15d agoView on Hugging Face
0likes340downloads
settings

This repository belongs to simpra on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namexh-tts-vixsd
visibilitypublic
licenceother
gatedno
ownersimpra
Account settings
simpra/xh-tts-vixsd · CoolFace