datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
aopcode-sarvinaz-work11-clean
AOPCODE Sarvinaz - Uyghur Text-to-Speech Dataset
Dataset Description
This is a high-quality, professionally curated Uyghur text-to-speech dataset provided by AOPCODE. The dataset contains 3,483 clean audio samples with corresponding Uyghur text transcriptions, specifically processed for TTS model training.
Dataset Highlights
🎯 Pure Uyghur Content: All entries contain only Uyghur text (no English letters or digits)
🔊 High Quality Audio: Professional voice… See the full description on the dataset page: https://huggingface.co/datasets/sayiwen/aopcode-sarvinaz-work11-clean.aopcode-sarvinaz-v1-clean
AOPCODE Sarvinaz V1 Clean - Uyghur Text-to-Speech Dataset
Dataset Description
This is a high-quality, professionally curated Uyghur text-to-speech dataset provided by AOPCODE. The dataset contains 22,052 clean audio samples with corresponding Uyghur text transcriptions, specifically processed for TTS model training.
Dataset Highlights
🎯 Pure Uyghur Content: All entries contain only Uyghur text (no English letters or digits)
🔊 High Quality Audio: Professional… See the full description on the dataset page: https://huggingface.co/datasets/sayiwen/aopcode-sarvinaz-v1-clean.
