humairawan/Urdu-aud0
Urdu-aud0 Dataset Dataset Summary The Urdu-aud0 dataset is a high-quality collection of synthetic Urdu audio samples paired with transcripts, designed for Text-to-Speech (TTS) model training, evaluation, and fine-tuning. The dataset contains approximately 45,000 audio clips generated using OpenAI's Audio API with the "dan" voice model. Each entry includes: High-quality audio in WAV format (22,050 Hz, mono, 16-bit) Corresponding Urdu text transcripts Generation… See the full description on the dataset page: https://huggingface.co/datasets/humairawan/Urdu-aud0.
07
No card is published for this repository, or it could not be fetched from Hugging Face right now.
