CoolFace
Datasetpublicgated

humairawan/Urdu-aud0

Urdu-aud0 Dataset Dataset Summary The Urdu-aud0 dataset is a high-quality collection of synthetic Urdu audio samples paired with transcripts, designed for Text-to-Speech (TTS) model training, evaluation, and fine-tuning. The dataset contains approximately 45,000 audio clips generated using OpenAI's Audio API with the "dan" voice model. Each entry includes: High-quality audio in WAV format (22,050 Hz, mono, 16-bit) Corresponding Urdu text transcripts Generation… See the full description on the dataset page: https://huggingface.co/datasets/humairawan/Urdu-aud0.

sourceHugging Faceapache-2.0updated 9mo agoView on Hugging Face
0likes7downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.