CoolFace
Datasetpublicgated

humair025/UrduMegaSpeech

UrduMegaSpeech-1M Dataset Summary UrduMegaSpeech-1M is a large-scale Urdu-English parallel speech corpus designed for automatic speech recognition (ASR), text-to-speech (TTS), and speech translation tasks. This dataset contains high-quality audio recordings paired with Urdu transcriptions and English source text, along with quality metrics for each sample. Dataset Composition Language: Urdu (transcriptions), English (source text) Total Samples:… See the full description on the dataset page: https://huggingface.co/datasets/humair025/UrduMegaSpeech.

sourceHugging Facecc-by-4.0updated 10mo agoView on Hugging Face
10likes23downloads

humair025/UrduMegaSpeech · main · files are served by the source, never re-hosted here

This repository is gated. The listing is public, but downloading a file means accepting the publisher’s terms at Hugging Face first — the links above take you there rather than around it.