CoolFace
Datasetpublic

zuhri025/urdu-eng-merged-dataset

--- language: - ur - en license: cc-by-4.0 task_categories: - automatic-speech-recognition tags: - urdu - english - speech - audio - asr - tts size_categories: - 10K<n<100K --- # Urdu + English merged speech dataset A merged dataset with: - **Urdu rows**: duration filter + normalization + cleaning - **English rows**: duration filter only, no text normalization ## Dataset Summary | Field | Value | |---|---| | **Samples** | 241,149 | | **Total audio** | 488.9 hours | |… See the full description on the dataset page: https://huggingface.co/datasets/zuhri025/urdu-eng-merged-dataset.

sourceHugging Faceupdated 4mo agoView on Hugging Face
0likes21downloads
settings

This repository belongs to zuhri025 on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameurdu-eng-merged-dataset
visibilitypublic
licencenot set
gatedno
ownerzuhri025
Account settings
zuhri025/urdu-eng-merged-dataset · CoolFace