zuhri025/urdu-eng-merged-dataset
--- language: - ur - en license: cc-by-4.0 task_categories: - automatic-speech-recognition tags: - urdu - english - speech - audio - asr - tts size_categories: - 10K<n<100K --- # Urdu + English merged speech dataset A merged dataset with: - **Urdu rows**: duration filter + normalization + cleaning - **English rows**: duration filter only, no text normalization ## Dataset Summary | Field | Value | |---|---| | **Samples** | 241,149 | | **Total audio** | 488.9 hours | |… See the full description on the dataset page: https://huggingface.co/datasets/zuhri025/urdu-eng-merged-dataset.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face