CoolFace
4 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01humair025 /Urdu-ONYX-WAV-realgatedaudio100K<n<1M1 likes23 downloads11mo agoHugging Face02humair025 /Urdu-ONYX-WAVgated Urdu-ONYX-WAV Urdu-ONYX-WAV is a high-quality Urdu Text-to-Speech (TTS) dataset consisting of audio recordings and corresponding transcripts. This dataset has been specifically prepared for training TTS models and conducting research in Urdu speech synthesis. 📊 Dataset Structure This dataset is distributed across multiple parts due to size constraints: Main repository: Base dataset with initial samples part2: Additional 2.56 GB of audio data (6 Arrow files) part3:… See the full description on the dataset page: https://huggingface.co/datasets/humair025/Urdu-ONYX-WAV.audiotext-to-speech100K<n<1M0 likes21 downloads11mo agoHugging Face03humair025 /Urdu-ONYX-WAV-kanade-V2gated Urdu-ONYX-WAV-kanade-Annotated-V2 Version 2.0 - Artifact-Free Edition 🎉 Overview This is an improved version of the Urdu-ONYX-WAV dataset, tokenized with the Kanade neural codec and optimized for artifact-free audio decoding. This dataset contains 143,627 samples of high-quality Urdu speech with comprehensive linguistic and acoustic annotations, totaling ~244 hours (~10 days) of continuous audio. Key Features 🎯 Large-Scale: 143K+ samples, 244+ hours of… See the full description on the dataset page: https://huggingface.co/datasets/humair025/Urdu-ONYX-WAV-kanade-V2.tabulartext-to-speech100K<n<1M0 likes4 downloads8mo agoHugging Face04Humair332 /Urdu-ONYX-WAV-Annotedgatedaudio10K<n<100K0 likes2 downloads8mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.