CoolFace
Datasetpublic

Prajwal-143/ASR-Tamil-cleaned

Dataset Card for Dataset Name Dataset Details Dataset Description This dataset is a combination of the Common Voice 16.0 and Open SLR datasets which is of 534 hours. It has been meticulously curated, normalized to a 16kHz sampling rate, and cleaned for better usability. This dataset aims to provide a comprehensive collection of speech data for various applications, including speech recognition, natural language processing, and machine learning… See the full description on the dataset page: https://huggingface.co/datasets/Prajwal-143/ASR-Tamil-cleaned.

sourceHugging Faceupdated 2y agoView on Hugging Face
3likes189downloads

Prajwal-143/ASR-Tamil-cleaned · main · files are served by the source, never re-hosted here