CoolFace
Datasetpublicgated

ragunath-ravi/TamilVoiceCorpus

Tamil Conversational ASR Dataset This is a dataset for Automatic Speech Recognition (ASR) focused on conversational Tamil, collected from various public sources on the web. Each sample is a short audio clip (averaging 10 seconds) paired with its corresponding transcription. Dataset Summary Language: Tamil (ta) Domain: Conversational speech Average Duration per Clip: ~10 seconds Format: Audio (.wav) + text Sample Rate: 16kHz recommended Total Examples: PureVox:… See the full description on the dataset page: https://huggingface.co/datasets/ragunath-ravi/TamilVoiceCorpus.

sourceHugging Faceapache-2.0updated 1y agoView on Hugging Face
5likes36downloads

ragunath-ravi/TamilVoiceCorpus · main · files are served by the source, never re-hosted here

This repository is gated. The listing is public, but downloading a file means accepting the publisher’s terms at Hugging Face first — the links above take you there rather than around it.