CoolFace
Datasetpublicgated

kapturecx/Chaashini

Chaashini (चाशनी) Chaashini — Hindi/Urdu for sugar syrup — is a continuously growing corpus of clean, single-speaker, studio-grade Indian-language speech built for training speech models (text-to-speech, speech recognition, speech language models). Every clip in the corpus has passed a strict multi-stage quality gate; the aim is purity over volume. Total: 1,371,203 clips · 2914.03 hours · 33 languages Format: mono 24 kHz FLAC (audio column) with a verbatim transcript and rich… See the full description on the dataset page: https://huggingface.co/datasets/kapturecx/Chaashini.

sourceHugging Faceapache-2.0updated 7m agoView on Hugging Face
1likes2.3kdownloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.

kapturecx/Chaashini · CoolFace