CoolFace
Datasetpublic

prathoshap/sushrota-sanskrit-asr-data

Su-śrotā — Sanskrit ASR Dataset Curated and consented Sanskrit speech with utterance-level transcriptions, used to train the Su-śrotā Sanskrit ASR model (finetuned IndicConformer-CTC). Focused on śāstric and recitational Sanskrit (chant and prose). Author: Prof. Prathosh A P, Indian Institute of Science, Bengaluru. Audio: 16 kHz mono WAV. Transcriptions: Devanāgarī. Splits split clips hours description train 6,438 17.4 full training set (all sources… See the full description on the dataset page: https://huggingface.co/datasets/prathoshap/sushrota-sanskrit-asr-data.

sourceHugging Facecc-by-4.0updated 1mo agoView on Hugging Face
5likes239downloads
3 commits on main
9ea73b51mo ago

Add dataset card

prathoshap
bed6e4e1mo ago

Upload dataset

prathoshap
3c9bdac1mo ago

initial commit

prathoshap