prathoshap/sushrota-sanskrit-asr-data
Su-śrotā — Sanskrit ASR Dataset Curated and consented Sanskrit speech with utterance-level transcriptions, used to train the Su-śrotā Sanskrit ASR model (finetuned IndicConformer-CTC). Focused on śāstric and recitational Sanskrit (chant and prose). Author: Prof. Prathosh A P, Indian Institute of Science, Bengaluru. Audio: 16 kHz mono WAV. Transcriptions: Devanāgarī. Splits split clips hours description train 6,438 17.4 full training set (all sources… See the full description on the dataset page: https://huggingface.co/datasets/prathoshap/sushrota-sanskrit-asr-data.
5284
