CoolFace
Datasetpublic

prathoshap/sushrota-sanskrit-asr-data

Su-śrotā — Sanskrit ASR Dataset Curated and consented Sanskrit speech with utterance-level transcriptions, used to train the Su-śrotā Sanskrit ASR model (finetuned IndicConformer-CTC). Focused on śāstric and recitational Sanskrit (chant and prose). Author: Prof. Prathosh A P, Indian Institute of Science, Bengaluru. Audio: 16 kHz mono WAV. Transcriptions: Devanāgarī. Splits split clips hours description train 6,438 17.4 full training set (all sources… See the full description on the dataset page: https://huggingface.co/datasets/prathoshap/sushrota-sanskrit-asr-data.

sourceHugging Facecc-by-4.0updated 1mo agoView on Hugging Face
5likes284downloads

prathoshap/sushrota-sanskrit-asr-data · main · files are served by the source, never re-hosted here