CoolFace
Datasetpublic

transitionGap/ASR_Marathi_Sentences

Marathi Sentence-Level ASR Dataset 📌 Overview This dataset contains sentence-level Marathi speech segments aligned with transcripts. The dataset was created by extracting subtitle timestamps (SRV3 format) from Marathi YouTube content and segmenting the corresponding audio using precise time alignment. Each sample contains: A WAV audio file (sentence-level) The corresponding Marathi transcript text This dataset is suitable for: Whisper fine-tuning Wav2Vec2 CTC… See the full description on the dataset page: https://huggingface.co/datasets/transitionGap/ASR_Marathi_Sentences.

sourceHugging Faceupdated 7mo agoView on Hugging Face
0likes37downloads
1 commits on main
af3d6a17mo ago

Duplicate from Prasad12344321/ASR_Marathi_Sentences

Prasad12344321