CoolFace
Datasetpublic

transitionGap/ASR_Marathi_Sentences

Marathi Sentence-Level ASR Dataset 📌 Overview This dataset contains sentence-level Marathi speech segments aligned with transcripts. The dataset was created by extracting subtitle timestamps (SRV3 format) from Marathi YouTube content and segmenting the corresponding audio using precise time alignment. Each sample contains: A WAV audio file (sentence-level) The corresponding Marathi transcript text This dataset is suitable for: Whisper fine-tuning Wav2Vec2 CTC… See the full description on the dataset page: https://huggingface.co/datasets/transitionGap/ASR_Marathi_Sentences.

sourceHugging Faceupdated 7mo agoView on Hugging Face
0likes37downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face