CoolFace
Datasetpublic

parambharat/bengali_asr_corpus

The corpus contains roughly 500 hours of audio and transcripts in Bangla language. The transcripts have beed de-duplicated using exact match deduplication and audio has be converted to 16000 samples

sourceHugging Facecc-by-4.0updated 3y agoView on Hugging Face
0likes36downloads

parambharat/bengali_asr_corpus · main · files are served by the source, never re-hosted here