CoolFace
Datasetpublic

IntisarUddin/Bengali_Long_form_ASR

Bengali Long-Form ASR Dataset Dataset Summary The Bengali Long-Form ASR Dataset is a large-scale collection of long-duration Bangla speech recordings paired with verified transcripts. The dataset is designed specifically for long-form Automatic Speech Recognition (ASR) research. Key Statistics Total duration: 310.06 hours Number of recordings: 382 Average duration per recording: ~48.7 minutes Language: Bengali (bn) Audio format: WAV Sampling rate:… See the full description on the dataset page: https://huggingface.co/datasets/IntisarUddin/Bengali_Long_form_ASR.

sourceHugging Facecc-by-4.0updated 6mo agoView on Hugging Face
0likes24downloads

IntisarUddin/Bengali_Long_form_ASR · main · files are served by the source, never re-hosted here