CoolFace
Datasetpublic

parambharat/kannada_asr_corpus

The corpus contains roughly 360 hours of audio and transcripts in Kannada language. The transcripts have beed de-duplicated using exact match deduplication.

sourceHugging Facecc-by-4.0updated 4y agoView on Hugging Face
0likes29downloads
5 commits on main
e4cb7694y ago

fix dataloading script issue

Bharat Ramanathan
629e2864y ago

fix broken mp3 files

Bharat Ramanathan
59261974y ago

add readme and loading script

Bharat Ramanathan
4c22fb14y ago

add data files

Bharat Ramanathan
e7ed5ff4y ago

initial commit

parambharat