CoolFace
Datasetpublic

sanchit-gandhi/common_voice_16_1_hi_pseudo_labelled

Common Voice 16.1 Hindi Pseudo-Labelled This is the Common Voice 16.1 Hindi split pseudo-labelled using the Whisper large-v3 model, according to the instructions detailed in the Distil-Whisper repository. To reproduce this pseudo-labelling run, follow the instructions detailed here.

sourceHugging Faceupdated 3y agoView on Hugging Face
0likes52downloads
6 commits on main
38497113y ago

Update README.md

sanchit-gandhi
92551a73y ago

Upload dataset

sanchit-gandhi
542cd293y ago

Saving final transcriptions for split test

Sanchit Gandhi
f552b8c3y ago

Saving final transcriptions for split validation

Sanchit Gandhi
9ea6e813y ago

Saving final transcriptions for split train

Sanchit Gandhi
e548f833y ago

initial commit

sanchit-gandhi