sanchit-gandhi/common_voice_16_1_hi_pseudo_labelled
Common Voice 16.1 Hindi Pseudo-Labelled This is the Common Voice 16.1 Hindi split pseudo-labelled using the Whisper large-v3 model, according to the instructions detailed in the Distil-Whisper repository. To reproduce this pseudo-labelling run, follow the instructions detailed here.
052
Update README.md
Upload dataset
Saving final transcriptions for split test
Saving final transcriptions for split validation
Saving final transcriptions for split train
initial commit
