CoolFace
Datasetpublic

sanchit-gandhi/common_voice_16_1_hi_pseudo_labelled

Common Voice 16.1 Hindi Pseudo-Labelled This is the Common Voice 16.1 Hindi split pseudo-labelled using the Whisper large-v3 model, according to the instructions detailed in the Distil-Whisper repository. To reproduce this pseudo-labelling run, follow the instructions detailed here.

sourceHugging Faceupdated 3y agoView on Hugging Face
0likes52downloads
Dataset Card

Common Voice 16.1 Hindi Pseudo-Labelled

This is the Common Voice 16.1 Hindi split pseudo-labelled using the Whisper large-v3 model, according to the instructions detailed in the Distil-Whisper repository. To reproduce this pseudo-labelling run, follow the instructions detailed here.