sapinsapin/whisper-small-pld-bcl
035
whisper-small-pld-bcl
`openai/whisper-small` finetuned on `sapinsapin/pld`.
Extended run: trained to convergence on the bcl portion of PLD read speech, selected on held-out CER. WER/CER are lowercased on the held-out split; CER is the model-selection metric because Philippine-language orthography varies at the word level.
Trained with finetune_asr.py from the halohalo pipeline; the dataset adapter normalizes each corpus to (audio@16k, text, speaker_id) so corpora are swappable with a --dataset flag.
