CoolFace
Datasetpublicgated

cillegio/az-asr-voa-305h

Labelling field value label_origin script speech_register broadcast channel wideband-16k provenance inferred Editorial broadcast text aligned to VOA audio. Terminal punctuation at 77.4% (against 98.1% for LocalDoc) is consistent with segments cut from continuous broadcast rather than at sentence ends. Adding this to a call model made it worse. Chinar-F8 v4 mixed in 147.8 h of it and strict WER on human-transcribed calls moved 44.23% -> 56.69%. Register, not… See the full description on the dataset page: https://huggingface.co/datasets/cillegio/az-asr-voa-305h.

sourceHugging Facecc-by-nc-4.0updated 7d agoView on Hugging Face
0likes12downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.