CoolFace
Datasetpublic

leduckhai/VietMed

VietMed: A Dataset and Benchmark for Automatic Speech Recognition of Vietnamese in the Medical Domain (LREC-COLING 2024, Oral) Description: We introduced a Vietnamese speech recognition dataset in the medical domain comprising 16h of labeled medical speech, 1000h of unlabeled medical speech and 1200h of unlabeled general-domain speech. To our best knowledge, VietMed is by far the world’s largest public medical speech recognition dataset in 7 aspects: total… See the full description on the dataset page: https://huggingface.co/datasets/leduckhai/VietMed.

sourceHugging Facemitupdated 4mo agoView on Hugging Face
25likes588downloads

leduckhai/VietMed · main · files are served by the source, never re-hosted here