CoolFace
Datasetpublic

tensorxt/ViMedCSS

๐Ÿฉบ ViMedCSS: A Vietnamese Medical Code-Switching Speech Dataset (LREC 2026) ๐Ÿ“– Overview ViMedCSS is a Vietnamese medical speech dataset for code-switching ASR, where each utterance contains at least one non-Vietnamese (mainly English) medical term embedded in Vietnamese speech. ๐Ÿ“Š Dataset Statistics Split Statistics (from ViMedCSS-Metadata) Split # Rows Duration (hours) Avg duration (s) Total CS terms train 11,832 24.30โ€ฆ See the full description on the dataset page: https://huggingface.co/datasets/tensorxt/ViMedCSS.

sourceHugging Facecc-by-4.0updated 7mo agoView on Hugging Face
18likes538downloads
18 commits on main
b6959a17mo ago

Update README.md

tensorxt
134d7cd7mo ago

Update README.md

tensorxt
32608417mo ago

Update README.md

tensorxt
835bec67mo ago

Update README.md

tensorxt
aa0da887mo ago

Update README.md

tensorxt
97aa8727mo ago

Update README.md

tensorxt
54c97337mo ago

Update README.md

tensorxt
6ce0ade7mo ago

Update README.md

tensorxt
b85d1c07mo ago

Update README.md

tensorxt
a41c13f7mo ago

Update README.md

tensorxt
bdc61247mo ago

Update README.md

tensorxt
79519987mo ago

Update README.md

tensorxt
72127e77mo ago

Upload dataset

tensorxt
8b2c8067mo ago

Update README.md

tensorxt
16214317mo ago

Update README.md

tensorxt
8621ff77mo ago

Upload 4 files

tensorxt
ea526c27mo ago

Upload dataset

tensorxt
dc8375c8mo ago

initial commit

tensorxt