CoolFace
Datasetpublic

nguyenvulebinh/spoken_norm_pattern

Vietnamese Inverse Text Normalization Inverse text normalization (ITN) is the task that transforms spoken to written styles. It is particularly useful in automatic speech recognition (ASR) systems where proper names are often miss-recognized by their pronunciations instead of the written forms. By applying ITN, we can improve the readability of the ASR system’s output significantly. This dataset provides data for doing ITN task in the Vietnamese language. For example:… See the full description on the dataset page: https://huggingface.co/datasets/nguyenvulebinh/spoken_norm_pattern.

sourceHugging Faceapache-2.0updated 1y agoView on Hugging Face
4likes341downloads
Dataset Card

Vietnamese Inverse Text Normalization

Inverse text normalization (ITN) is the task that transforms spoken to written styles. It is particularly useful in automatic speech recognition (ASR) systems where proper names are often miss-recognized by their pronunciations instead of the written forms. By applying ITN, we can improve the readability of the ASR system’s output significantly. This dataset provides data for doing ITN task in the Vietnamese language.

For example:

Spoken (src)Written (tgt)Types
tám giờ chín phút ngày ba tháng tư năm hai nghìn8h9 3/4/2000time and date
tám mét khối năm mươi ki lô gam8m3 50 kgnumber and unit of measure
không chín sáu hai bảy bảy chín chín không bốn0962779904phone number

Dataset

The ITN dataset has 3 splits: train, validation, and test.

Dataset SplitNumber of Instances in Split
Train500,000
Validation2,500
Test2,500