CoolFace
Datasetpublic

SEACrowd/thai_romanization

The Thai Romanization dataset contains 648,241 Thai words that were transliterated into English, making Thai pronounciation easier for non-native Thai speakers. This is a valuable dataset for Thai language learners and researchers working on Thai language processing task. Each word in the Thai Romanization dataset is paired with its English phonetic representation, enabling accurate pronunciation guidance. This facilitates the learning and practice of Thai pronunciation for individuals who may not be familiar with the Thai script. The dataset aids in improving the accessibility and usability of Thai language resources, supporting applications such as speech recognition, text-to-speech synthesis, and machine translation. It enables the development of Thai language tools that can benefit Thai learners, tourists, and those interested in Thai culture and language.

sourceHugging Facecc-by-sa-3.0updated 2y agoView on Hugging Face
0likes25downloads
7 commits on main
b8e709a2y ago

Upload README.md with huggingface_hub

holylovenia
7f86aec2y ago

Upload __init__.py with huggingface_hub

holylovenia
ce0541d2y ago

Upload thai_romanization.py with huggingface_hub

holylovenia
28eb9972y ago

Upload README.md with huggingface_hub

holylovenia
cd738cf2y ago

Upload LICENSE with huggingface_hub

holylovenia
7bad0b52y ago

Upload requirements.txt with huggingface_hub

holylovenia
effd13d2y ago

initial commit

holylovenia