tedim
Datasets
All datasets matching “tedim”ted-im2im-zhcn-ented-im2im-es-enzomi-tedim-bible-corpus
Zomi Tedim–Burmese Parallel Corpus (Bible Verses)
Dataset Summary
A parallel text corpus of Bible verses in Tedim (Zomi) and Burmese,
extracted from USX source files. This dataset is intended as a foundational
resource for Tedim-language NLP research, including machine translation,
language modeling, and text generation for this low-resource language.
Dataset Structure
Each record contains the following fields:
Field
Description
id
Unique… See the full description on the dataset page: https://huggingface.co/datasets/zomi-tedim-ai/zomi-tedim-bible-corpus.ted-im2im-de-en
