ateso
Datasets
All datasets matching “ateso”Synthetic_Ateso_MMS
Dataset Card for "Synthetic_Atest_MMS"
More Information needed
Synthetic_Ateso_VITS_22.5k
Dataset Card for "Synthetic_Ateso_VITS_22.5k"
More Information needed
ateso-crowd-validated-paths
Dataset Card for "ateso-crowd-validated-paths"
More Information needed
ateso-english-bible-corpus
Ateso–English Bible Parallel Corpus
The largest publicly available sentence-level parallel corpus for the Ateso
(Teso) language.
Ateso is an Eastern Nilotic language spoken by approximately 3 million people —
the Iteso of eastern Uganda and western Kenya. Despite the size of that
community, Ateso remains thinly represented in the NLP ecosystem. Meta's
NLLB-200 covers 200 languages and includes no Eastern Nilotic language at
all — not Ateso, Turkana, Karamojong, Toposa, Nyangatom… See the full description on the dataset page: https://huggingface.co/datasets/Egonyu/ateso-english-bible-corpus.ateso-data-v1Ateso_news_articles
Ateso News Articles
Ateso (teo) is one of the most spoken languages in Uganda
Dataset Details
Artictles were scrapped from https://www.aicerit.co.ug
