CoolFace
Datasetpublic

Okwu/african-language-parallel-corpus

African Language Parallel Corpus Human-created, human-validated parallel sentence pairs for three African languages, released openly by Okwu. Version 1.0. Dataset summary A parallel corpus of everyday-register sentence pairs for Yorùbá, Swahili, and Nigerian Pidgin, each paired with English. The core is derived from NKENNE's own language-learning curriculum — content authored and reviewed by native-speaker educators — supplemented for Swahili with public-domain… See the full description on the dataset page: https://huggingface.co/datasets/Okwu/african-language-parallel-corpus.

sourceHugging Facecdla-permissive-2.0updated 2d agoView on Hugging Face
0likes139downloads
settings

This repository belongs to Okwu on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameafrican-language-parallel-corpus
visibilitypublic
licencecdla-permissive-2.0
gatedno
ownerOkwu
Account settings
Okwu/african-language-parallel-corpus · CoolFace