CoolFace
Datasetpublic

Okwu/african-language-parallel-corpus

African Language Parallel Corpus Human-created, human-validated parallel sentence pairs for three African languages, released openly by Okwu. Version 1.0. Dataset summary A parallel corpus of everyday-register sentence pairs for Yorùbá, Swahili, and Nigerian Pidgin, each paired with English. The core is derived from NKENNE's own language-learning curriculum — content authored and reviewed by native-speaker educators — supplemented for Swahili with public-domain… See the full description on the dataset page: https://huggingface.co/datasets/Okwu/african-language-parallel-corpus.

sourceHugging Facecdla-permissive-2.0updated 3d agoView on Hugging Face
0likes140downloads

Okwu/african-language-parallel-corpus · main · files are served by the source, never re-hosted here