CoolFace
Datasetpublic

mbaye930/wolof-arabic-parallel-corpus

MudawanSn: A Gold-Standard Wolof--Arabic Parallel Corpus for Machine Translation A publicly available parallel corpus for the Wolof–Arabic language pair, a gold-standard resource containing 1,271 sentence-aligned pairs. The corpus consists of manual translations from Wolof into Modern Standard Arabic (MSA). The source texts are drawn from the MasakhaNER corpus, covering politics, society, religion, and sports in Senegalese news discourse. Dataset Structure The… See the full description on the dataset page: https://huggingface.co/datasets/mbaye930/wolof-arabic-parallel-corpus.

sourceHugging Facecc-by-nc-sa-4.0updated 4mo agoView on Hugging Face
3likes38downloads
8 commits on main
99dc1424mo ago

Update README.md

mbaye930
03c7ae04mo ago

Update README.md

mbaye930
bf314484mo ago

Update README.md

mbaye930
c7977d04mo ago

Upload 3 files

mbaye930
f40a28a4mo ago

Delete .gitattributes

mbaye930
79085054mo ago

Update README.md

mbaye930
36737054mo ago

Initial upload of train, dev, and test splits

mbaye930
62296e04mo ago

initial commit

mbaye930