tufaax/somali-multilingual-infopankki
Somali Multilingual Infopankki somali-multilingual-infopankki is a parallel corpus containing multilingual translation pairs that involve the Somali (so) language. This dataset has been filtered and extracted from the original Helsinki-NLP/opus_infopankki corpus. It is designed to support machine translation (NMT), multilingual sentence alignment, and Somali natural language processing (NLP) research. Dataset Details Source Dataset: Helsinki-NLP/opus_infopankki… See the full description on the dataset page: https://huggingface.co/datasets/tufaax/somali-multilingual-infopankki.
This repository belongs to tufaax on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
