CoolFace
Datasetpublic

nickoo004/kaa-parallel-corpus

Kaa Karakalpak-English Parallel Corpus (FineTranslations) 📌 Overview This repository contains a high-quality, curated parallel corpus for the Karakalpak (kaa) language, paired with English (en). Karakalpak is a low-resource Turkic language spoken primarily in the Republic of Karakalpakstan. This dataset is a specialized subset extracted from the massive HuggingFaceFW/finetranslations project. The goal of this repo is to provide a dedicated and easy-to-access… See the full description on the dataset page: https://huggingface.co/datasets/nickoo004/kaa-parallel-corpus.

sourceHugging Faceapache-2.0updated 5mo agoView on Hugging Face
0likes41downloads

nickoo004/kaa-parallel-corpus · main · files are served by the source, never re-hosted here