CoolFace
Datasetpublic

adeshkin/khakas-russian-parallel-corpus

Khakas-Russian Parallel Corpus The creation of this dataset is aimed at supporting the development of natural language processing (NLP) tools and machine translation for the Khakas language, which is classified as a "Definitely Endangered" language. By providing high-quality parallel data, this project helps preserve the linguistic heritage of the Khakas people. Dataset Overlap: The Khakas sentences in this corpus do not overlap with those in the Khakas… See the full description on the dataset page: https://huggingface.co/datasets/adeshkin/khakas-russian-parallel-corpus.

sourceHugging Facecc-by-4.0updated 17d agoView on Hugging Face
2likes255downloads

adeshkin/khakas-russian-parallel-corpus · main · files are served by the source, never re-hosted here